RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Charles, Li, Qiyang, Luo, Jianlan, Levine, Sergey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2024)
por: Luo, Jianlan, et al.
Publicado: (2024)
Q-learning with Adjoint Matching
por: Li, Qiyang, et al.
Publicado: (2026)
por: Li, Qiyang, et al.
Publicado: (2026)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
por: Feng, Yunhai, et al.
Publicado: (2025)
por: Feng, Yunhai, et al.
Publicado: (2025)
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
por: Li, Lanpei, et al.
Publicado: (2024)
por: Li, Lanpei, et al.
Publicado: (2024)
A Taxonomy for Evaluating Generalist Robot Manipulation Policies
por: Gao, Jensen, et al.
Publicado: (2025)
por: Gao, Jensen, et al.
Publicado: (2025)
RLIF: Interactive Imitation Learning as Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2023)
por: Luo, Jianlan, et al.
Publicado: (2023)
Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression
por: Acero, Fernando, et al.
Publicado: (2024)
por: Acero, Fernando, et al.
Publicado: (2024)
Octo: An Open-Source Generalist Robot Policy
por: Octo Model Team, et al.
Publicado: (2024)
por: Octo Model Team, et al.
Publicado: (2024)
Foundation Policies with Hilbert Representations
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Real-Time Execution of Action Chunking Flow Policies
por: Black, Kevin, et al.
Publicado: (2025)
por: Black, Kevin, et al.
Publicado: (2025)
SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2024)
por: Luo, Jianlan, et al.
Publicado: (2024)
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
por: Jones, Joshua, et al.
Publicado: (2025)
por: Jones, Joshua, et al.
Publicado: (2025)
Turning Video Models into Generalist Robot Policies
por: Li, Sizhe Lester, et al.
Publicado: (2026)
por: Li, Sizhe Lester, et al.
Publicado: (2026)
Reinforcement Learning via Auxiliary Task Distillation
por: Harish, Abhinav Narayan, et al.
Publicado: (2024)
por: Harish, Abhinav Narayan, et al.
Publicado: (2024)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
por: Zhou, Zhiyuan, et al.
Publicado: (2025)
por: Zhou, Zhiyuan, et al.
Publicado: (2025)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
por: Stachowicz, Kyle, et al.
Publicado: (2024)
por: Stachowicz, Kyle, et al.
Publicado: (2024)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
Flow Q-Learning
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
Embodiment-Aware Generalist Specialist Distillation for Unified Humanoid Whole-Body Control
por: Peng, Quanquan, et al.
Publicado: (2026)
por: Peng, Quanquan, et al.
Publicado: (2026)
Robot Policy Transfer with Online Demonstrations: An Active Reinforcement Learning Approach
por: Hou, Muhan, et al.
Publicado: (2025)
por: Hou, Muhan, et al.
Publicado: (2025)
Co-jump: Cooperative Jumping with Quadrupedal Robots via Multi-Agent Reinforcement Learning
por: Dong, Shihao, et al.
Publicado: (2026)
por: Dong, Shihao, et al.
Publicado: (2026)
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
Learning Visuotactile Skills with Two Multifingered Hands
por: Lin, Toru, et al.
Publicado: (2024)
por: Lin, Toru, et al.
Publicado: (2024)
RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
por: Nasiriany, Soroush, et al.
Publicado: (2024)
por: Nasiriany, Soroush, et al.
Publicado: (2024)
Uncertainty-Based Smooth Policy Regularisation for Reinforcement Learning with Few Demonstrations
por: Zhu, Yujie, et al.
Publicado: (2025)
por: Zhu, Yujie, et al.
Publicado: (2025)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
por: Nakamoto, Mitsuhiko, et al.
Publicado: (2024)
por: Nakamoto, Mitsuhiko, et al.
Publicado: (2024)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
por: Liu, Tenglong, et al.
Publicado: (2024)
por: Liu, Tenglong, et al.
Publicado: (2024)
Robot Policy Learning with Temporal Optimal Transport Reward
por: Fu, Yuwei, et al.
Publicado: (2024)
por: Fu, Yuwei, et al.
Publicado: (2024)
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
por: Li, Keyu, et al.
Publicado: (2021)
por: Li, Keyu, et al.
Publicado: (2021)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic Routing
por: Wang, Yixiao, et al.
Publicado: (2025)
por: Wang, Yixiao, et al.
Publicado: (2025)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
por: Tang, Grace, et al.
Publicado: (2024)
por: Tang, Grace, et al.
Publicado: (2024)
Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections
por: Zha, Lihan, et al.
Publicado: (2023)
por: Zha, Lihan, et al.
Publicado: (2023)
Research on Autonomous Robots Navigation based on Reinforcement Learning
por: Wang, Zixiang, et al.
Publicado: (2024)
por: Wang, Zixiang, et al.
Publicado: (2024)
RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots
por: Nasiriany, Soroush, et al.
Publicado: (2026)
por: Nasiriany, Soroush, et al.
Publicado: (2026)
Rethinking the Intermediate Features in Adversarial Attacks: Misleading Robotic Models via Adversarial Distillation
por: Zhao, Ke, et al.
Publicado: (2024)
por: Zhao, Ke, et al.
Publicado: (2024)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
por: Lei, Kun, et al.
Publicado: (2025)
por: Lei, Kun, et al.
Publicado: (2025)
Ejemplares similares
-
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025) -
Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2024) -
Q-learning with Adjoint Matching
por: Li, Qiyang, et al.
Publicado: (2026) -
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
por: Feng, Yunhai, et al.
Publicado: (2025) -
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025)