EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Shengjie, Liu, Shaohuai, Ye, Weirui, You, Jiacheng, Gao, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Tasks, Not Samples: Mastering Humanoid Control through Multi-Task Model-Based Reinforcement Learning
by: Liu, Shaohuai, et al.
Published: (2026)
by: Liu, Shaohuai, et al.
Published: (2026)
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
by: Ye, Weirui, et al.
Published: (2023)
by: Ye, Weirui, et al.
Published: (2023)
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
by: Ye, Weirui, et al.
Published: (2025)
by: Ye, Weirui, et al.
Published: (2025)
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
by: Evers, Thomas, et al.
Published: (2026)
by: Evers, Thomas, et al.
Published: (2026)
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025)
by: Wang, Shengjie, et al.
Published: (2025)
PRIME: Scaffolding Manipulation Tasks with Behavior Primitives for Data-Efficient Imitation Learning
by: Gao, Tian, et al.
Published: (2024)
by: Gao, Tian, et al.
Published: (2024)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
by: Nie, Buqing, et al.
Published: (2025)
by: Nie, Buqing, et al.
Published: (2025)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Leveraging Locality to Boost Sample Efficiency in Robotic Manipulation
by: Zhang, Tong, et al.
Published: (2024)
by: Zhang, Tong, et al.
Published: (2024)
Learning Pareto Set for Multi-Objective Continuous Robot Control
by: Shu, Tianye, et al.
Published: (2024)
by: Shu, Tianye, et al.
Published: (2024)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
by: Gao, Jensen, et al.
Published: (2024)
by: Gao, Jensen, et al.
Published: (2024)
V2X-VLM: End-to-End V2X Cooperative Autonomous Driving Through Large Vision-Language Models
by: You, Junwei, et al.
Published: (2024)
by: You, Junwei, et al.
Published: (2024)
OmniDrones: An Efficient and Flexible Platform for Reinforcement Learning in Drone Control
by: Xu, Botian, et al.
Published: (2023)
by: Xu, Botian, et al.
Published: (2023)
Scaling Algorithm Distillation for Continuous Control with Mamba
by: Beaussant, Samuel, et al.
Published: (2025)
by: Beaussant, Samuel, et al.
Published: (2025)
Data-Efficient Learning from Human Interventions for Mobile Robots
by: Peng, Zhenghao, et al.
Published: (2025)
by: Peng, Zhenghao, et al.
Published: (2025)
Growing Q-Networks: Solving Continuous Control Tasks with Adaptive Control Resolution
by: Seyde, Tim, et al.
Published: (2024)
by: Seyde, Tim, et al.
Published: (2024)
Meta-Controller: Few-Shot Imitation of Unseen Embodiments and Tasks in Continuous Control
by: Cho, Seongwoong, et al.
Published: (2024)
by: Cho, Seongwoong, et al.
Published: (2024)
RDT2: Exploring the Scaling Limit of UMI Data Towards Zero-Shot Cross-Embodiment Generalization
by: Liu, Songming, et al.
Published: (2026)
by: Liu, Songming, et al.
Published: (2026)
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
by: Das, Rocktim Jyoti, et al.
Published: (2025)
by: Das, Rocktim Jyoti, et al.
Published: (2025)
Sim2Dust: Mastering Dynamic Waypoint Tracking on Granular Media
by: Orsula, Andrej, et al.
Published: (2025)
by: Orsula, Andrej, et al.
Published: (2025)
Confounding Robust Continuous Control via Automatic Reward Shaping
by: Juliani, Mateo, et al.
Published: (2026)
by: Juliani, Mateo, et al.
Published: (2026)
EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data
by: Zhang, Jesse, et al.
Published: (2024)
by: Zhang, Jesse, et al.
Published: (2024)
Continual Learning for Multimodal Data Fusion of a Soft Gripper
by: Kushawaha, Nilay, et al.
Published: (2024)
by: Kushawaha, Nilay, et al.
Published: (2024)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
by: Zhou, Zehao
Published: (2024)
by: Zhou, Zehao
Published: (2024)
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
by: Korkmaz, Yigit, et al.
Published: (2025)
by: Korkmaz, Yigit, et al.
Published: (2025)
ExBody2: Advanced Expressive Humanoid Whole-Body Control
by: Ji, Mazeyu, et al.
Published: (2024)
by: Ji, Mazeyu, et al.
Published: (2024)
FACET: Force-Adaptive Control via Impedance Reference Tracking for Legged Robots
by: Xu, Botian, et al.
Published: (2025)
by: Xu, Botian, et al.
Published: (2025)
RLtools: A Fast, Portable Deep Reinforcement Learning Library for Continuous Control
by: Eschmann, Jonas, et al.
Published: (2023)
by: Eschmann, Jonas, et al.
Published: (2023)
Offline Reinforcement Learning with Discrete Diffusion Skills
by: Qiao, RuiXi, et al.
Published: (2025)
by: Qiao, RuiXi, et al.
Published: (2025)
Discrete Variational Autoencoding via Policy Search
by: Drolet, Michael, et al.
Published: (2025)
by: Drolet, Michael, et al.
Published: (2025)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
by: Li, Lanpei, et al.
Published: (2024)
by: Li, Lanpei, et al.
Published: (2024)
Spiking Neural Networks for Continuous Control via End-to-End Model-Based Learning
by: Huebotter, Justus, et al.
Published: (2025)
by: Huebotter, Justus, et al.
Published: (2025)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
by: Chen, Shirui, et al.
Published: (2026)
by: Chen, Shirui, et al.
Published: (2026)
An Integrated Imitation and Reinforcement Learning Methodology for Robust Agile Aircraft Control with Limited Pilot Demonstration Data
by: Sever, Gulay Goktas, et al.
Published: (2023)
by: Sever, Gulay Goktas, et al.
Published: (2023)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
by: Liu, Huihan, et al.
Published: (2026)
by: Liu, Huihan, et al.
Published: (2026)
DHP: Discrete Hierarchical Planning for Hierarchical Reinforcement Learning Agents
by: Sharma, Shashank, et al.
Published: (2025)
by: Sharma, Shashank, et al.
Published: (2025)
EasyInsert: A Data-Efficient and Generalizable Insertion Policy
by: Li, Guanghe, et al.
Published: (2025)
by: Li, Guanghe, et al.
Published: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
by: Patel, Bhrij, et al.
Published: (2023)
by: Patel, Bhrij, et al.
Published: (2023)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
by: Yue, Yang, et al.
Published: (2024)
by: Yue, Yang, et al.
Published: (2024)
Similar Items
-
Scaling Tasks, Not Samples: Mastering Humanoid Control through Multi-Task Model-Based Reinforcement Learning
by: Liu, Shaohuai, et al.
Published: (2026) -
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
by: Ye, Weirui, et al.
Published: (2023) -
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
by: Ye, Weirui, et al.
Published: (2025) -
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
by: Evers, Thomas, et al.
Published: (2026) -
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025)