Saved in:
| Main Authors: | Tang, Yinzhou, Shang, Yu, Chen, Yinuo, Wei, Bingwen, Zhang, Xin, Yu, Shu'ang, Shi, Liangzhi, Yu, Chao, Gao, Chen, Wu, Wei, Li, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.03556 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
JuggleRL: Mastering Ball Juggling with a Quadrotor via Deep Reinforcement Learning
by: Ji, Shilong, et al.
Published: (2025)
by: Ji, Shilong, et al.
Published: (2025)
LongScape: Advancing Long-Horizon Embodied World Models with Context-Aware MoE
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
AirScape: An Aerial Generative World Model with Motion Controllability
by: Zhao, Baining, et al.
Published: (2025)
by: Zhao, Baining, et al.
Published: (2025)
What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study
by: Chen, Jiayu, et al.
Published: (2024)
by: Chen, Jiayu, et al.
Published: (2024)
What Can RL Bring to VLA Generalization? An Empirical Study
by: Liu, Jijia, et al.
Published: (2025)
by: Liu, Jijia, et al.
Published: (2025)
Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models
by: Shi, Liangzhi, et al.
Published: (2026)
by: Shi, Liangzhi, et al.
Published: (2026)
RoboHanger: Learning Generalizable Robotic Hanger Insertion for Diverse Garments
by: Chen, Yuxing, et al.
Published: (2024)
by: Chen, Yuxing, et al.
Published: (2024)
Multi-Robot System for Cooperative Exploration in Unknown Environments: A Survey
by: Wang, Chuqi, et al.
Published: (2025)
by: Wang, Chuqi, et al.
Published: (2025)
Keyframe-Guided Structured Rewards for Reinforcement Learning in Long-Horizon Laboratory Robotics
by: Qiu, Yibo, et al.
Published: (2026)
by: Qiu, Yibo, et al.
Published: (2026)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
by: Wu, Hao, et al.
Published: (2026)
by: Wu, Hao, et al.
Published: (2026)
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
by: Geng, Haoran, et al.
Published: (2025)
by: Geng, Haoran, et al.
Published: (2025)
Unified Algorithms for RL with Decision-Estimation Coefficients: PAC, Reward-Free, Preference-Based Learning, and Beyond
by: Chen, Fan, et al.
Published: (2022)
by: Chen, Fan, et al.
Published: (2022)
RLinf-USER: A Unified and Extensible System for Real-World Online Policy Learning in Embodied AI
by: Zang, Hongzhi, et al.
Published: (2026)
by: Zang, Hongzhi, et al.
Published: (2026)
Localization matters too: How localization error affects UAV flight
by: Zhang, Suquan, et al.
Published: (2024)
by: Zhang, Suquan, et al.
Published: (2024)
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
by: Huang, Yu, et al.
Published: (2026)
by: Huang, Yu, et al.
Published: (2026)
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
by: Jiang, Zhennan, et al.
Published: (2026)
by: Jiang, Zhennan, et al.
Published: (2026)
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
by: Shang, Yu, et al.
Published: (2026)
by: Shang, Yu, et al.
Published: (2026)
RoboCodeX: Multimodal Code Generation for Robotic Behavior Synthesis
by: Mu, Yao, et al.
Published: (2024)
by: Mu, Yao, et al.
Published: (2024)
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation
by: Feng, Xuelu, et al.
Published: (2025)
by: Feng, Xuelu, et al.
Published: (2025)
Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning
by: Liao, Xinyao, et al.
Published: (2025)
by: Liao, Xinyao, et al.
Published: (2025)
Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training
by: Ye, Chenlu, et al.
Published: (2025)
by: Ye, Chenlu, et al.
Published: (2025)
MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models
by: Guo, Zile, et al.
Published: (2026)
by: Guo, Zile, et al.
Published: (2026)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
RewardAnything: Generalizable Principle-Following Reward Models
by: Yu, Zhuohao, et al.
Published: (2025)
by: Yu, Zhuohao, et al.
Published: (2025)
SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling
by: Zhang, Yixian, et al.
Published: (2025)
by: Zhang, Yixian, et al.
Published: (2025)
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation
by: Zhang, Zongzheng, et al.
Published: (2025)
by: Zhang, Zongzheng, et al.
Published: (2025)
RoboDreamer: Learning Compositional World Models for Robot Imagination
by: Zhou, Siyuan, et al.
Published: (2024)
by: Zhou, Siyuan, et al.
Published: (2024)
RoboGrasp: A Universal Grasping Policy for Robust Robotic Control
by: Huang, Yiqi, et al.
Published: (2025)
by: Huang, Yiqi, et al.
Published: (2025)
RoboPilot: Generalizable Dynamic Robotic Manipulation with Dual-thinking Modes
by: Liu, Xinyi, et al.
Published: (2025)
by: Liu, Xinyi, et al.
Published: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
Generalizable Dense Reward for Long-Horizon Robotic Tasks
by: Yong, Silong, et al.
Published: (2026)
by: Yong, Silong, et al.
Published: (2026)
MARBLE: Multi-Aspect Reward Balance for Diffusion RL
by: Zhao, Canyu, et al.
Published: (2026)
by: Zhao, Canyu, et al.
Published: (2026)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
by: Jiang, Zhennan, et al.
Published: (2025)
by: Jiang, Zhennan, et al.
Published: (2025)
RoboWits: Unexpected Challenges for Robotic Creative Problem Solving
by: Lin, Chunru, et al.
Published: (2026)
by: Lin, Chunru, et al.
Published: (2026)
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
by: Bhattarai, Manish, et al.
Published: (2026)
by: Bhattarai, Manish, et al.
Published: (2026)
RoboRouter: Training-Free Policy Routing for Robotic Manipulation
by: Chen, Yiteng, et al.
Published: (2026)
by: Chen, Yiteng, et al.
Published: (2026)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
by: Huang, Wenlong, et al.
Published: (2026)
by: Huang, Wenlong, et al.
Published: (2026)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Similar Items
-
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025) -
JuggleRL: Mastering Ball Juggling with a Quadrotor via Deep Reinforcement Learning
by: Ji, Shilong, et al.
Published: (2025) -
LongScape: Advancing Long-Horizon Embodied World Models with Context-Aware MoE
by: Shang, Yu, et al.
Published: (2025) -
AirScape: An Aerial Generative World Model with Motion Controllability
by: Zhao, Baining, et al.
Published: (2025) -
What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study
by: Chen, Jiayu, et al.
Published: (2024)