World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Zhennan, Liu, Kai, Qin, Yuxin, Tian, Shuai, Zheng, Yupeng, Zhou, Mingcai, Yu, Chao, Li, Haoran, Zhao, Dongbin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
by: Jiang, Zhennan, et al.
Published: (2026)
by: Jiang, Zhennan, et al.
Published: (2026)
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning
by: Liu, Yuan, et al.
Published: (2026)
by: Liu, Yuan, et al.
Published: (2026)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
TeViR: Text-to-Video Reward with Diffusion Models for Efficient Reinforcement Learning
by: Chen, Yuhui, et al.
Published: (2025)
by: Chen, Yuhui, et al.
Published: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Posterior Optimization with Clipped Objective for Bridging Efficiency and Stability in Generative Policy Learning
by: Chen, Yuhui, et al.
Published: (2026)
by: Chen, Yuhui, et al.
Published: (2026)
Beyond Action Residuals: Real-World Robot Policy Steering via Bottleneck Latent Reinforcement Learning
by: Yu, Dongjie, et al.
Published: (2026)
by: Yu, Dongjie, et al.
Published: (2026)
OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation
by: Zheng, Yuhang, et al.
Published: (2026)
by: Zheng, Yuhang, et al.
Published: (2026)
ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
by: Chen, Yuhui, et al.
Published: (2025)
by: Chen, Yuhui, et al.
Published: (2025)
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation
by: Xu, Qinwen, et al.
Published: (2026)
by: Xu, Qinwen, et al.
Published: (2026)
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
by: Yuan, Ge, et al.
Published: (2026)
by: Yuan, Ge, et al.
Published: (2026)
Empowering Multi-Robot Cooperation via Sequential World Models
by: Zhao, Zijie, et al.
Published: (2025)
by: Zhao, Zijie, et al.
Published: (2025)
Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
by: Bo, Zitong, et al.
Published: (2025)
by: Bo, Zitong, et al.
Published: (2025)
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
DiffuDepGrasp: Diffusion-based Depth Noise Modeling Empowers Sim2Real Robotic Grasping
by: Zhou, Yingting, et al.
Published: (2025)
by: Zhou, Yingting, et al.
Published: (2025)
RM-RL: Role-Model Reinforcement Learning for Precise Robot Manipulation
by: Chen, Xiangyu, et al.
Published: (2025)
by: Chen, Xiangyu, et al.
Published: (2025)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026)
by: Liu, Xin, et al.
Published: (2026)
STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation
by: Tian, Yuxuan, et al.
Published: (2026)
by: Tian, Yuxuan, et al.
Published: (2026)
World Models for Robotic Manipulation: A Survey
by: Wang, Fangyuan, et al.
Published: (2026)
by: Wang, Fangyuan, et al.
Published: (2026)
NeuronsGym: A Hybrid Framework and Benchmark for Robot Tasks with Sim2Real Policy Learning
by: Li, Haoran, et al.
Published: (2023)
by: Li, Haoran, et al.
Published: (2023)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation
by: Liu, Jiahao, et al.
Published: (2026)
by: Liu, Jiahao, et al.
Published: (2026)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
by: Qian, Zezhong, et al.
Published: (2025)
by: Qian, Zezhong, et al.
Published: (2025)
Iterative On-Policy Refinement of Hierarchical Diffusion Policies for Language-Conditioned Manipulation
by: Grislain, Clemence, et al.
Published: (2026)
by: Grislain, Clemence, et al.
Published: (2026)
DreamerAD: Efficient Reinforcement Learning via Latent World Model for Autonomous Driving
by: Yang, Pengxuan, et al.
Published: (2026)
by: Yang, Pengxuan, et al.
Published: (2026)
Adaptive Diffusion Policy Optimization for Robotic Manipulation
by: Jiang, Huiyun, et al.
Published: (2025)
by: Jiang, Huiyun, et al.
Published: (2025)
Evaluating Real-World Robot Manipulation Policies in Simulation
by: Li, Xuanlin, et al.
Published: (2024)
by: Li, Xuanlin, et al.
Published: (2024)
LongBench: Evaluating Robotic Manipulation Policies on Real-World Long-Horizon Tasks
by: Chen, Xueyao, et al.
Published: (2026)
by: Chen, Xueyao, et al.
Published: (2026)
Scaling World Model for Hierarchical Manipulation Policies
by: Long, Qian, et al.
Published: (2026)
by: Long, Qian, et al.
Published: (2026)
GaussianDream: A Feed-Forward 3D Gaussian World Model for Robotic Manipulation
by: Zhang, Zijian, et al.
Published: (2026)
by: Zhang, Zijian, et al.
Published: (2026)
Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation
by: Liao, Yue, et al.
Published: (2025)
by: Liao, Yue, et al.
Published: (2025)
MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment
by: Zhang, Ruicheng, et al.
Published: (2025)
by: Zhang, Ruicheng, et al.
Published: (2025)
RISE: Self-Improving Robot Policy with Compositional World Model
by: Yang, Jiazhi, et al.
Published: (2026)
by: Yang, Jiazhi, et al.
Published: (2026)
SAMP: Spatial Anchor-based Motion Policy for Collision-Aware Robotic Manipulators
by: Chen, Kai, et al.
Published: (2025)
by: Chen, Kai, et al.
Published: (2025)
Genie Centurion: Accelerating Scalable Real-World Robot Training with Human Rewind-and-Refine Guidance
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
WorldRFT: Latent World Model Planning with Reinforcement Fine-Tuning for Autonomous Driving
by: Yang, Pengxuan, et al.
Published: (2025)
by: Yang, Pengxuan, et al.
Published: (2025)
EMMA: Generalizing Real-World Robot Manipulation via Generative Visual Transfer
by: Dong, Zhehao, et al.
Published: (2025)
by: Dong, Zhehao, et al.
Published: (2025)
Similar Items
-
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
by: Jiang, Zhennan, et al.
Published: (2026) -
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning
by: Liu, Yuan, et al.
Published: (2026) -
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025) -
TeViR: Text-to-Video Reward with Diffusion Models for Efficient Reinforcement Learning
by: Chen, Yuhui, et al.
Published: (2025) -
Survey of Vision-Language-Action Models for Embodied Manipulation
by: Li, Haoran, et al.
Published: (2025)