Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Chuning, Yu, Raymond, Feng, Siyuan, Burchfiel, Benjamin, Shah, Paarth, Gupta, Abhishek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic World Models
von: Berg, Jacob, et al.
Veröffentlicht: (2025)
von: Berg, Jacob, et al.
Veröffentlicht: (2025)
PolyTouch: A Robust Multi-Modal Tactile Sensor for Contact-rich Manipulation Using Tactile-Diffusion Policies
von: Zhao, Jialiang, et al.
Veröffentlicht: (2025)
von: Zhao, Jialiang, et al.
Veröffentlicht: (2025)
Robot Learning as an Empirical Science: Best Practices for Policy Evaluation
von: Kress-Gazit, Hadas, et al.
Veröffentlicht: (2024)
von: Kress-Gazit, Hadas, et al.
Veröffentlicht: (2024)
Geometry-aware 4D Video Generation for Robot Manipulation
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
von: Chi, Cheng, et al.
Veröffentlicht: (2023)
von: Chi, Cheng, et al.
Veröffentlicht: (2023)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
PoseDiff: A Unified Diffusion Model Bridging Robot Pose Estimation and Video-to-Action Control
von: Zhang, Haozhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Haozhuo, et al.
Veröffentlicht: (2025)
Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation
von: Levy, Jacob, et al.
Veröffentlicht: (2026)
von: Levy, Jacob, et al.
Veröffentlicht: (2026)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
ASID: Active Exploration for System Identification in Robotic Manipulation
von: Memmel, Marius, et al.
Veröffentlicht: (2024)
von: Memmel, Marius, et al.
Veröffentlicht: (2024)
Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
Video Generators are Robot Policies
von: Liang, Junbang, et al.
Veröffentlicht: (2025)
von: Liang, Junbang, et al.
Veröffentlicht: (2025)
$τ_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
von: Zhou, Pengfei, et al.
Veröffentlicht: (2026)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2026)
An Ontology for Unified Modeling of Tasks, Actions, Environments, and Capabilities in Personal Service Robotics
von: Martorana, Margherita, et al.
Veröffentlicht: (2025)
von: Martorana, Margherita, et al.
Veröffentlicht: (2025)
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
von: Guo, Jun, et al.
Veröffentlicht: (2026)
von: Guo, Jun, et al.
Veröffentlicht: (2026)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
Adaptive Compliance Policy: Learning Approximate Compliance for Diffusion Guided Control
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
BUMBLE: Unifying Reasoning and Acting with Vision-Language Models for Building-wide Mobile Manipulation
von: Shah, Rutav, et al.
Veröffentlicht: (2024)
von: Shah, Rutav, et al.
Veröffentlicht: (2024)
ChronoDreamer: Action-Conditioned World Model as an Online Simulator for Robotic Planning
von: Zhou, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenhao, et al.
Veröffentlicht: (2025)
WorldVLA: Towards Autoregressive Action World Model
von: Cen, Jun, et al.
Veröffentlicht: (2025)
von: Cen, Jun, et al.
Veröffentlicht: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
von: Li, Ying, et al.
Veröffentlicht: (2025)
von: Li, Ying, et al.
Veröffentlicht: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
von: Zhai, Shaopeng, et al.
Veröffentlicht: (2025)
von: Zhai, Shaopeng, et al.
Veröffentlicht: (2025)
DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
Commanding Humanoid by Free-form Language: A Large Language Action Model with Unified Motion Vocabulary
von: Liu, Zhirui, et al.
Veröffentlicht: (2025)
von: Liu, Zhirui, et al.
Veröffentlicht: (2025)
CoPAL: Corrective Planning of Robot Actions with Large Language Models
von: Joublin, Frank, et al.
Veröffentlicht: (2023)
von: Joublin, Frank, et al.
Veröffentlicht: (2023)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
ViSA-Flow: Accelerating Robot Skill Learning via Large-Scale Video Semantic Action Flow
von: Chen, Changhe, et al.
Veröffentlicht: (2025)
von: Chen, Changhe, et al.
Veröffentlicht: (2025)
SLAC: Simulation-Pretrained Latent Action Space for Whole-Body Real-World RL
von: Hu, Jiaheng, et al.
Veröffentlicht: (2025)
von: Hu, Jiaheng, et al.
Veröffentlicht: (2025)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
DreamGen: Unlocking Generalization in Robot Learning through Video World Models
von: Jang, Joel, et al.
Veröffentlicht: (2025)
von: Jang, Joel, et al.
Veröffentlicht: (2025)
ManiWAV: Learning Robot Manipulation from In-the-Wild Audio-Visual Data
von: Liu, Zeyi, et al.
Veröffentlicht: (2024)
von: Liu, Zeyi, et al.
Veröffentlicht: (2024)
Multi-Task Interactive Robot Fleet Learning with Visual World Models
von: Liu, Huihan, et al.
Veröffentlicht: (2024)
von: Liu, Huihan, et al.
Veröffentlicht: (2024)
Efficient Continual Adaptation of Pretrained Robotic Policy with Online Meta-Learned Adapters
von: Zhu, Ruiqi, et al.
Veröffentlicht: (2025)
von: Zhu, Ruiqi, et al.
Veröffentlicht: (2025)
Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
WMPO: World Model-based Policy Optimization for Vision-Language-Action Models
von: Zhu, Fangqi, et al.
Veröffentlicht: (2025)
von: Zhu, Fangqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Semantic World Models
von: Berg, Jacob, et al.
Veröffentlicht: (2025) -
PolyTouch: A Robust Multi-Modal Tactile Sensor for Contact-rich Manipulation Using Tactile-Diffusion Policies
von: Zhao, Jialiang, et al.
Veröffentlicht: (2025) -
Robot Learning as an Empirical Science: Best Practices for Policy Evaluation
von: Kress-Gazit, Hadas, et al.
Veröffentlicht: (2024) -
Geometry-aware 4D Video Generation for Robot Manipulation
von: Liu, Zeyi, et al.
Veröffentlicht: (2025) -
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
von: Huang, Kevin, et al.
Veröffentlicht: (2025)