Large Video Planner Enables Generalizable Robot Control
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Boyuan, Zhang, Tianyuan, Geng, Haoran, Zhang, Caiyi, Li, Peihao, Song, Kiwhan, Freeman, William T., Malik, Jitendra, Abbeel, Pieter, Tedrake, Russ, Sitzmann, Vincent, Du, Yilun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
History-Guided Video Diffusion
di: Song, Kiwhan, et al.
Pubblicazione: (2025)
di: Song, Kiwhan, et al.
Pubblicazione: (2025)
Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation
di: Heng, Liang, et al.
Pubblicazione: (2025)
di: Heng, Liang, et al.
Pubblicazione: (2025)
Rodrigues Network for Learning Robot Actions
di: Zhang, Jialiang, et al.
Pubblicazione: (2025)
di: Zhang, Jialiang, et al.
Pubblicazione: (2025)
Selective Underfitting in Diffusion Models
di: Song, Kiwhan, et al.
Pubblicazione: (2025)
di: Song, Kiwhan, et al.
Pubblicazione: (2025)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
di: Wang, Yuran, et al.
Pubblicazione: (2025)
di: Wang, Yuran, et al.
Pubblicazione: (2025)
Interactive Task Planning with Language Models
di: Li, Boyi, et al.
Pubblicazione: (2023)
di: Li, Boyi, et al.
Pubblicazione: (2023)
DittoGym: Learning to Control Soft Shape-Shifting Robots
di: Huang, Suning, et al.
Pubblicazione: (2024)
di: Huang, Suning, et al.
Pubblicazione: (2024)
PoCo: Policy Composition from and for Heterogeneous Robot Learning
di: Wang, Lirui, et al.
Pubblicazione: (2024)
di: Wang, Lirui, et al.
Pubblicazione: (2024)
Deep Sensorimotor Control by Imitating Predictive Models of Human Motion
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2025)
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2025)
DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization
di: Tang, Yikai, et al.
Pubblicazione: (2025)
di: Tang, Yikai, et al.
Pubblicazione: (2025)
Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning
di: Cheng, Ziheng, et al.
Pubblicazione: (2026)
di: Cheng, Ziheng, et al.
Pubblicazione: (2026)
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
di: Wang, Renhao, et al.
Pubblicazione: (2025)
di: Wang, Renhao, et al.
Pubblicazione: (2025)
SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending
di: Kuang, Yuxuan, et al.
Pubblicazione: (2025)
di: Kuang, Yuxuan, et al.
Pubblicazione: (2025)
Hand-Object Interaction Pretraining from Videos
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
How to Peel with a Knife: Aligning Fine-Grained Manipulation with Human Preference
di: Lin, Toru, et al.
Pubblicazione: (2026)
di: Lin, Toru, et al.
Pubblicazione: (2026)
Twisting Lids Off with Two Hands
di: Lin, Toru, et al.
Pubblicazione: (2024)
di: Lin, Toru, et al.
Pubblicazione: (2024)
World Model for Robot Learning: A Comprehensive Survey
di: Hou, Bohan, et al.
Pubblicazione: (2026)
di: Hou, Bohan, et al.
Pubblicazione: (2026)
Object-centric 3D Motion Field for Robot Learning from Human Videos
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2025)
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2025)
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
di: Geng, Haoran, et al.
Pubblicazione: (2025)
di: Geng, Haoran, et al.
Pubblicazione: (2025)
Visual Imitation Enables Contextual Humanoid Control
di: Allshire, Arthur, et al.
Pubblicazione: (2025)
di: Allshire, Arthur, et al.
Pubblicazione: (2025)
End-to-end RL Improves Dexterous Grasping Policies
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
Robot Fleet Learning via Policy Merging
di: Wang, Lirui, et al.
Pubblicazione: (2023)
di: Wang, Lirui, et al.
Pubblicazione: (2023)
pixelSplat: 3D Gaussian Splats from Image Pairs for Scalable Generalizable 3D Reconstruction
di: Charatan, David, et al.
Pubblicazione: (2023)
di: Charatan, David, et al.
Pubblicazione: (2023)
D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping
di: Lou, Haozhe, et al.
Pubblicazione: (2026)
di: Lou, Haozhe, et al.
Pubblicazione: (2026)
Sampling-Based Motion Planning with Discrete Configuration-Space Symmetries
di: Cohn, Thomas, et al.
Pubblicazione: (2025)
di: Cohn, Thomas, et al.
Pubblicazione: (2025)
FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
di: Luo, Jianlan, et al.
Pubblicazione: (2024)
di: Luo, Jianlan, et al.
Pubblicazione: (2024)
Generative View Stitching
di: Song, Chonghyuk, et al.
Pubblicazione: (2025)
di: Song, Chonghyuk, et al.
Pubblicazione: (2025)
Video as the New Language for Real-World Decision Making
di: Yang, Sherry, et al.
Pubblicazione: (2024)
di: Yang, Sherry, et al.
Pubblicazione: (2024)
Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation
di: Chen, Yongchao, et al.
Pubblicazione: (2025)
di: Chen, Yongchao, et al.
Pubblicazione: (2025)
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
di: Zhang, Junyu, et al.
Pubblicazione: (2025)
di: Zhang, Junyu, et al.
Pubblicazione: (2025)
Turning Video Models into Generalist Robot Policies
di: Li, Sizhe Lester, et al.
Pubblicazione: (2026)
di: Li, Sizhe Lester, et al.
Pubblicazione: (2026)
Generalizable Reasoning through Compositional Energy Minimization
di: Oarga, Alexandru, et al.
Pubblicazione: (2025)
di: Oarga, Alexandru, et al.
Pubblicazione: (2025)
Empirical Analysis of Sim-and-Real Cotraining of Diffusion Policies for Planar Pushing from Pixels
di: Wei, Adam, et al.
Pubblicazione: (2025)
di: Wei, Adam, et al.
Pubblicazione: (2025)
How Well do Diffusion Policies Learn Kinematic Constraint Manifolds?
di: Foland, Lexi, et al.
Pubblicazione: (2025)
di: Foland, Lexi, et al.
Pubblicazione: (2025)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
di: Huang, Huang, et al.
Pubblicazione: (2025)
di: Huang, Huang, et al.
Pubblicazione: (2025)
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
di: Seo, Younggyo, et al.
Pubblicazione: (2025)
di: Seo, Younggyo, et al.
Pubblicazione: (2025)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
di: Seo, Younggyo, et al.
Pubblicazione: (2024)
di: Seo, Younggyo, et al.
Pubblicazione: (2024)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
di: Tian, Yi, et al.
Pubblicazione: (2022)
di: Tian, Yi, et al.
Pubblicazione: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
di: Tian, Yi, et al.
Pubblicazione: (2026)
di: Tian, Yi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
History-Guided Video Diffusion
di: Song, Kiwhan, et al.
Pubblicazione: (2025) -
Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
di: Chen, Boyuan, et al.
Pubblicazione: (2024) -
ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation
di: Heng, Liang, et al.
Pubblicazione: (2025) -
Rodrigues Network for Learning Robot Actions
di: Zhang, Jialiang, et al.
Pubblicazione: (2025) -
Selective Underfitting in Diffusion Models
di: Song, Kiwhan, et al.
Pubblicazione: (2025)