Beyond Dense Futures: World Models as Structured Planners for Robotic Manipulation
Fuente:
arXiv
Salvato in:
| Autori principali: | Jin, Minghao, Liao, Mozheng, Han, Mingfei, Li, Zhihui, Chang, Xiaojun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
di: Dai, Tingjun, et al.
Pubblicazione: (2026)
di: Dai, Tingjun, et al.
Pubblicazione: (2026)
Image Generation as a Visual Planner for Robotic Manipulation
di: Pang, Ye
Pubblicazione: (2025)
di: Pang, Ye
Pubblicazione: (2025)
LatentPilot: Scene-Aware Vision-and-Language Navigation by Dreaming Ahead with Latent Visual Reasoning
di: Hao, Haihong, et al.
Pubblicazione: (2026)
di: Hao, Haihong, et al.
Pubblicazione: (2026)
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data
di: Chen, Harold Haodong, et al.
Pubblicazione: (2026)
di: Chen, Harold Haodong, et al.
Pubblicazione: (2026)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
di: Qian, Zezhong, et al.
Pubblicazione: (2025)
di: Qian, Zezhong, et al.
Pubblicazione: (2025)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
di: Chen, Yuzhi, et al.
Pubblicazione: (2026)
di: Chen, Yuzhi, et al.
Pubblicazione: (2026)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
di: Zhou, Kaichen, et al.
Pubblicazione: (2026)
di: Zhou, Kaichen, et al.
Pubblicazione: (2026)
A Step Toward World Models: A Survey on Robotic Manipulation
di: Zhang, Peng-Fei, et al.
Pubblicazione: (2025)
di: Zhang, Peng-Fei, et al.
Pubblicazione: (2025)
STORM: Search-Guided Generative World Models for Robotic Manipulation
di: Lin, Wenjun, et al.
Pubblicazione: (2025)
di: Lin, Wenjun, et al.
Pubblicazione: (2025)
Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation
di: Liao, Yue, et al.
Pubblicazione: (2025)
di: Liao, Yue, et al.
Pubblicazione: (2025)
Large Video Planner Enables Generalizable Robot Control
di: Chen, Boyuan, et al.
Pubblicazione: (2025)
di: Chen, Boyuan, et al.
Pubblicazione: (2025)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
di: Li, Yajie, et al.
Pubblicazione: (2026)
di: Li, Yajie, et al.
Pubblicazione: (2026)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
di: Duan, Jiafei, et al.
Pubblicazione: (2024)
di: Duan, Jiafei, et al.
Pubblicazione: (2024)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
di: Barcellona, Leonardo, et al.
Pubblicazione: (2024)
di: Barcellona, Leonardo, et al.
Pubblicazione: (2024)
EnerVerse: Envisioning Embodied Future Space for Robotics Manipulation
di: Huang, Siyuan, et al.
Pubblicazione: (2025)
di: Huang, Siyuan, et al.
Pubblicazione: (2025)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
di: Guo, Jun, et al.
Pubblicazione: (2025)
di: Guo, Jun, et al.
Pubblicazione: (2025)
ST-$π$: Structured SpatioTemporal VLA for Robotic Manipulation
di: Ma, Chuanhao, et al.
Pubblicazione: (2026)
di: Ma, Chuanhao, et al.
Pubblicazione: (2026)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
di: Chai, Ying, et al.
Pubblicazione: (2025)
di: Chai, Ying, et al.
Pubblicazione: (2025)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
di: Li, Huiqiong, et al.
Pubblicazione: (2026)
di: Li, Huiqiong, et al.
Pubblicazione: (2026)
MIND-V: Hierarchical World Model for Long-Horizon Robotic Manipulation with RL-based Physical Alignment
di: Zhang, Ruicheng, et al.
Pubblicazione: (2025)
di: Zhang, Ruicheng, et al.
Pubblicazione: (2025)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
di: Li, Yi, et al.
Pubblicazione: (2025)
di: Li, Yi, et al.
Pubblicazione: (2025)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
di: Huang, Wenlong, et al.
Pubblicazione: (2026)
di: Huang, Wenlong, et al.
Pubblicazione: (2026)
Causal World Modeling for Robot Control
di: Li, Lin, et al.
Pubblicazione: (2026)
di: Li, Lin, et al.
Pubblicazione: (2026)
PointSLAM++: Robust Dense Neural Gaussian Point Cloud-based SLAM
di: Wang, Xu, et al.
Pubblicazione: (2026)
di: Wang, Xu, et al.
Pubblicazione: (2026)
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation
di: Han, Songhao, et al.
Pubblicazione: (2025)
di: Han, Songhao, et al.
Pubblicazione: (2025)
Implicit Geometry Representations for Vision-and-Language Navigation from Web Videos
di: Han, Mingfei, et al.
Pubblicazione: (2026)
di: Han, Mingfei, et al.
Pubblicazione: (2026)
Occupancy World Model for Robots
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models
di: Ruan, Bo-Kai, et al.
Pubblicazione: (2026)
di: Ruan, Bo-Kai, et al.
Pubblicazione: (2026)
Robo-ABC: Affordance Generalization Beyond Categories via Semantic Correspondence for Robot Manipulation
di: Ju, Yuanchen, et al.
Pubblicazione: (2024)
di: Ju, Yuanchen, et al.
Pubblicazione: (2024)
IRASim: A Fine-Grained World Model for Robot Manipulation
di: Zhu, Fangqi, et al.
Pubblicazione: (2024)
di: Zhu, Fangqi, et al.
Pubblicazione: (2024)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
di: Yenamandra, Sriram, et al.
Pubblicazione: (2024)
di: Yenamandra, Sriram, et al.
Pubblicazione: (2024)
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
di: Li, Qixiu, et al.
Pubblicazione: (2024)
di: Li, Qixiu, et al.
Pubblicazione: (2024)
CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation
di: Hao, Haihong, et al.
Pubblicazione: (2025)
di: Hao, Haihong, et al.
Pubblicazione: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
di: Huang, Haifeng, et al.
Pubblicazione: (2025)
di: Huang, Haifeng, et al.
Pubblicazione: (2025)
Evaluating Real-World Robot Manipulation Policies in Simulation
di: Li, Xuanlin, et al.
Pubblicazione: (2024)
di: Li, Xuanlin, et al.
Pubblicazione: (2024)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
di: Wang, Guokang, et al.
Pubblicazione: (2024)
di: Wang, Guokang, et al.
Pubblicazione: (2024)
A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM
di: Han, ByungOk, et al.
Pubblicazione: (2024)
di: Han, ByungOk, et al.
Pubblicazione: (2024)
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
di: Wen, Youpeng, et al.
Pubblicazione: (2024)
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
di: Ye, Guo, et al.
Pubblicazione: (2025)
di: Ye, Guo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
di: Dai, Tingjun, et al.
Pubblicazione: (2026) -
Image Generation as a Visual Planner for Robotic Manipulation
di: Pang, Ye
Pubblicazione: (2025) -
LatentPilot: Scene-Aware Vision-and-Language Navigation by Dreaming Ahead with Latent Visual Reasoning
di: Hao, Haihong, et al.
Pubblicazione: (2026) -
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data
di: Chen, Harold Haodong, et al.
Pubblicazione: (2026) -
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
di: Qian, Zezhong, et al.
Pubblicazione: (2025)