ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Ying, Wei, Xiaobao, Chi, Xiaowei, Li, Yuming, Zhao, Zhongyu, Wang, Hao, Ma, Ningning, Lu, Ming, Zhang, Shanghang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ManipDreamer3D : Synthesizing Plausible Robotic Manipulation Video with Occupancy-aware 3D Trajectory
von: Li, Ying, et al.
Veröffentlicht: (2025)
von: Li, Ying, et al.
Veröffentlicht: (2025)
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
von: Qian, Zezhong, et al.
Veröffentlicht: (2025)
von: Qian, Zezhong, et al.
Veröffentlicht: (2025)
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
iManip: Skill-Incremental Learning for Robotic Manipulation
von: Zheng, Zexin, et al.
Veröffentlicht: (2025)
von: Zheng, Zexin, et al.
Veröffentlicht: (2025)
Action-Sketcher: From Reasoning to Action via Visual Sketches for Long-Horizon Robotic Manipulation
von: Tan, Huajie, et al.
Veröffentlicht: (2026)
von: Tan, Huajie, et al.
Veröffentlicht: (2026)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
von: Li, Puhao, et al.
Veröffentlicht: (2024)
von: Li, Puhao, et al.
Veröffentlicht: (2024)
RoboArmGS: High-Quality Robotic Arm Splatting via Bézier Curve Refinement
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation
von: Huang, Chengyue, et al.
Veröffentlicht: (2026)
von: Huang, Chengyue, et al.
Veröffentlicht: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
von: Guo, Jun, et al.
Veröffentlicht: (2025)
von: Guo, Jun, et al.
Veröffentlicht: (2025)
Manip4Care: Robotic Manipulation of Human Limbs for Solving Assistive Tasks
von: Koh, Yubin, et al.
Veröffentlicht: (2025)
von: Koh, Yubin, et al.
Veröffentlicht: (2025)
HeteroGenManip: Generalizable Manipulation For Heterogeneous Object Interactions
von: Shen, Zhenhao, et al.
Veröffentlicht: (2026)
von: Shen, Zhenhao, et al.
Veröffentlicht: (2026)
STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation
von: Tian, Yuxuan, et al.
Veröffentlicht: (2026)
von: Tian, Yuxuan, et al.
Veröffentlicht: (2026)
ChronoDreamer: Action-Conditioned World Model as an Online Simulator for Robotic Planning
von: Zhou, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenhao, et al.
Veröffentlicht: (2025)
BiPreManip: Learning Affordance-Based Bimanual Preparatory Manipulation through Anticipatory Collaboration
von: Shen, Yan, et al.
Veröffentlicht: (2026)
von: Shen, Yan, et al.
Veröffentlicht: (2026)
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
RoboDreamer: Learning Compositional World Models for Robot Imagination
von: Zhou, Siyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Siyuan, et al.
Veröffentlicht: (2024)
AdaManip: Adaptive Articulated Object Manipulation Environments and Policy Learning
von: Wang, Yuanfei, et al.
Veröffentlicht: (2025)
von: Wang, Yuanfei, et al.
Veröffentlicht: (2025)
ManipArena: Comprehensive Real-world Evaluation of Reasoning-Oriented Generalist Robot Manipulation
von: Sun, Yu, et al.
Veröffentlicht: (2026)
von: Sun, Yu, et al.
Veröffentlicht: (2026)
UniDoorManip: Learning Universal Door Manipulation Policy Over Large-scale and Diverse Door Manipulation Environments
von: Li, Yu, et al.
Veröffentlicht: (2024)
von: Li, Yu, et al.
Veröffentlicht: (2024)
VTAO-BiManip: Masked Visual-Tactile-Action Pre-training with Object Understanding for Bimanual Dexterous Manipulation
von: Sun, Zhengnan, et al.
Veröffentlicht: (2025)
von: Sun, Zhengnan, et al.
Veröffentlicht: (2025)
HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models
von: Feng, Qiuxuan, et al.
Veröffentlicht: (2026)
von: Feng, Qiuxuan, et al.
Veröffentlicht: (2026)
High-Fidelity Simulated Data Generation for Real-World Zero-Shot Robotic Manipulation Learning with Gaussian Splatting
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2025)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models
von: Huang, Siyuan, et al.
Veröffentlicht: (2024)
von: Huang, Siyuan, et al.
Veröffentlicht: (2024)
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
von: Li, Kailin, et al.
Veröffentlicht: (2025)
von: Li, Kailin, et al.
Veröffentlicht: (2025)
RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments
von: Murooka, Masaki, et al.
Veröffentlicht: (2025)
von: Murooka, Masaki, et al.
Veröffentlicht: (2025)
Mask World Model: Predicting What Matters for Robust Robot Policy Learning
von: Lou, Yunfan, et al.
Veröffentlicht: (2026)
von: Lou, Yunfan, et al.
Veröffentlicht: (2026)
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
von: Liu, Yushan, et al.
Veröffentlicht: (2026)
von: Liu, Yushan, et al.
Veröffentlicht: (2026)
GAPartManip: A Large-scale Part-centric Dataset for Material-Agnostic Articulated Object Manipulation
von: Cui, Wenbo, et al.
Veröffentlicht: (2024)
von: Cui, Wenbo, et al.
Veröffentlicht: (2024)
Unifying Perception and Action: A Hybrid-Modality Pipeline with Implicit Visual Chain-of-Thought for Robotic Action Generation
von: Ma, Xiangkai, et al.
Veröffentlicht: (2025)
von: Ma, Xiangkai, et al.
Veröffentlicht: (2025)
Long-horizon Locomotion and Manipulation on a Quadrupedal Robot with Large Language Models
von: Ouyang, Yutao, et al.
Veröffentlicht: (2024)
von: Ouyang, Yutao, et al.
Veröffentlicht: (2024)
Lang2Manip: A Tool for LLM-Based Symbolic-to-Geometric Planning for Manipulation
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
CycleManip: Enabling Cyclic Task Manipulation via Effective Historical Perception and Understanding
von: Wei, Yi-Lin, et al.
Veröffentlicht: (2025)
von: Wei, Yi-Lin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ManipDreamer3D : Synthesizing Plausible Robotic Manipulation Video with Occupancy-aware 3D Trajectory
von: Li, Ying, et al.
Veröffentlicht: (2025) -
WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation
von: Qian, Zezhong, et al.
Veröffentlicht: (2025) -
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
von: Tang, Weiliang, et al.
Veröffentlicht: (2025) -
iManip: Skill-Incremental Learning for Robotic Manipulation
von: Zheng, Zexin, et al.
Veröffentlicht: (2025) -
Action-Sketcher: From Reasoning to Action via Visual Sketches for Long-Horizon Robotic Manipulation
von: Tan, Huajie, et al.
Veröffentlicht: (2026)