iMoWM: Taming Interactive Multi-Modal World Model for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Chuanrui, Wu, Zhengxian, Lu, Guanxing, Tang, Yansong, Wang, Ziwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ManiGaussian: Dynamic Gaussian Splatting for Multi-task Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation
by: Xue, Yuquan, et al.
Published: (2025)
by: Xue, Yuquan, et al.
Published: (2025)
ManiGaussian++: General Robotic Bimanual Manipulation with Hierarchical Gaussian World Model
by: Yu, Tengbo, et al.
Published: (2025)
by: Yu, Tengbo, et al.
Published: (2025)
Human-in-the-loop Online Rejection Sampling for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
ManiCM: Real-time 3D Diffusion Policy via Consistency Model for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
AnyBimanual: Transferring Unimanual Policy for General Bimanual Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
by: Guo, Wenkai, et al.
Published: (2025)
by: Guo, Wenkai, et al.
Published: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
by: Jiang, Feng, et al.
Published: (2026)
by: Jiang, Feng, et al.
Published: (2026)
$τ_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
by: Zhou, Pengfei, et al.
Published: (2026)
by: Zhou, Pengfei, et al.
Published: (2026)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
H-WM: Robotic Task and Motion Planning Guided by Hierarchical World Model
by: Huang, Jinbang, et al.
Published: (2026)
by: Huang, Jinbang, et al.
Published: (2026)
PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation
by: Li, Wenxuan, et al.
Published: (2025)
by: Li, Wenxuan, et al.
Published: (2025)
LaDi-WM: A Latent Diffusion-based World Model for Predictive Manipulation
by: Huang, Yuhang, et al.
Published: (2025)
by: Huang, Yuhang, et al.
Published: (2025)
MoTo: A Zero-shot Plug-in Interaction-aware Navigation for General Mobile Manipulation
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
WM-DAgger: Enabling Efficient Data Aggregation for Imitation Learning with World Models
by: Yu, Anlan, et al.
Published: (2026)
by: Yu, Anlan, et al.
Published: (2026)
ResWM: Residual-Action World Model for Visual RL
by: Zhang, Jseen, et al.
Published: (2026)
by: Zhang, Jseen, et al.
Published: (2026)
World Models for Robotic Manipulation: A Survey
by: Wang, Fangyuan, et al.
Published: (2026)
by: Wang, Fangyuan, et al.
Published: (2026)
AdaWM: Adaptive World Model based Planning for Autonomous Driving
by: Wang, Hang, et al.
Published: (2025)
by: Wang, Hang, et al.
Published: (2025)
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
by: Zhang, Yizheng, et al.
Published: (2025)
by: Zhang, Yizheng, et al.
Published: (2025)
iManip: Skill-Incremental Learning for Robotic Manipulation
by: Zheng, Zexin, et al.
Published: (2025)
by: Zheng, Zexin, et al.
Published: (2025)
DDP-WM: Disentangled Dynamics Prediction for Efficient World Models
by: Yin, Shicheng, et al.
Published: (2026)
by: Yin, Shicheng, et al.
Published: (2026)
Learning Robot Manipulation from Audio World Models
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
MP1: MeanFlow Tames Policy Learning in 1-step for Robotic Manipulation
by: Sheng, Juyi, et al.
Published: (2025)
by: Sheng, Juyi, et al.
Published: (2025)
iLearnRobot: An Interactive Learning-Based Multi-Modal Robot with Continuous Improvement
by: Wang, Kohou, et al.
Published: (2025)
by: Wang, Kohou, et al.
Published: (2025)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
Learning Multi-Modal Trajectory Policies for Data-Efficient Robotic Manipulation
by: Chen, Zijia, et al.
Published: (2026)
by: Chen, Zijia, et al.
Published: (2026)
ContactGaussian-WM: Learning Physics-Grounded World Model from Videos
by: Wang, Meizhong, et al.
Published: (2026)
by: Wang, Meizhong, et al.
Published: (2026)
Say, Dream, and Act: Learning Video World Models for Instruction-Driven Robot Manipulation
by: Gu, Songen, et al.
Published: (2026)
by: Gu, Songen, et al.
Published: (2026)
Behavior Tree Generation using Large Language Models for Sequential Manipulation Planning with Human Instructions and Feedback
by: Ao, Jicong, et al.
Published: (2024)
by: Ao, Jicong, et al.
Published: (2024)
Hybrid Consistency Policy: Decoupling Multi-Modal Diversity and Real-Time Efficiency in Robotic Manipulation
by: Zhao, Qianyou, et al.
Published: (2025)
by: Zhao, Qianyou, et al.
Published: (2025)
Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation
by: Song, Zijian, et al.
Published: (2026)
by: Song, Zijian, et al.
Published: (2026)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
by: Chen, Yuzhi, et al.
Published: (2026)
by: Chen, Yuzhi, et al.
Published: (2026)
OSVI-WM: One-Shot Visual Imitation for Unseen Tasks using World-Model-Guided Trajectory Generation
by: Goswami, Raktim Gautam, et al.
Published: (2025)
by: Goswami, Raktim Gautam, et al.
Published: (2025)
ParticleFormer: A 3D Point Cloud World Model for Multi-Object, Multi-Material Robotic Manipulation
by: Huang, Suning, et al.
Published: (2025)
by: Huang, Suning, et al.
Published: (2025)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
by: Zhou, Gaoyue, et al.
Published: (2024)
by: Zhou, Gaoyue, et al.
Published: (2024)
STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation
by: Tian, Yuxuan, et al.
Published: (2026)
by: Tian, Yuxuan, et al.
Published: (2026)
Multi-Task Interactive Robot Fleet Learning with Visual World Models
by: Liu, Huihan, et al.
Published: (2024)
by: Liu, Huihan, et al.
Published: (2024)
Design and Benchmarking of A Multi-Modality Sensor for Robotic Manipulation with GAN-Based Cross-Modality Interpretation
by: Zhang, Dandan, et al.
Published: (2025)
by: Zhang, Dandan, et al.
Published: (2025)
Similar Items
-
ManiGaussian: Dynamic Gaussian Splatting for Multi-task Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024) -
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025) -
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation
by: Xue, Yuquan, et al.
Published: (2025) -
ManiGaussian++: General Robotic Bimanual Manipulation with Hierarchical Gaussian World Model
by: Yu, Tengbo, et al.
Published: (2025) -
Human-in-the-loop Online Rejection Sampling for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025)