IRASim: A Fine-Grained World Model for Robot Manipulation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhu, Fangqi, Wu, Hongtao, Guo, Song, Liu, Yuxiao, Cheang, Chilam, Kong, Tao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GR-3 Technical Report
por: Cheang, Chilam, et al.
Publicado: (2025)
por: Cheang, Chilam, et al.
Publicado: (2025)
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
por: Li, Peiyan, et al.
Publicado: (2024)
por: Li, Peiyan, et al.
Publicado: (2024)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
por: Huang, Wenlong, et al.
Publicado: (2026)
por: Huang, Wenlong, et al.
Publicado: (2026)
GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
por: Cheang, Chi-Lam, et al.
Publicado: (2024)
por: Cheang, Chi-Lam, et al.
Publicado: (2024)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
por: Li, Yi, et al.
Publicado: (2025)
por: Li, Yi, et al.
Publicado: (2025)
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
por: Liu, Minghuan, et al.
Publicado: (2025)
por: Liu, Minghuan, et al.
Publicado: (2025)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
por: Ren, Pengzhen, et al.
Publicado: (2023)
por: Ren, Pengzhen, et al.
Publicado: (2023)
Physically Grounded Vision-Language Models for Robotic Manipulation
por: Gao, Jensen, et al.
Publicado: (2023)
por: Gao, Jensen, et al.
Publicado: (2023)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
por: Liu, Yijun, et al.
Publicado: (2025)
por: Liu, Yijun, et al.
Publicado: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
por: Jiang, Guangqi, et al.
Publicado: (2024)
por: Jiang, Guangqi, et al.
Publicado: (2024)
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
por: Lu, Haoran, et al.
Publicado: (2026)
por: Lu, Haoran, et al.
Publicado: (2026)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
por: Rasouli, Amir, et al.
Publicado: (2025)
por: Rasouli, Amir, et al.
Publicado: (2025)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
por: Lu, Guanxing, et al.
Publicado: (2025)
por: Lu, Guanxing, et al.
Publicado: (2025)
Object-Centric World Model for Language-Guided Manipulation
por: Jeong, Youngjoon, et al.
Publicado: (2025)
por: Jeong, Youngjoon, et al.
Publicado: (2025)
Chameleon: Episodic Memory for Long-Horizon Robotic Manipulation
por: Guo, Xinying, et al.
Publicado: (2026)
por: Guo, Xinying, et al.
Publicado: (2026)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
por: Wu, Yiming, et al.
Publicado: (2025)
por: Wu, Yiming, et al.
Publicado: (2025)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
por: Lancaster, Patrick, et al.
Publicado: (2023)
por: Lancaster, Patrick, et al.
Publicado: (2023)
SkiP: When to Skip and When to Refine for Efficient Robot Manipulation
por: Dai, Mingtong, et al.
Publicado: (2026)
por: Dai, Mingtong, et al.
Publicado: (2026)
Robot Learning from a Physical World Model
por: Mao, Jiageng, et al.
Publicado: (2025)
por: Mao, Jiageng, et al.
Publicado: (2025)
BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities
por: Jiang, Yunfan, et al.
Publicado: (2025)
por: Jiang, Yunfan, et al.
Publicado: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
por: Shen, Yichao, et al.
Publicado: (2025)
por: Shen, Yichao, et al.
Publicado: (2025)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
por: Tang, Yihe, et al.
Publicado: (2025)
por: Tang, Yihe, et al.
Publicado: (2025)
Language-Conditioned World Modeling for Visual Navigation
por: Dong, Yifei, et al.
Publicado: (2026)
por: Dong, Yifei, et al.
Publicado: (2026)
SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation
por: Li, Xin, et al.
Publicado: (2024)
por: Li, Xin, et al.
Publicado: (2024)
PIVOT-R: Primitive-Driven Waypoint-Aware World Model for Robotic Manipulation
por: Zhang, Kaidong, et al.
Publicado: (2024)
por: Zhang, Kaidong, et al.
Publicado: (2024)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
por: Kim, Ju-Young, et al.
Publicado: (2025)
por: Kim, Ju-Young, et al.
Publicado: (2025)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
por: Xu, Wenjiang, et al.
Publicado: (2025)
por: Xu, Wenjiang, et al.
Publicado: (2025)
GSWorld: Closed-Loop Photo-Realistic Simulation Suite for Robotic Manipulation
por: Jiang, Guangqi, et al.
Publicado: (2025)
por: Jiang, Guangqi, et al.
Publicado: (2025)
ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics
por: Wei, Ziyu, et al.
Publicado: (2026)
por: Wei, Ziyu, et al.
Publicado: (2026)
Visual IRL for Human-Like Robotic Manipulation
por: Asali, Ehsan, et al.
Publicado: (2024)
por: Asali, Ehsan, et al.
Publicado: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
por: Yenamandra, Sriram, et al.
Publicado: (2023)
por: Yenamandra, Sriram, et al.
Publicado: (2023)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
por: Xiong, Chuyan, et al.
Publicado: (2024)
por: Xiong, Chuyan, et al.
Publicado: (2024)
SEM: Enhancing Spatial Understanding for Robust Robot Manipulation
por: Lin, Xuewu, et al.
Publicado: (2025)
por: Lin, Xuewu, et al.
Publicado: (2025)
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
por: Zhang, Kaidong, et al.
Publicado: (2025)
por: Zhang, Kaidong, et al.
Publicado: (2025)
Geometry-aware 4D Video Generation for Robot Manipulation
por: Liu, Zeyi, et al.
Publicado: (2025)
por: Liu, Zeyi, et al.
Publicado: (2025)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
por: Patel, Shivansh, et al.
Publicado: (2025)
por: Patel, Shivansh, et al.
Publicado: (2025)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
por: Chen, Shizhe, et al.
Publicado: (2025)
por: Chen, Shizhe, et al.
Publicado: (2025)
AnyPlace: Learning Generalized Object Placement for Robot Manipulation
por: Zhao, Yuchi, et al.
Publicado: (2025)
por: Zhao, Yuchi, et al.
Publicado: (2025)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
por: Dharmarajan, Karthik, et al.
Publicado: (2025)
por: Dharmarajan, Karthik, et al.
Publicado: (2025)
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation
por: Lian, Shijie, et al.
Publicado: (2026)
por: Lian, Shijie, et al.
Publicado: (2026)
Ejemplares similares
-
GR-3 Technical Report
por: Cheang, Chilam, et al.
Publicado: (2025) -
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
por: Li, Peiyan, et al.
Publicado: (2024) -
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
por: Huang, Wenlong, et al.
Publicado: (2026) -
GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
por: Cheang, Chi-Lam, et al.
Publicado: (2024) -
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
por: Li, Yi, et al.
Publicado: (2025)