AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiong, Chuyan, Shen, Chengyu, Li, Xiaoqi, Zhou, Kaichen, Liu, Jeremy, Wang, Ruiping, Dong, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping
von: Zhang, Mingxu, et al.
Veröffentlicht: (2025)
von: Zhang, Mingxu, et al.
Veröffentlicht: (2025)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
RoboPCA: Pose-centered Affordance Learning from Human Demonstrations for Robot Manipulation
von: Xiao, Zhanqi, et al.
Veröffentlicht: (2026)
von: Xiao, Zhanqi, et al.
Veröffentlicht: (2026)
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
SEM: Enhancing Spatial Understanding for Robust Robot Manipulation
von: Lin, Xuewu, et al.
Veröffentlicht: (2025)
von: Lin, Xuewu, et al.
Veröffentlicht: (2025)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation
von: Li, Chenxuan, et al.
Veröffentlicht: (2024)
von: Li, Chenxuan, et al.
Veröffentlicht: (2024)
SIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLM
von: Zhang, Chuanrui, et al.
Veröffentlicht: (2026)
von: Zhang, Chuanrui, et al.
Veröffentlicht: (2026)
Robust MLLM Unlearning via Visual Knowledge Distillation
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
von: Wang, Yuhang, et al.
Veröffentlicht: (2025)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning
von: Gu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Gu, Tiancheng, et al.
Veröffentlicht: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under Occlusions
von: Wu, Ruihai, et al.
Veröffentlicht: (2023)
von: Wu, Ruihai, et al.
Veröffentlicht: (2023)
RPMArt: Towards Robust Perception and Manipulation for Articulated Objects
von: Wang, Junbo, et al.
Veröffentlicht: (2024)
von: Wang, Junbo, et al.
Veröffentlicht: (2024)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
ESearch-R1: Learning Cost-Aware MLLM Agents for Interactive Embodied Search via Reinforcement Learning
von: Zhou, Weijie, et al.
Veröffentlicht: (2025)
von: Zhou, Weijie, et al.
Veröffentlicht: (2025)
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
von: Rasouli, Amir, et al.
Veröffentlicht: (2025)
von: Rasouli, Amir, et al.
Veröffentlicht: (2025)
GSWorld: Closed-Loop Photo-Realistic Simulation Suite for Robotic Manipulation
von: Jiang, Guangqi, et al.
Veröffentlicht: (2025)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2025)
M4Diffuser: Multi-View Diffusion Policy with Manipulability-Aware Control for Robust Mobile Manipulation
von: Dong, Ju, et al.
Veröffentlicht: (2025)
von: Dong, Ju, et al.
Veröffentlicht: (2025)
Robotic Visual Instruction
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
Learning to Tune Like an Expert: Interpretable and Scene-Aware Navigation via MLLM Reasoning and CVAE-Based Adaptation
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yanbo, et al.
Veröffentlicht: (2025)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
IRASim: A Fine-Grained World Model for Robot Manipulation
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
von: Liu, Minghuan, et al.
Veröffentlicht: (2025)
von: Liu, Minghuan, et al.
Veröffentlicht: (2025)
LIME: Less Is More for MLLM Evaluation
von: Zhu, King, et al.
Veröffentlicht: (2024)
von: Zhu, King, et al.
Veröffentlicht: (2024)
OPENTOUCH: Bringing Full-Hand Touch to Real-World Interaction
von: Song, Yuxin Ray, et al.
Veröffentlicht: (2025)
von: Song, Yuxin Ray, et al.
Veröffentlicht: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics
von: Wei, Ziyu, et al.
Veröffentlicht: (2026)
von: Wei, Ziyu, et al.
Veröffentlicht: (2026)
Visual IRL for Human-Like Robotic Manipulation
von: Asali, Ehsan, et al.
Veröffentlicht: (2024)
von: Asali, Ehsan, et al.
Veröffentlicht: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
SkiP: When to Skip and When to Refine for Efficient Robot Manipulation
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
Phoenix: A Motion-based Self-Reflection Framework for Fine-grained Robotic Action Correction
von: Xia, Wenke, et al.
Veröffentlicht: (2025)
von: Xia, Wenke, et al.
Veröffentlicht: (2025)
Chameleon: Episodic Memory for Long-Horizon Robotic Manipulation
von: Guo, Xinying, et al.
Veröffentlicht: (2026)
von: Guo, Xinying, et al.
Veröffentlicht: (2026)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025) -
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping
von: Zhang, Mingxu, et al.
Veröffentlicht: (2025) -
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
von: Zhong, Linqing, et al.
Veröffentlicht: (2024) -
RoboPCA: Pose-centered Affordance Learning from Human Demonstrations for Robot Manipulation
von: Xiao, Zhanqi, et al.
Veröffentlicht: (2026) -
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2025)