MMRo: Are Multimodal LLMs Eligible as the Brain for In-Home Robotics?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jinming, Zhu, Yichen, Xu, Zhiyuan, Gu, Jindong, Zhu, Minjie, Liu, Xin, Liu, Ning, Peng, Yaxin, Feng, Feifei, Tang, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
Object-Centric Instruction Augmentation for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
Visual Robotic Manipulation with Depth-Aware Pretraining
von: Wang, Wanying, et al.
Veröffentlicht: (2024)
von: Wang, Wanying, et al.
Veröffentlicht: (2024)
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
CoA-VLA: Improving Vision-Language-Action Models via Visual-Textual Chain-of-Affordance
von: Li, Jinming, et al.
Veröffentlicht: (2024)
von: Li, Jinming, et al.
Veröffentlicht: (2024)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
Let Me Show You: Learning by Retrieving from Egocentric Video for Robotic Manipulation
von: Zhu, Yichen, et al.
Veröffentlicht: (2025)
von: Zhu, Yichen, et al.
Veröffentlicht: (2025)
Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
A Survey on Robotics with Foundation Models: toward Embodied AI
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
Learning from Imperfect Demonstrations with Self-Supervision for Robotic Manipulation
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt
von: Zhu, Xiang, et al.
Veröffentlicht: (2025)
von: Zhu, Xiang, et al.
Veröffentlicht: (2025)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
HACTS: a Human-As-Copilot Teleoperation System for Robot Learning
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Embodied Agents
von: Zhu, Yichen, et al.
Veröffentlicht: (2024)
von: Zhu, Yichen, et al.
Veröffentlicht: (2024)
Load-Aware Locomotion Control for Humanoid Robots in Industrial Transportation Tasks
von: Fu, Lequn, et al.
Veröffentlicht: (2026)
von: Fu, Lequn, et al.
Veröffentlicht: (2026)
Designing Robots to Support Parent-Child Connections: Opportunities Through Robot-Mediated Communication
von: Xu, Michael F, et al.
Veröffentlicht: (2026)
von: Xu, Michael F, et al.
Veröffentlicht: (2026)
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
Goal State Generation for Robotic Manipulation Based on Linguistically Guided Hybrid Gaussian Diffusion
von: Xu, Yichen, et al.
Veröffentlicht: (2024)
von: Xu, Yichen, et al.
Veröffentlicht: (2024)
Bringing Robots Home: The Rise of AI Robots in Consumer Electronics
von: Dong, Haiwei, et al.
Veröffentlicht: (2024)
von: Dong, Haiwei, et al.
Veröffentlicht: (2024)
MotuBrain: An Advanced World Action Model for Robot Control
von: MotuBrain Team, et al.
Veröffentlicht: (2026)
von: MotuBrain Team, et al.
Veröffentlicht: (2026)
A Microgravity Simulation Experimental Platform For Small Space Robots In Orbit
von: Luo, Hang, et al.
Veröffentlicht: (2025)
von: Luo, Hang, et al.
Veröffentlicht: (2025)
Designing Telepresence Robots to Support Place Attachment
von: Hu, Yaxin, et al.
Veröffentlicht: (2025)
von: Hu, Yaxin, et al.
Veröffentlicht: (2025)
BadRobot: Jailbreaking Embodied LLMs in the Physical World
von: Zhang, Hangtao, et al.
Veröffentlicht: (2024)
von: Zhang, Hangtao, et al.
Veröffentlicht: (2024)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching
von: Chen, Xinwang, et al.
Veröffentlicht: (2024)
von: Chen, Xinwang, et al.
Veröffentlicht: (2024)
RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
von: Wang, Xinhua, et al.
Veröffentlicht: (2026)
von: Wang, Xinhua, et al.
Veröffentlicht: (2026)
Modelling and Optimization of Magnetic Navigation Systems for Passive Robots in Minimally Invasive Brain Surgery
von: Xu Tang, et al.
Veröffentlicht: (2025)
von: Xu Tang, et al.
Veröffentlicht: (2025)
LLMs for Coding and Robotics Education
von: Shu, Peng, et al.
Veröffentlicht: (2024)
von: Shu, Peng, et al.
Veröffentlicht: (2024)
OmniNxt: A Fully Open-source and Compact Aerial Robot with Omnidirectional Visual Perception
von: Liu, Peize, et al.
Veröffentlicht: (2024)
von: Liu, Peize, et al.
Veröffentlicht: (2024)
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
Failure Mechanisms and Risk Estimation for Legged Robot Locomotion on Granular Slopes
von: Liao, Xingjue, et al.
Veröffentlicht: (2026)
von: Liao, Xingjue, et al.
Veröffentlicht: (2026)
Embodiment Transfer Learning for Vision-Language-Action Models
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
The Effects of Communication Delay on Human Performance and Neurocognitive Responses in Mobile Robot Teleoperation
von: Chen, Zhaokun, et al.
Veröffentlicht: (2025)
von: Chen, Zhaokun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
von: Zhu, Minjie, et al.
Veröffentlicht: (2024) -
Object-Centric Instruction Augmentation for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024) -
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
von: Zhu, Minjie, et al.
Veröffentlicht: (2024) -
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024) -
Visual Robotic Manipulation with Depth-Aware Pretraining
von: Wang, Wanying, et al.
Veröffentlicht: (2024)