IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xiaoyu, Guo, Junliang, He, Tianyu, Zhang, Chuheng, Zhang, Pushi, Yang, Derek Cathera, Zhao, Li, Bian, Jiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
How Do VLAs Effectively Inherit from VLMs?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
PIG-Nav: Key Insights for Pretrained Image Goal Navigation Models
von: Wan, Jiansong, et al.
Veröffentlicht: (2025)
von: Wan, Jiansong, et al.
Veröffentlicht: (2025)
Beyond Human Demonstrations: Diffusion-Based Reinforcement Learning to Generate Data for VLA Training
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
RynnBrain: Open Embodied Foundation Models
von: Dang, Ronghao, et al.
Veröffentlicht: (2026)
von: Dang, Ronghao, et al.
Veröffentlicht: (2026)
Body Discovery of Embodied AI
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
von: Sun, Zhe, et al.
Veröffentlicht: (2025)
Embodied Navigation Foundation Model
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2025)
What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models
von: Peng, Yuanfang, et al.
Veröffentlicht: (2026)
von: Peng, Yuanfang, et al.
Veröffentlicht: (2026)
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
Embodied Arena: A Comprehensive, Unified, and Evolving Evaluation Platform for Embodied AI
von: Ni, Fei, et al.
Veröffentlicht: (2025)
von: Ni, Fei, et al.
Veröffentlicht: (2025)
What Do Latent Action Models Actually Learn?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
Image Quality Assessment for Embodied AI
von: Li, Chunyi, et al.
Veröffentlicht: (2025)
von: Li, Chunyi, et al.
Veröffentlicht: (2025)
A Brain-inspired Embodied Intelligence for Fluid and Fast Reflexive Robotics Control
von: Guo, Weiyu, et al.
Veröffentlicht: (2026)
von: Guo, Weiyu, et al.
Veröffentlicht: (2026)
Learning Additively Compositional Latent Actions for Embodied AI
von: Wei, Hangxing, et al.
Veröffentlicht: (2026)
von: Wei, Hangxing, et al.
Veröffentlicht: (2026)
Modeling the Mental World for Embodied AI: A Comprehensive Review
von: Liu, Biyuan, et al.
Veröffentlicht: (2025)
von: Liu, Biyuan, et al.
Veröffentlicht: (2025)
Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning
von: Liang, Wenlong, et al.
Veröffentlicht: (2025)
von: Liang, Wenlong, et al.
Veröffentlicht: (2025)
Embodied Visuomotor Representation
von: Burner, Levi, et al.
Veröffentlicht: (2024)
von: Burner, Levi, et al.
Veröffentlicht: (2024)
EO-1: An Open Unified Embodied Foundation Model for General Robot Control
von: Qu, Delin, et al.
Veröffentlicht: (2025)
von: Qu, Delin, et al.
Veröffentlicht: (2025)
A Survey on Robotics with Foundation Models: toward Embodied AI
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
Embodied Robot Manipulation in the Era of Foundation Models: Planning and Learning Perspectives
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation
von: Li, Dongjiang, et al.
Veröffentlicht: (2025)
von: Li, Dongjiang, et al.
Veröffentlicht: (2025)
DynSyn: Dynamical Synergistic Representation for Efficient Learning and Control in Overactuated Embodied Systems
von: He, Kaibo, et al.
Veröffentlicht: (2024)
von: He, Kaibo, et al.
Veröffentlicht: (2024)
DM0: An Embodied-Native Vision-Language-Action Model towards Physical AI
von: Yu, En, et al.
Veröffentlicht: (2026)
von: Yu, En, et al.
Veröffentlicht: (2026)
Asking Before Acting: Gather Information in Embodied Decision Making with Language Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2023)
Joint Moment Estimation for Hip Exoskeleton Control: A Generalized Moment Feature Generation Method
von: Zhang, Yuanwen, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanwen, et al.
Veröffentlicht: (2024)
The Essential Role of Causality in Foundation World Models for Embodied AI
von: Gupta, Tarun, et al.
Veröffentlicht: (2024)
von: Gupta, Tarun, et al.
Veröffentlicht: (2024)
MIPD: A Multi-sensory Interactive Perception Dataset for Embodied Intelligent Driving
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
Unified Noise Steering for Efficient Human-Guided VLA Adaptation
von: Lu, Junjie, et al.
Veröffentlicht: (2026)
von: Lu, Junjie, et al.
Veröffentlicht: (2026)
MiMo-Embodied: X-Embodied Foundation Model Technical Report
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
UniDex: A Robot Foundation Suite for Universal Dexterous Hand Control from Egocentric Human Videos
von: Zhang, Gu, et al.
Veröffentlicht: (2026)
von: Zhang, Gu, et al.
Veröffentlicht: (2026)
INF‐SLiM: Large‐Scale Implicit Neural Fields for Semantic LiDAR Mapping of Embodied AI Agents
von: Jianyuan Zhang, et al.
Veröffentlicht: (2025)
von: Jianyuan Zhang, et al.
Veröffentlicht: (2025)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
Transforming Monolithic Foundation Models into Embodied Multi-Agent Architectures for Human-Robot Collaboration
von: Sun, Nan, et al.
Veröffentlicht: (2025)
von: Sun, Nan, et al.
Veröffentlicht: (2025)
BEINGS: Bayesian Embodied Image-goal Navigation with Gaussian Splatting
von: Meng, Wugang, et al.
Veröffentlicht: (2024)
von: Meng, Wugang, et al.
Veröffentlicht: (2024)
Self-Improving Embodied Foundation Models
von: Ghasemipour, Seyed Kamyar Seyed, et al.
Veröffentlicht: (2025)
von: Ghasemipour, Seyed Kamyar Seyed, et al.
Veröffentlicht: (2025)
Embodied Intelligence for Flexible Manufacturing: A Survey
von: Xu, Kai, et al.
Veröffentlicht: (2025)
von: Xu, Kai, et al.
Veröffentlicht: (2025)
Embodied Learning of Reward for Musculoskeletal Control with Vision Language Models
von: Soedarmadji, Saraswati, et al.
Veröffentlicht: (2025)
von: Soedarmadji, Saraswati, et al.
Veröffentlicht: (2025)
LCMF: Lightweight Cross-Modality Mambaformer for Embodied Robotics VQA
von: Kang, Zeyi, et al.
Veröffentlicht: (2025)
von: Kang, Zeyi, et al.
Veröffentlicht: (2025)
Safety Control of Service Robots with LLMs and Embodied Knowledge Graphs
von: Qi, Yong, et al.
Veröffentlicht: (2024)
von: Qi, Yong, et al.
Veröffentlicht: (2024)
Wanderland: Geometrically Grounded Simulation for Open-World Embodied AI
von: Liu, Xinhao, et al.
Veröffentlicht: (2025)
von: Liu, Xinhao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025) -
How Do VLAs Effectively Inherit from VLMs?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025) -
PIG-Nav: Key Insights for Pretrained Image Goal Navigation Models
von: Wan, Jiansong, et al.
Veröffentlicht: (2025) -
Beyond Human Demonstrations: Diffusion-Based Reinforcement Learning to Generate Data for VLA Training
von: Yang, Rushuai, et al.
Veröffentlicht: (2025) -
RynnBrain: Open Embodied Foundation Models
von: Dang, Ronghao, et al.
Veröffentlicht: (2026)