Self-Improving Embodied Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ghasemipour, Seyed Kamyar Seyed, Wahid, Ayzaan, Tompson, Jonathan, Sanketi, Pannag, Mordatch, Igor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ALOHA Unleashed: A Simple Recipe for Robot Dexterity
von: Zhao, Tony Z., et al.
Veröffentlicht: (2024)
von: Zhao, Tony Z., et al.
Veröffentlicht: (2024)
Robot Data Curation with Mutual Information Estimators
von: Hejna, Joey, et al.
Veröffentlicht: (2025)
von: Hejna, Joey, et al.
Veröffentlicht: (2025)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
Learning Diverse Robot Striking Motions with Diffusion Models and Kinematically Constrained Gradient Guidance
von: Lee, Kin Man, et al.
Veröffentlicht: (2024)
von: Lee, Kin Man, et al.
Veröffentlicht: (2024)
Vision Language Models are In-Context Value Learners
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
von: Ahn, Michael, et al.
Veröffentlicht: (2024)
von: Ahn, Michael, et al.
Veröffentlicht: (2024)
OpenVLA: An Open-Source Vision-Language-Action Model
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation
von: ALOHA 2 Team, et al.
Veröffentlicht: (2024)
von: ALOHA 2 Team, et al.
Veröffentlicht: (2024)
Embodied Red Teaming for Auditing Robotic Foundation Models
von: Karnik, Sathwik, et al.
Veröffentlicht: (2024)
von: Karnik, Sathwik, et al.
Veröffentlicht: (2024)
Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers
von: Jain, Vidhi, et al.
Veröffentlicht: (2024)
von: Jain, Vidhi, et al.
Veröffentlicht: (2024)
Robo-DM: Data Management For Large Robot Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making
von: Wang, Yucen, et al.
Veröffentlicht: (2025)
von: Wang, Yucen, et al.
Veröffentlicht: (2025)
Pelican-VL 1.0: A Foundation Brain Model for Embodied Intelligence
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
Octo: An Open-Source Generalist Robot Policy
von: Octo Model Team, et al.
Veröffentlicht: (2024)
von: Octo Model Team, et al.
Veröffentlicht: (2024)
The Essential Role of Causality in Foundation World Models for Embodied AI
von: Gupta, Tarun, et al.
Veröffentlicht: (2024)
von: Gupta, Tarun, et al.
Veröffentlicht: (2024)
Model Adaptation for Time Constrained Embodied Control
von: Song, Jaehyun, et al.
Veröffentlicht: (2024)
von: Song, Jaehyun, et al.
Veröffentlicht: (2024)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning with Enhanced PPO for Safe Mobile Robot Navigation
von: Taheri, Hamid, et al.
Veröffentlicht: (2024)
von: Taheri, Hamid, et al.
Veröffentlicht: (2024)
Vidarc: Embodied Video Diffusion Model for Closed-loop Control
von: Feng, Yao, et al.
Veröffentlicht: (2025)
von: Feng, Yao, et al.
Veröffentlicht: (2025)
KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Temporal and Semantic Evaluation Metrics for Foundation Models in Post-Hoc Analysis of Robotic Sub-tasks
von: Salfity, Jonathan, et al.
Veröffentlicht: (2024)
von: Salfity, Jonathan, et al.
Veröffentlicht: (2024)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own
von: Ye, Weirui, et al.
Veröffentlicht: (2023)
von: Ye, Weirui, et al.
Veröffentlicht: (2023)
Integrating DeepRL with Robust Low-Level Control in Robotic Manipulators for Non-Repetitive Reaching Tasks
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2024)
von: Shahna, Mehdi Heydari, et al.
Veröffentlicht: (2024)
OrbiSim: World Models as Differentiable Physics Engines for Embodied Intelligence
von: Li, Jiajian, et al.
Veröffentlicht: (2026)
von: Li, Jiajian, et al.
Veröffentlicht: (2026)
MEM: Multi-Scale Embodied Memory for Vision Language Action Models
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
von: Guan, Weifan, et al.
Veröffentlicht: (2025)
HEAL: An Empirical Study on Hallucinations in Embodied Agents Driven by Large Language Models
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2025)
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2025)
Learning the RoPEs: Better 2D and 3D Position Encodings with STRING
von: Schenck, Connor, et al.
Veröffentlicht: (2025)
von: Schenck, Connor, et al.
Veröffentlicht: (2025)
On the Strengths and Weaknesses of Data for Open-set Embodied Assistance
von: Tambwekar, Pradyumna, et al.
Veröffentlicht: (2026)
von: Tambwekar, Pradyumna, et al.
Veröffentlicht: (2026)
From Inference Efficiency to Embodied Efficiency: Revisiting Efficiency Metrics for Vision-Language-Action Models
von: Li, Zhuofan, et al.
Veröffentlicht: (2026)
von: Li, Zhuofan, et al.
Veröffentlicht: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
VLP: Vision-Language Preference Learning for Embodied Manipulation
von: Liu, Runze, et al.
Veröffentlicht: (2025)
von: Liu, Runze, et al.
Veröffentlicht: (2025)
Robotic Control via Embodied Chain-of-Thought Reasoning
von: Zawalski, Michał, et al.
Veröffentlicht: (2024)
von: Zawalski, Michał, et al.
Veröffentlicht: (2024)
Foundation Models for Rapid Autonomy Validation
von: Farid, Alec, et al.
Veröffentlicht: (2024)
von: Farid, Alec, et al.
Veröffentlicht: (2024)
Safe Imitation Learning of Nonlinear Model Predictive Control for Flexible Robots
von: Mamedov, Shamil, et al.
Veröffentlicht: (2022)
von: Mamedov, Shamil, et al.
Veröffentlicht: (2022)
Safe Bayesian Optimization for the Control of High-Dimensional Embodied Systems
von: Wei, Yunyue, et al.
Veröffentlicht: (2024)
von: Wei, Yunyue, et al.
Veröffentlicht: (2024)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
von: Yuan, Yifu, et al.
Veröffentlicht: (2025)
von: Yuan, Yifu, et al.
Veröffentlicht: (2025)
Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation
von: Xie, Quanting, et al.
Veröffentlicht: (2024)
von: Xie, Quanting, et al.
Veröffentlicht: (2024)
ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ALOHA Unleashed: A Simple Recipe for Robot Dexterity
von: Zhao, Tony Z., et al.
Veröffentlicht: (2024) -
Robot Data Curation with Mutual Information Estimators
von: Hejna, Joey, et al.
Veröffentlicht: (2025) -
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025) -
Learning Diverse Robot Striking Motions with Diffusion Models and Kinematically Constrained Gradient Guidance
von: Lee, Kin Man, et al.
Veröffentlicht: (2024) -
Vision Language Models are In-Context Value Learners
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)