AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Ahn, Michael, Dwibedi, Debidatta, Finn, Chelsea, Arenas, Montse Gonzalez, Gopalakrishnan, Keerthana, Hausman, Karol, Ichter, Brian, Irpan, Alex, Joshi, Nikhil, Julian, Ryan, Kirmani, Sean, Leal, Isabel, Lee, Edward, Levine, Sergey, Lu, Yao, Maddineni, Sharath, Rao, Kanishka, Sadigh, Dorsa, Sanketi, Pannag, Sermanet, Pierre, Vuong, Quan, Welker, Stefan, Xia, Fei, Xiao, Ted, Xu, Peng, Xu, Steve, Xu, Zhuo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RT-H: Action Hierarchies Using Language
by: Belkhale, Suneel, et al.
Published: (2024)
by: Belkhale, Suneel, et al.
Published: (2024)
RT-Sketch: Goal-Conditioned Imitation Learning from Hand-Drawn Sketches
by: Sundaresan, Priya, et al.
Published: (2024)
by: Sundaresan, Priya, et al.
Published: (2024)
Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers
by: Jain, Vidhi, et al.
Published: (2024)
by: Jain, Vidhi, et al.
Published: (2024)
Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
by: Bharadhwaj, Homanga, et al.
Published: (2024)
by: Bharadhwaj, Homanga, et al.
Published: (2024)
A Short Note on Evaluating RepNet for Temporal Repetition Counting in Videos
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
by: Nasiriany, Soroush, et al.
Published: (2024)
by: Nasiriany, Soroush, et al.
Published: (2024)
SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities
by: Chen, Boyuan, et al.
Published: (2024)
by: Chen, Boyuan, et al.
Published: (2024)
Robot Data Curation with Mutual Information Estimators
by: Hejna, Joey, et al.
Published: (2025)
by: Hejna, Joey, et al.
Published: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
by: Gao, Jensen, et al.
Published: (2023)
by: Gao, Jensen, et al.
Published: (2023)
Chain of Code: Reasoning with a Language Model-Augmented Code Emulator
by: Li, Chengshu, et al.
Published: (2023)
by: Li, Chengshu, et al.
Published: (2023)
OVR: A Dataset for Open Vocabulary Temporal Repetition Counting in Videos
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
What's the Move? Hybrid Imitation Learning via Salient Points
by: Sundaresan, Priya, et al.
Published: (2024)
by: Sundaresan, Priya, et al.
Published: (2024)
Octo: An Open-Source Generalist Robot Policy
by: Octo Model Team, et al.
Published: (2024)
by: Octo Model Team, et al.
Published: (2024)
Abordagem do impacto psicossocial no adoecer da mama
by: Isabel Leal
Published: (2004)
by: Isabel Leal
Published: (2004)
BONDING AND PREMATURITY: EXPLORATORY STUDY ON EARLY PATERNAL INVOLVEMENT IN HOSPITALIZATION CONTEXTS
by: Isabel Leal
Published: (2014)
by: Isabel Leal
Published: (2014)
STEER: Flexible Robotic Manipulation via Dense Language Grounding
by: Smith, Laura, et al.
Published: (2024)
by: Smith, Laura, et al.
Published: (2024)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
by: Gao, Jensen, et al.
Published: (2024)
by: Gao, Jensen, et al.
Published: (2024)
OpenVLA: An Open-Source Vision-Language-Action Model
by: Kim, Moo Jin, et al.
Published: (2024)
by: Kim, Moo Jin, et al.
Published: (2024)
FlexCap: Describe Anything in Images in Controllable Detail
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
Iconismo y narratario competente en El nombre de la Rosa
by: Isabel Leal F
Published: (2012)
by: Isabel Leal F
Published: (2012)
ESTUDO DA ESCALA DE DEPRESSÃO, ANSIEDADE E STRESSE PARA CRIANÇAS (EADS- C)
by: Isabel P. Leal
Published: (2009)
by: Isabel P. Leal
Published: (2009)
EDITORIAL
by: Isabel P. Leal
Published: (2009)
by: Isabel P. Leal
Published: (2009)
A esperança nos pais de crianças com cancro. uma análise fenomenológica interpretativa da relação com profissionais de saúde.
by: Isabel Pereira Leal
Published: (2001)
by: Isabel Pereira Leal
Published: (2001)
Self-Improving Embodied Foundation Models
by: Ghasemipour, Seyed Kamyar Seyed, et al.
Published: (2025)
by: Ghasemipour, Seyed Kamyar Seyed, et al.
Published: (2025)
HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies
by: Xie, Amber, et al.
Published: (2026)
by: Xie, Amber, et al.
Published: (2026)
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)
by: Bıyık, Erdem, et al.
Published: (2024)
How to Train Your Robots? The Impact of Demonstration Modality on Imitation Learning
by: Li, Haozhuo, et al.
Published: (2025)
by: Li, Haozhuo, et al.
Published: (2025)
Imitation Bootstrapped Reinforcement Learning
by: Hu, Hengyuan, et al.
Published: (2023)
by: Hu, Hengyuan, et al.
Published: (2023)
Invariance Co-training for Robot Visual Generalization
by: Yang, Jonathan, et al.
Published: (2025)
by: Yang, Jonathan, et al.
Published: (2025)
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning
by: Clark, Jaden, et al.
Published: (2025)
by: Clark, Jaden, et al.
Published: (2025)
Data Analogies Enable Efficient Cross-Embodiment Transfer
by: Yang, Jonathan, et al.
Published: (2026)
by: Yang, Jonathan, et al.
Published: (2026)
Polychromic Objectives for Reinforcement Learning
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
Evaluating Gemini Robotics Policies in a Veo World Simulator
by: Gemini Robotics Team, et al.
Published: (2025)
by: Gemini Robotics Team, et al.
Published: (2025)
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
by: Nasiriany, Soroush, et al.
Published: (2024)
by: Nasiriany, Soroush, et al.
Published: (2024)
VADER: Visual Affordance Detection and Error Recovery for Multi Robot Human Collaboration
by: Ahn, Michael, et al.
Published: (2024)
by: Ahn, Michael, et al.
Published: (2024)
Action-Free Reasoning for Policy Generalization
by: Clark, Jaden, et al.
Published: (2025)
by: Clark, Jaden, et al.
Published: (2025)
Latent Diffusion Planning for Imitation Learning
by: Xie, Amber, et al.
Published: (2025)
by: Xie, Amber, et al.
Published: (2025)
Unified Video Action Model
by: Li, Shuang, et al.
Published: (2025)
by: Li, Shuang, et al.
Published: (2025)
Similar Items
-
RT-H: Action Hierarchies Using Language
by: Belkhale, Suneel, et al.
Published: (2024) -
RT-Sketch: Goal-Conditioned Imitation Learning from Hand-Drawn Sketches
by: Sundaresan, Priya, et al.
Published: (2024) -
Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers
by: Jain, Vidhi, et al.
Published: (2024) -
Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
by: Bharadhwaj, Homanga, et al.
Published: (2024) -
A Short Note on Evaluating RepNet for Temporal Repetition Counting in Videos
by: Dwibedi, Debidatta, et al.
Published: (2024)