Robotic Control via Embodied Chain-of-Thought Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zawalski, Michał, Chen, William, Pertsch, Karl, Mees, Oier, Finn, Chelsea, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FAST: Efficient Action Tokenization for Vision-Language-Action Models
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
von: Pertsch, Karl, et al.
Veröffentlicht: (2025)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
Training Strategies for Efficient Embodied Reasoning
von: Chen, William, et al.
Veröffentlicht: (2025)
von: Chen, William, et al.
Veröffentlicht: (2025)
Scaling Cross-Embodied Learning: One Policy for Manipulation, Navigation, Locomotion and Aviation
von: Doshi, Ria, et al.
Veröffentlicht: (2024)
von: Doshi, Ria, et al.
Veröffentlicht: (2024)
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
The Ingredients for Robotic Diffusion Transformers
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
Octo: An Open-Source Generalist Robot Policy
von: Octo Model Team, et al.
Veröffentlicht: (2024)
von: Octo Model Team, et al.
Veröffentlicht: (2024)
Evaluating Real-World Robot Manipulation Policies in Simulation
von: Li, Xuanlin, et al.
Veröffentlicht: (2024)
von: Li, Xuanlin, et al.
Veröffentlicht: (2024)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
Affordance-Guided Reinforcement Learning via Visual Prompting
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
von: Lee, Olivia Y., et al.
Veröffentlicht: (2024)
Adapt On-the-Go: Behavior Modulation for Single-Life Robot Deployment
von: Chen, Annie S., et al.
Veröffentlicht: (2023)
von: Chen, Annie S., et al.
Veröffentlicht: (2023)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
von: Lee, Tony, et al.
Veröffentlicht: (2026)
von: Lee, Tony, et al.
Veröffentlicht: (2026)
MEM: Multi-Scale Embodied Memory for Vision Language Action Models
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
von: Torne, Marcel, et al.
Veröffentlicht: (2026)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
MemER: Scaling Up Memory for Robot Control via Experience Retrieval
von: Sridhar, Ajay, et al.
Veröffentlicht: (2025)
von: Sridhar, Ajay, et al.
Veröffentlicht: (2025)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
von: Chen, William, et al.
Veröffentlicht: (2024)
von: Chen, William, et al.
Veröffentlicht: (2024)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
von: Pai, Jonas, et al.
Veröffentlicht: (2025)
von: Pai, Jonas, et al.
Veröffentlicht: (2025)
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control
von: Chen, William, et al.
Veröffentlicht: (2026)
von: Chen, William, et al.
Veröffentlicht: (2026)
$π_0$: A Vision-Language-Action Flow Model for General Robot Control
von: Black, Kevin, et al.
Veröffentlicht: (2024)
von: Black, Kevin, et al.
Veröffentlicht: (2024)
Multimodal Spatial Language Maps for Robot Navigation and Manipulation
von: Huang, Chenguang, et al.
Veröffentlicht: (2025)
von: Huang, Chenguang, et al.
Veröffentlicht: (2025)
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
von: Jones, Joshua, et al.
Veröffentlicht: (2025)
von: Jones, Joshua, et al.
Veröffentlicht: (2025)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
von: Kareer, Simar, et al.
Veröffentlicht: (2025)
von: Kareer, Simar, et al.
Veröffentlicht: (2025)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
von: Gao, Jensen, et al.
Veröffentlicht: (2024)
von: Gao, Jensen, et al.
Veröffentlicht: (2024)
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
von: Blank, Nils, et al.
Veröffentlicht: (2024)
von: Blank, Nils, et al.
Veröffentlicht: (2024)
Chain-of-Thought Predictive Control
von: Jia, Zhiwei, et al.
Veröffentlicht: (2023)
von: Jia, Zhiwei, et al.
Veröffentlicht: (2023)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
von: Yuan, Yifu, et al.
Veröffentlicht: (2025)
von: Yuan, Yifu, et al.
Veröffentlicht: (2025)
OpenVLA: An Open-Source Vision-Language-Action Model
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
von: Kim, Moo Jin, et al.
Veröffentlicht: (2024)
BridgeData V2: A Dataset for Robot Learning at Scale
von: Walke, Homer, et al.
Veröffentlicht: (2023)
von: Walke, Homer, et al.
Veröffentlicht: (2023)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
von: Xu, Charles, et al.
Veröffentlicht: (2024)
von: Xu, Charles, et al.
Veröffentlicht: (2024)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
von: Zhao, Qingqing, et al.
Veröffentlicht: (2025)
von: Zhao, Qingqing, et al.
Veröffentlicht: (2025)
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
Robust Finetuning of Vision-Language-Action Robot Policies via Parameter Merging
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
Autonomous Improvement of Instruction Following Skills via Foundation Models
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
Learning Long-Context Diffusion Policies via Past-Token Prediction
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
Re-Mix: Optimizing Data Mixtures for Large Scale Imitation Learning
von: Hejna, Joey, et al.
Veröffentlicht: (2024)
von: Hejna, Joey, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FAST: Efficient Action Tokenization for Vision-Language-Action Models
von: Pertsch, Karl, et al.
Veröffentlicht: (2025) -
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024) -
Training Strategies for Efficient Embodied Reasoning
von: Chen, William, et al.
Veröffentlicht: (2025) -
Scaling Cross-Embodied Learning: One Policy for Manipulation, Navigation, Locomotion and Aviation
von: Doshi, Ria, et al.
Veröffentlicht: (2024) -
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
von: Myers, Vivek, et al.
Veröffentlicht: (2024)