Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Blank, Nils, Reuss, Moritz, Rühle, Marcel, Yağmurlu, Ömer Erdinç, Wenzel, Fabian, Mees, Oier, Lioutikov, Rudolf |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
by: Reuss, Moritz, et al.
Published: (2025)
by: Reuss, Moritz, et al.
Published: (2025)
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
by: Nakamoto, Mitsuhiko, et al.
Published: (2024)
by: Nakamoto, Mitsuhiko, et al.
Published: (2024)
Interpretable Affordance Detection on 3D Point Clouds with Probabilistic Prototypes
by: Li, Maximilian Xiling, et al.
Published: (2025)
by: Li, Maximilian Xiling, et al.
Published: (2025)
Towards Diverse Behaviors: A Benchmark for Imitation Learning with Human Demonstrations
by: Jia, Xiaogang, et al.
Published: (2024)
by: Jia, Xiaogang, et al.
Published: (2024)
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
by: Myers, Vivek, et al.
Published: (2024)
by: Myers, Vivek, et al.
Published: (2024)
Scaling Cross-Embodied Learning: One Policy for Manipulation, Navigation, Locomotion and Aviation
by: Doshi, Ria, et al.
Published: (2024)
by: Doshi, Ria, et al.
Published: (2024)
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
by: Jones, Joshua, et al.
Published: (2025)
by: Jones, Joshua, et al.
Published: (2025)
NaviTrace: Evaluating Embodied Navigation of Vision-Language Models
by: Windecker, Tim, et al.
Published: (2025)
by: Windecker, Tim, et al.
Published: (2025)
Autonomous Improvement of Instruction Following Skills via Foundation Models
by: Zhou, Zhiyuan, et al.
Published: (2024)
by: Zhou, Zhiyuan, et al.
Published: (2024)
LeLaN: Learning A Language-Conditioned Navigation Policy from In-the-Wild Videos
by: Hirose, Noriaki, et al.
Published: (2024)
by: Hirose, Noriaki, et al.
Published: (2024)
Robotic Control via Embodied Chain-of-Thought Reasoning
by: Zawalski, Michał, et al.
Published: (2024)
by: Zawalski, Michał, et al.
Published: (2024)
Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)
by: Li, Ge, et al.
Published: (2024)
Multimodal Spatial Language Maps for Robot Navigation and Manipulation
by: Huang, Chenguang, et al.
Published: (2025)
by: Huang, Chenguang, et al.
Published: (2025)
TOP-ERL: Transformer-based Off-Policy Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)
by: Li, Ge, et al.
Published: (2024)
PointMapPolicy: Structured Point Cloud Processing for Multi-Modal Imitation Learning
by: Jia, Xiaogang, et al.
Published: (2025)
by: Jia, Xiaogang, et al.
Published: (2025)
The Ingredients for Robotic Diffusion Transformers
by: Dasari, Sudeep, et al.
Published: (2024)
by: Dasari, Sudeep, et al.
Published: (2024)
Beyond Visuals: Investigating Force Feedback in Extended Reality for Robot Data Collection
by: Li, Xueyin, et al.
Published: (2025)
by: Li, Xueyin, et al.
Published: (2025)
Use the Force, Bot! -- Force-Aware ProDMP with Event-Based Replanning
by: Lödige, Paul Werner, et al.
Published: (2024)
by: Lödige, Paul Werner, et al.
Published: (2024)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
by: Pai, Jonas, et al.
Published: (2025)
by: Pai, Jonas, et al.
Published: (2025)
BMP: Bridging the Gap between B-Spline and Movement Primitives
by: Liao, Weiran, et al.
Published: (2024)
by: Liao, Weiran, et al.
Published: (2024)
Perception Stitching: Zero-Shot Perception Encoder Transfer for Visuomotor Robot Policies
by: Jian, Pingcheng, et al.
Published: (2024)
by: Jian, Pingcheng, et al.
Published: (2024)
Movement Primitive Diffusion: Learning Gentle Robotic Manipulation of Deformable Objects
by: Scheikl, Paul Maria, et al.
Published: (2023)
by: Scheikl, Paul Maria, et al.
Published: (2023)
A Robotic Skill Learning System Built Upon Diffusion Policies and Foundation Models
by: Ingelhag, Nils, et al.
Published: (2024)
by: Ingelhag, Nils, et al.
Published: (2024)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
Training Strategies for Efficient Embodied Reasoning
by: Chen, William, et al.
Published: (2025)
by: Chen, William, et al.
Published: (2025)
Robot Utility Models: General Policies for Zero-Shot Deployment in New Environments
by: Etukuru, Haritheja, et al.
Published: (2024)
by: Etukuru, Haritheja, et al.
Published: (2024)
Octo: An Open-Source Generalist Robot Policy
by: Octo Model Team, et al.
Published: (2024)
by: Octo Model Team, et al.
Published: (2024)
X-IL: Exploring the Design Space of Imitation Learning Policies
by: Jia, Xiaogang, et al.
Published: (2025)
by: Jia, Xiaogang, et al.
Published: (2025)
Comp-LTL: Temporal Logic Planning via Zero-Shot Policy Composition
by: Bergeron, Taylor, et al.
Published: (2024)
by: Bergeron, Taylor, et al.
Published: (2024)
Multi-Floor Zero-Shot Object Navigation Policy
by: Zhang, Lingfeng, et al.
Published: (2024)
by: Zhang, Lingfeng, et al.
Published: (2024)
ZeroCAP: Zero-Shot Multi-Robot Context Aware Pattern Formation via Large Language Models
by: Venkatesh, Vishnunandan L. N., et al.
Published: (2024)
by: Venkatesh, Vishnunandan L. N., et al.
Published: (2024)
MoRe-ERL: Learning Motion Residuals using Episodic Reinforcement Learning
by: Huang, Xi, et al.
Published: (2025)
by: Huang, Xi, et al.
Published: (2025)
Evaluating Real-World Robot Manipulation Policies in Simulation
by: Li, Xuanlin, et al.
Published: (2024)
by: Li, Xuanlin, et al.
Published: (2024)
An Overview of Prototype Formulations for Interpretable Deep Learning
by: Li, Maximilian Xiling, et al.
Published: (2024)
by: Li, Maximilian Xiling, et al.
Published: (2024)
Robotic Scene Cloning:Advancing Zero-Shot Robotic Scene Adaptation in Manipulation via Visual Prompt Editing
by: Huang, Binyuan, et al.
Published: (2026)
by: Huang, Binyuan, et al.
Published: (2026)
Spotting the Unfriendly Robot -- Towards better Metrics for Interactions
by: Wenzel, Raphael, et al.
Published: (2025)
by: Wenzel, Raphael, et al.
Published: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
Similar Items
-
Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
by: Reuss, Moritz, et al.
Published: (2024) -
FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
by: Reuss, Moritz, et al.
Published: (2025) -
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024) -
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
by: Zhou, Hongyi, et al.
Published: (2025) -
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
by: Nakamoto, Mitsuhiko, et al.
Published: (2024)