Scaffolding Dexterous Manipulation with Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | de Bakker, Vincent, Hejna, Joey, Lum, Tyler Ga Wei, Celik, Onur, Taranovic, Aleksandar, Blessing, Denis, Neumann, Gerhard, Bohg, Jeannette, Sadigh, Dorsa |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Acquiring Diverse Skills using Curriculum Reinforcement Learning with Mixture of Experts
by: Celik, Onur, et al.
Published: (2024)
by: Celik, Onur, et al.
Published: (2024)
SimToolReal: An Object-Centric Policy for Zero-Shot Dexterous Tool Manipulation
by: Kedia, Kushal, et al.
Published: (2026)
by: Kedia, Kushal, et al.
Published: (2026)
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning
by: Clark, Jaden, et al.
Published: (2025)
by: Clark, Jaden, et al.
Published: (2025)
Data Retrieval with Importance Weights for Few-Shot Imitation Learning
by: Xie, Amber, et al.
Published: (2025)
by: Xie, Amber, et al.
Published: (2025)
MotIF: Motion Instruction Fine-tuning
by: Hwang, Minyoung, et al.
Published: (2024)
by: Hwang, Minyoung, et al.
Published: (2024)
Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
Re-Mix: Optimizing Data Mixtures for Large Scale Imitation Learning
by: Hejna, Joey, et al.
Published: (2024)
by: Hejna, Joey, et al.
Published: (2024)
Neural Attention Field: Emerging Point Relevance in 3D Scenes for One-Shot Dexterous Grasping
by: Wang, Qianxu, et al.
Published: (2024)
by: Wang, Qianxu, et al.
Published: (2024)
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
by: Ren, Juntao, et al.
Published: (2025)
by: Ren, Juntao, et al.
Published: (2025)
What's the Move? Hybrid Imitation Learning via Salient Points
by: Sundaresan, Priya, et al.
Published: (2024)
by: Sundaresan, Priya, et al.
Published: (2024)
DexForce: Extracting Force-informed Actions from Kinesthetic Demonstrations for Dexterous Manipulation
by: Chen, Claire, et al.
Published: (2025)
by: Chen, Claire, et al.
Published: (2025)
So You Think You Can Scale Up Autonomous Robot Data Collection?
by: Mirchandani, Suvir, et al.
Published: (2024)
by: Mirchandani, Suvir, et al.
Published: (2024)
HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies
by: Xie, Amber, et al.
Published: (2026)
by: Xie, Amber, et al.
Published: (2026)
SpringGrasp: Synthesizing Compliant, Dexterous Grasps under Shape Uncertainty
by: Chen, Sirui, et al.
Published: (2024)
by: Chen, Sirui, et al.
Published: (2024)
Get a Grip: Multi-Finger Grasp Evaluation at Scale Enables Robust Sim-to-Real Transfer
by: Lum, Tyler Ga Wei, et al.
Published: (2024)
by: Lum, Tyler Ga Wei, et al.
Published: (2024)
Robot-Powered Data Flywheels: Deploying Robots in the Wild for Continual Data Collection and Foundation Model Adaptation
by: Grannen, Jennifer, et al.
Published: (2025)
by: Grannen, Jennifer, et al.
Published: (2025)
HoMeR: Learning In-the-Wild Mobile Manipulation via Hybrid Imitation and Whole-Body Control
by: Sundaresan, Priya, et al.
Published: (2025)
by: Sundaresan, Priya, et al.
Published: (2025)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
Towards Fusing Point Cloud and Visual Representations for Imitation Learning
by: Donat, Atalay, et al.
Published: (2025)
by: Donat, Atalay, et al.
Published: (2025)
DexDrummer: In-Hand, Contact-Rich, and Long-Horizon Dexterous Robot Drumming
by: Fang, Hung-Chieh, et al.
Published: (2026)
by: Fang, Hung-Chieh, et al.
Published: (2026)
Robot Data Curation with Mutual Information Estimators
by: Hejna, Joey, et al.
Published: (2025)
by: Hejna, Joey, et al.
Published: (2025)
DextrAH-G: Pixels-to-Action Dexterous Arm-Hand Grasping with Geometric Fabrics
by: Lum, Tyler Ga Wei, et al.
Published: (2024)
by: Lum, Tyler Ga Wei, et al.
Published: (2024)
Physically Grounded Vision-Language Models for Robotic Manipulation
by: Gao, Jensen, et al.
Published: (2023)
by: Gao, Jensen, et al.
Published: (2023)
Masquerade: Learning from In-the-wild Human Videos using Data-Editing
by: Lepert, Marion, et al.
Published: (2025)
by: Lepert, Marion, et al.
Published: (2025)
Phantom: Training Robots Without Robots Using Only Human Videos
by: Lepert, Marion, et al.
Published: (2025)
by: Lepert, Marion, et al.
Published: (2025)
Shadow: Leveraging Segmentation Masks for Cross-Embodiment Policy Transfer
by: Lepert, Marion, et al.
Published: (2025)
by: Lepert, Marion, et al.
Published: (2025)
COAST: Constraints and Streams for Task and Motion Planning
by: Vu, Brandon, et al.
Published: (2024)
by: Vu, Brandon, et al.
Published: (2024)
MaIL: Improving Imitation Learning with Mamba
by: Jia, Xiaogang, et al.
Published: (2024)
by: Jia, Xiaogang, et al.
Published: (2024)
SLAG: Scalable Language-Augmented Gaussian Splatting
by: Szilagyi, Laszlo, et al.
Published: (2025)
by: Szilagyi, Laszlo, et al.
Published: (2025)
EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models
by: Dong, Perry, et al.
Published: (2026)
by: Dong, Perry, et al.
Published: (2026)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
by: Gao, Jensen, et al.
Published: (2024)
by: Gao, Jensen, et al.
Published: (2024)
How to Train Your Robots? The Impact of Demonstration Modality on Imitation Learning
by: Li, Haozhuo, et al.
Published: (2025)
by: Li, Haozhuo, et al.
Published: (2025)
Vision Language Models are In-Context Value Learners
by: Ma, Yecheng Jason, et al.
Published: (2024)
by: Ma, Yecheng Jason, et al.
Published: (2024)
Vision in Action: Learning Active Perception from Human Demonstrations
by: Xiong, Haoyu, et al.
Published: (2025)
by: Xiong, Haoyu, et al.
Published: (2025)
RT-Sketch: Goal-Conditioned Imitation Learning from Hand-Drawn Sketches
by: Sundaresan, Priya, et al.
Published: (2024)
by: Sundaresan, Priya, et al.
Published: (2024)
EquivAct: SIM(3)-Equivariant Visuomotor Policies beyond Rigid Object Manipulation
by: Yang, Jingyun, et al.
Published: (2023)
by: Yang, Jingyun, et al.
Published: (2023)
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)
by: Bıyık, Erdem, et al.
Published: (2024)
Domain-Specific Fine-Tuning of Large Language Models for Interactive Robot Programming
by: Alt, Benjamin, et al.
Published: (2023)
by: Alt, Benjamin, et al.
Published: (2023)
Grounding Robot Generalization in Training Data via Retrieval-Augmented VLMs
by: Gao, Jensen, et al.
Published: (2026)
by: Gao, Jensen, et al.
Published: (2026)
Will People Enjoy a Robot Trainer? A Case Study with Snoopie the Pacerbot
by: Du, Maximilian, et al.
Published: (2026)
by: Du, Maximilian, et al.
Published: (2026)
Similar Items
-
Acquiring Diverse Skills using Curriculum Reinforcement Learning with Mixture of Experts
by: Celik, Onur, et al.
Published: (2024) -
SimToolReal: An Object-Centric Policy for Zero-Shot Dexterous Tool Manipulation
by: Kedia, Kushal, et al.
Published: (2026) -
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning
by: Clark, Jaden, et al.
Published: (2025) -
Data Retrieval with Importance Weights for Few-Shot Imitation Learning
by: Xie, Amber, et al.
Published: (2025) -
MotIF: Motion Instruction Fine-tuning
by: Hwang, Minyoung, et al.
Published: (2024)