Salvato in:
| Autori principali: | Li, Chengshu, Liang, Jacky, Zeng, Andy, Chen, Xinyun, Hausman, Karol, Sadigh, Dorsa, Levine, Sergey, Fei-Fei, Li, Xia, Fei, Ichter, Brian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2312.04474 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
di: Chen, Boyuan, et al.
Pubblicazione: (2024)
Physically Grounded Vision-Language Models for Robotic Manipulation
di: Gao, Jensen, et al.
Pubblicazione: (2023)
di: Gao, Jensen, et al.
Pubblicazione: (2023)
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
di: Nasiriany, Soroush, et al.
Pubblicazione: (2024)
di: Nasiriany, Soroush, et al.
Pubblicazione: (2024)
GenCHiP: Generating Robot Policy Code for High-Precision and Contact-Rich Manipulation Tasks
di: Burns, Kaylee, et al.
Pubblicazione: (2024)
di: Burns, Kaylee, et al.
Pubblicazione: (2024)
Generative Expressive Robot Behaviors using Large Language Models
di: Mahadevan, Karthik, et al.
Pubblicazione: (2024)
di: Mahadevan, Karthik, et al.
Pubblicazione: (2024)
Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections
di: Zha, Lihan, et al.
Pubblicazione: (2023)
di: Zha, Lihan, et al.
Pubblicazione: (2023)
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
di: Ahn, Michael, et al.
Pubblicazione: (2024)
di: Ahn, Michael, et al.
Pubblicazione: (2024)
Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization
di: Yang, Jonathan, et al.
Pubblicazione: (2025)
di: Yang, Jonathan, et al.
Pubblicazione: (2025)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Vocal Sandbox: Continual Learning and Adaptation for Situated Human-Robot Collaboration
di: Grannen, Jennifer, et al.
Pubblicazione: (2024)
di: Grannen, Jennifer, et al.
Pubblicazione: (2024)
ProVox: Personalization and Proactive Planning for Situated Human-Robot Collaboration
di: Grannen, Jennifer, et al.
Pubblicazione: (2025)
di: Grannen, Jennifer, et al.
Pubblicazione: (2025)
Grounding Robot Generalization in Training Data via Retrieval-Augmented VLMs
di: Gao, Jensen, et al.
Pubblicazione: (2026)
di: Gao, Jensen, et al.
Pubblicazione: (2026)
How to Train Your Robots? The Impact of Demonstration Modality on Imitation Learning
di: Li, Haozhuo, et al.
Pubblicazione: (2025)
di: Li, Haozhuo, et al.
Pubblicazione: (2025)
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning
di: Clark, Jaden, et al.
Pubblicazione: (2025)
di: Clark, Jaden, et al.
Pubblicazione: (2025)
Action-Free Reasoning for Policy Generalization
di: Clark, Jaden, et al.
Pubblicazione: (2025)
di: Clark, Jaden, et al.
Pubblicazione: (2025)
HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies
di: Xie, Amber, et al.
Pubblicazione: (2026)
di: Xie, Amber, et al.
Pubblicazione: (2026)
CoNVOI: Context-aware Navigation using Vision Language Models in Outdoor and Indoor Environments
di: Sathyamoorthy, Adarsh Jagan, et al.
Pubblicazione: (2024)
di: Sathyamoorthy, Adarsh Jagan, et al.
Pubblicazione: (2024)
Precise Robot Command Understanding Using Grammar-Constrained Large Language Models
di: Huo, Xinyun, et al.
Pubblicazione: (2026)
di: Huo, Xinyun, et al.
Pubblicazione: (2026)
RT-Sketch: Goal-Conditioned Imitation Learning from Hand-Drawn Sketches
di: Sundaresan, Priya, et al.
Pubblicazione: (2024)
di: Sundaresan, Priya, et al.
Pubblicazione: (2024)
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs
di: Hu, Zichao, et al.
Pubblicazione: (2024)
di: Hu, Zichao, et al.
Pubblicazione: (2024)
Batch Active Learning of Reward Functions from Human Preferences
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
Unified Video Action Model
di: Li, Shuang, et al.
Pubblicazione: (2025)
di: Li, Shuang, et al.
Pubblicazione: (2025)
SteerVLA: Steering Vision-Language-Action Models in Long-Tail Driving Scenarios
di: Gao, Tian, et al.
Pubblicazione: (2026)
di: Gao, Tian, et al.
Pubblicazione: (2026)
Joint Action Language Modelling for Transparent Policy Execution
di: Wulff, Theodor, et al.
Pubblicazione: (2025)
di: Wulff, Theodor, et al.
Pubblicazione: (2025)
Data Analogies Enable Efficient Cross-Embodiment Transfer
di: Yang, Jonathan, et al.
Pubblicazione: (2026)
di: Yang, Jonathan, et al.
Pubblicazione: (2026)
Toward Grounded Commonsense Reasoning
di: Kwon, Minae, et al.
Pubblicazione: (2023)
di: Kwon, Minae, et al.
Pubblicazione: (2023)
FAST: Efficient Action Tokenization for Vision-Language-Action Models
di: Pertsch, Karl, et al.
Pubblicazione: (2025)
di: Pertsch, Karl, et al.
Pubblicazione: (2025)
Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation
di: Yang, Jonathan, et al.
Pubblicazione: (2024)
di: Yang, Jonathan, et al.
Pubblicazione: (2024)
Will People Enjoy a Robot Trainer? A Case Study with Snoopie the Pacerbot
di: Du, Maximilian, et al.
Pubblicazione: (2026)
di: Du, Maximilian, et al.
Pubblicazione: (2026)
Invariance Co-training for Robot Visual Generalization
di: Yang, Jonathan, et al.
Pubblicazione: (2025)
di: Yang, Jonathan, et al.
Pubblicazione: (2025)
Language Guided Skill Discovery
di: Rho, Seungeun, et al.
Pubblicazione: (2024)
di: Rho, Seungeun, et al.
Pubblicazione: (2024)
Towards No-Code Programming of Cobots: Experiments with Code Synthesis by Large Code Models for Conversational Programming
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
Vision Language Models are In-Context Value Learners
di: Ma, Yecheng Jason, et al.
Pubblicazione: (2024)
di: Ma, Yecheng Jason, et al.
Pubblicazione: (2024)
Latent Diffusion Planning for Imitation Learning
di: Xie, Amber, et al.
Pubblicazione: (2025)
di: Xie, Amber, et al.
Pubblicazione: (2025)
Data Retrieval with Importance Weights for Few-Shot Imitation Learning
di: Xie, Amber, et al.
Pubblicazione: (2025)
di: Xie, Amber, et al.
Pubblicazione: (2025)
What Matters for Batch Online Reinforcement Learning in Robotics?
di: Dong, Perry, et al.
Pubblicazione: (2025)
di: Dong, Perry, et al.
Pubblicazione: (2025)
AgentThink: A Unified Framework for Tool-Augmented Chain-of-Thought Reasoning in Vision-Language Models for Autonomous Driving
di: Qian, Kangan, et al.
Pubblicazione: (2025)
di: Qian, Kangan, et al.
Pubblicazione: (2025)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
di: Tang, Yihe, et al.
Pubblicazione: (2025)
di: Tang, Yihe, et al.
Pubblicazione: (2025)
EXPO-FT: Sample-Efficient Reinforcement Learning Finetuning for Vision-Language-Action Models
di: Dong, Perry, et al.
Pubblicazione: (2026)
di: Dong, Perry, et al.
Pubblicazione: (2026)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
di: Qu, Kaixian, et al.
Pubblicazione: (2024)
di: Qu, Kaixian, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities
di: Chen, Boyuan, et al.
Pubblicazione: (2024) -
Physically Grounded Vision-Language Models for Robotic Manipulation
di: Gao, Jensen, et al.
Pubblicazione: (2023) -
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
di: Nasiriany, Soroush, et al.
Pubblicazione: (2024) -
GenCHiP: Generating Robot Policy Code for High-Precision and Contact-Rich Manipulation Tasks
di: Burns, Kaylee, et al.
Pubblicazione: (2024) -
Generative Expressive Robot Behaviors using Large Language Models
di: Mahadevan, Karthik, et al.
Pubblicazione: (2024)