ClevrSkills: Compositional Language and Visual Reasoning in Robotics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Haresh, Sanjay, Dijkman, Daniel, Bhattacharyya, Apratim, Memisevic, Roland |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Notes-to-Self: Scratchpad Augmented VLAs for Memory Dependent Manipulation Tasks
von: Haresh, Sanjay, et al.
Veröffentlicht: (2026)
von: Haresh, Sanjay, et al.
Veröffentlicht: (2026)
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
von: Bendikas, Rokas, et al.
Veröffentlicht: (2025)
von: Bendikas, Rokas, et al.
Veröffentlicht: (2025)
Delayed Attention Training Improves Length Generalization in Transformer--RNN Hybrids
von: Phan, Buu, et al.
Veröffentlicht: (2025)
von: Phan, Buu, et al.
Veröffentlicht: (2025)
Aligning Robot Navigation Behaviors with Human Intentions and Preferences
von: Karnan, Haresh
Veröffentlicht: (2024)
von: Karnan, Haresh
Veröffentlicht: (2024)
Look, Remember and Reason: Grounded reasoning in videos with language models
von: Bhattacharyya, Apratim, et al.
Veröffentlicht: (2023)
von: Bhattacharyya, Apratim, et al.
Veröffentlicht: (2023)
From Code to Action: Hierarchical Learning of Diffusion-VLM Policies
von: Peschl, Markus, et al.
Veröffentlicht: (2025)
von: Peschl, Markus, et al.
Veröffentlicht: (2025)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
Hybrid Training for Vision-Language-Action Models
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2025)
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2025)
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
Uni-Skill: Building Self-Evolving Skill Repository for Generalizable Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2026)
von: Xie, Senwei, et al.
Veröffentlicht: (2026)
RoboCoder: Robotic Learning from Basic Skills to General Tasks with Large Language Models
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
Self-Refining Vision Language Model for Robotic Failure Detection and Reasoning
von: Qi, Carl, et al.
Veröffentlicht: (2026)
von: Qi, Carl, et al.
Veröffentlicht: (2026)
MuTT: A Multimodal Trajectory Transformer for Robot Skills
von: Kienle, Claudius, et al.
Veröffentlicht: (2024)
von: Kienle, Claudius, et al.
Veröffentlicht: (2024)
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
von: Kim, Seungsu, et al.
Veröffentlicht: (2026)
von: Kim, Seungsu, et al.
Veröffentlicht: (2026)
Placeit! A Framework for Learning Robot Object Placement Skills
von: Ferrad, Amina, et al.
Veröffentlicht: (2025)
von: Ferrad, Amina, et al.
Veröffentlicht: (2025)
Generalize by Touching: Tactile Ensemble Skill Transfer for Robotic Furniture Assembly
von: Lin, Haohong, et al.
Veröffentlicht: (2024)
von: Lin, Haohong, et al.
Veröffentlicht: (2024)
OSMO: Open-Source Tactile Glove for Human-to-Robot Skill Transfer
von: Yin, Jessica, et al.
Veröffentlicht: (2025)
von: Yin, Jessica, et al.
Veröffentlicht: (2025)
Skill Expansion and Composition in Parameter Space
von: Liu, Tenglong, et al.
Veröffentlicht: (2025)
von: Liu, Tenglong, et al.
Veröffentlicht: (2025)
SPECI: Skill Prompts based Hierarchical Continual Imitation Learning for Robot Manipulation
von: Xu, Jingkai, et al.
Veröffentlicht: (2025)
von: Xu, Jingkai, et al.
Veröffentlicht: (2025)
Iterative Compositional Data Generation for Robot Control
von: Pham, Anh-Quan, et al.
Veröffentlicht: (2025)
von: Pham, Anh-Quan, et al.
Veröffentlicht: (2025)
STAR: Learning Diverse Robot Skill Abstractions through Rotation-Augmented Vector Quantization
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Learning Safe Autonomous Driving Policies Using Predictive Safety Representations
von: Keswani, Mahesh, et al.
Veröffentlicht: (2025)
von: Keswani, Mahesh, et al.
Veröffentlicht: (2025)
Robotic Control via Embodied Chain-of-Thought Reasoning
von: Zawalski, Michał, et al.
Veröffentlicht: (2024)
von: Zawalski, Michał, et al.
Veröffentlicht: (2024)
Verification of Visual Controllers via Compositional Geometric Transformations
von: Estornell, Alexander, et al.
Veröffentlicht: (2025)
von: Estornell, Alexander, et al.
Veröffentlicht: (2025)
PoCo: Policy Composition from and for Heterogeneous Robot Learning
von: Wang, Lirui, et al.
Veröffentlicht: (2024)
von: Wang, Lirui, et al.
Veröffentlicht: (2024)
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
von: Haarnoja, Tuomas, et al.
Veröffentlicht: (2023)
Learning Human-Like Badminton Skills for Humanoid Robots
von: Chen, Yeke, et al.
Veröffentlicht: (2026)
von: Chen, Yeke, et al.
Veröffentlicht: (2026)
Learning to Transfer Human Hand Skills for Robot Manipulations
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
von: Park, Sungjae, et al.
Veröffentlicht: (2025)
Skills Made to Order: Efficient Acquisition of Robot Cooking Skills Guided by Multiple Forms of Internet Data
von: Verghese, Mrinal, et al.
Veröffentlicht: (2024)
von: Verghese, Mrinal, et al.
Veröffentlicht: (2024)
A Compositional Paradigm for Foundation Models: Towards Smarter Robotic Agents
von: Quarantiello, Luigi, et al.
Veröffentlicht: (2025)
von: Quarantiello, Luigi, et al.
Veröffentlicht: (2025)
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
von: Ryu, Kanghyun, et al.
Veröffentlicht: (2024)
Flow-based Domain Randomization for Learning and Sequencing Robotic Skills
von: Curtis, Aidan, et al.
Veröffentlicht: (2025)
von: Curtis, Aidan, et al.
Veröffentlicht: (2025)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
ViReSkill: Vision-Grounded Replanning with Skill Memory for LLM-Based Planning in Lifelong Robot Learning
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2025)
von: Kagaya, Tomoyuki, et al.
Veröffentlicht: (2025)
Robotic Skill Diversification via Active Mutation of Reward Functions in Reinforcement Learning During a Liquid Pouring Task
von: van Buuren, Jannick, et al.
Veröffentlicht: (2025)
von: van Buuren, Jannick, et al.
Veröffentlicht: (2025)
Interpretable Robotic Manipulation from Language
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025)
von: Tang, Nan, et al.
Veröffentlicht: (2025)
LGR2: Language Guided Reward Relabeling for Accelerating Hierarchical Reinforcement Learning
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Notes-to-Self: Scratchpad Augmented VLAs for Memory Dependent Manipulation Tasks
von: Haresh, Sanjay, et al.
Veröffentlicht: (2026) -
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
von: Bendikas, Rokas, et al.
Veröffentlicht: (2025) -
Delayed Attention Training Improves Length Generalization in Transformer--RNN Hybrids
von: Phan, Buu, et al.
Veröffentlicht: (2025) -
Aligning Robot Navigation Behaviors with Human Intentions and Preferences
von: Karnan, Haresh
Veröffentlicht: (2024) -
Look, Remember and Reason: Grounded reasoning in videos with language models
von: Bhattacharyya, Apratim, et al.
Veröffentlicht: (2023)