Joint Action Language Modelling for Transparent Policy Execution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wulff, Theodor, Maharjan, Rahul Singh, Chi, Xinyun, Cangelosi, Angelo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment
von: Wulff, Theodor, et al.
Veröffentlicht: (2026)
von: Wulff, Theodor, et al.
Veröffentlicht: (2026)
Noise-Free Explanation for Driving Action Prediction
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024)
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024)
Signs of Language: Embodied Sign Language Fingerspelling Acquisition from Demonstrations for Human-Robot Interaction
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
von: Tavella, Federico, et al.
Veröffentlicht: (2022)
From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
Precise Robot Command Understanding Using Grammar-Constrained Large Language Models
von: Huo, Xinyun, et al.
Veröffentlicht: (2026)
von: Huo, Xinyun, et al.
Veröffentlicht: (2026)
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
von: Tavella, Federico, et al.
Veröffentlicht: (2025)
von: Tavella, Federico, et al.
Veröffentlicht: (2025)
DeGuV: Depth-Guided Visual Reinforcement Learning for Generalization and Interpretability in Manipulation
von: Pham, Tien, et al.
Veröffentlicht: (2025)
von: Pham, Tien, et al.
Veröffentlicht: (2025)
Attributes-aware Visual Emotion Representation Learning
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
von: Maharjan, Rahul Singh, et al.
Veröffentlicht: (2025)
Stable Language Guidance for Vision-Language-Action Models
von: Zhan, Zhihao, et al.
Veröffentlicht: (2026)
von: Zhan, Zhihao, et al.
Veröffentlicht: (2026)
SRPO: Self-Referential Policy Optimization for Vision-Language-Action Models
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
CASPER: Cognitive Architecture for Social Perception and Engagement in Robots
von: Vinanzi, Samuele, et al.
Veröffentlicht: (2022)
von: Vinanzi, Samuele, et al.
Veröffentlicht: (2022)
EdgeVLA: Efficient Vision-Language-Action Models
von: Budzianowski, Paweł, et al.
Veröffentlicht: (2025)
von: Budzianowski, Paweł, et al.
Veröffentlicht: (2025)
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
von: Zhu, Wang, et al.
Veröffentlicht: (2024)
von: Zhu, Wang, et al.
Veröffentlicht: (2024)
Red-Teaming Vision-Language-Action Models via Quality Diversity Prompt Generation for Robust Robot Policies
von: Srikanth, Siddharth, et al.
Veröffentlicht: (2026)
von: Srikanth, Siddharth, et al.
Veröffentlicht: (2026)
Bridging the Communication Gap: Artificial Agents Learning Sign Language through Imitation
von: Tavella, Federico, et al.
Veröffentlicht: (2024)
von: Tavella, Federico, et al.
Veröffentlicht: (2024)
ASMR: Augmenting Life Scenario using Large Generative Models for Robotic Action Reflection
von: Tsai, Shang-Chi, et al.
Veröffentlicht: (2025)
von: Tsai, Shang-Chi, et al.
Veröffentlicht: (2025)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
Driving Everywhere with Large Language Model Policy Adaptation
von: Li, Boyi, et al.
Veröffentlicht: (2024)
von: Li, Boyi, et al.
Veröffentlicht: (2024)
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
Phonology Recognition in American Sign Language
von: Tavella, Federico, et al.
Veröffentlicht: (2021)
von: Tavella, Federico, et al.
Veröffentlicht: (2021)
CARTIER: Cartographic lAnguage Reasoning Targeted at Instruction Execution for Robots
von: Rivkin, Dmitriy, et al.
Veröffentlicht: (2023)
von: Rivkin, Dmitriy, et al.
Veröffentlicht: (2023)
LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
Policy-Guided World Model Planning for Language-Conditioned Visual Navigation
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2026)
von: Chahe, Amirhosein, et al.
Veröffentlicht: (2026)
Chain of Code: Reasoning with a Language Model-Augmented Code Emulator
von: Li, Chengshu, et al.
Veröffentlicht: (2023)
von: Li, Chengshu, et al.
Veröffentlicht: (2023)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
Anticipate & Act : Integrating LLMs and Classical Planning for Efficient Task Execution in Household Environments
von: Arora, Raghav, et al.
Veröffentlicht: (2025)
von: Arora, Raghav, et al.
Veröffentlicht: (2025)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
From Forecasting to Planning: Policy World Model for Collaborative State-Action Prediction
von: Zhao, Zhida, et al.
Veröffentlicht: (2025)
von: Zhao, Zhida, et al.
Veröffentlicht: (2025)
In-Context Learning Enables Robot Action Prediction in LLMs
von: Yin, Yida, et al.
Veröffentlicht: (2024)
von: Yin, Yida, et al.
Veröffentlicht: (2024)
ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models
von: Darabi, Nastaran, et al.
Veröffentlicht: (2026)
von: Darabi, Nastaran, et al.
Veröffentlicht: (2026)
AlphaSpace: Enabling Robotic Actions through Semantic Tokenization and Symbolic Reasoning
von: Dao, Alan, et al.
Veröffentlicht: (2025)
von: Dao, Alan, et al.
Veröffentlicht: (2025)
Evaluation of Habitat Robotics using Large Language Models
von: Li, William, et al.
Veröffentlicht: (2025)
von: Li, William, et al.
Veröffentlicht: (2025)
Statler: State-Maintaining Language Models for Embodied Reasoning
von: Yoneda, Takuma, et al.
Veröffentlicht: (2023)
von: Yoneda, Takuma, et al.
Veröffentlicht: (2023)
Instruct Large Language Models to Drive like Humans
von: Zhang, Ruijun, et al.
Veröffentlicht: (2024)
von: Zhang, Ruijun, et al.
Veröffentlicht: (2024)
PoseLess: Depth-Free Vision-to-Joint Control via Direct Image Mapping with VLM
von: Dao, Alan, et al.
Veröffentlicht: (2025)
von: Dao, Alan, et al.
Veröffentlicht: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
World Action Models: The Next Frontier in Embodied AI
von: Wang, Siyin, et al.
Veröffentlicht: (2026)
von: Wang, Siyin, et al.
Veröffentlicht: (2026)
Benchmarking Local Language Models for Social Robots using Edge Devices
von: Lamouille, Dorian, et al.
Veröffentlicht: (2026)
von: Lamouille, Dorian, et al.
Veröffentlicht: (2026)
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
von: Liu, Yibin, et al.
Veröffentlicht: (2025)
von: Liu, Yibin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment
von: Wulff, Theodor, et al.
Veröffentlicht: (2026) -
Noise-Free Explanation for Driving Action Prediction
von: Zhu, Hongbo, et al.
Veröffentlicht: (2024) -
Signs of Language: Embodied Sign Language Fingerspelling Acquisition from Demonstrations for Human-Robot Interaction
von: Tavella, Federico, et al.
Veröffentlicht: (2022) -
From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning
von: Xie, Haodong, et al.
Veröffentlicht: (2024) -
Precise Robot Command Understanding Using Grammar-Constrained Large Language Models
von: Huo, Xinyun, et al.
Veröffentlicht: (2026)