Enhancing Speech Instruction Understanding and Disambiguation in Robotics via Speech Prosody
Fuente:
arXiv
Saved in:
| Main Authors: | Sasu, David, Yamoah, Kweku Andoh, Quartey, Benedict, Schluter, Natalie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploiting Contextual Structure to Generate Useful Auxiliary Tasks
by: Quartey, Benedict, et al.
Published: (2023)
by: Quartey, Benedict, et al.
Published: (2023)
LLMs for Robotic Object Disambiguation
by: Jiang, Connie, et al.
Published: (2024)
by: Jiang, Connie, et al.
Published: (2024)
Verifiably Following Complex Robot Instructions with Foundation Models
by: Quartey, Benedict, et al.
Published: (2024)
by: Quartey, Benedict, et al.
Published: (2024)
ConformalNL2LTL: Translating Natural Language Instructions into Temporal Logic Formulas with Conformal Correctness Guarantees
by: Sundarsingh, David Smith, et al.
Published: (2025)
by: Sundarsingh, David Smith, et al.
Published: (2025)
λ: A Benchmark for Data-Efficiency in Long-Horizon Indoor Mobile Manipulation Robotics
by: Jaafar, Ahmed, et al.
Published: (2024)
by: Jaafar, Ahmed, et al.
Published: (2024)
Diffusion Language Models for Speech Recognition
by: Naveriani, Davyd, et al.
Published: (2026)
by: Naveriani, Davyd, et al.
Published: (2026)
Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following
by: Yang, Brian, et al.
Published: (2024)
by: Yang, Brian, et al.
Published: (2024)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
by: Sasu, David, et al.
Published: (2025)
by: Sasu, David, et al.
Published: (2025)
Diagnosing Robotics Systems Issues with Large Language Models
by: Herrmann, Jordis Emilia, et al.
Published: (2024)
by: Herrmann, Jordis Emilia, et al.
Published: (2024)
Can Large Language Models Solve Robot Routing?
by: Huang, Zhehui, et al.
Published: (2024)
by: Huang, Zhehui, et al.
Published: (2024)
Understanding Social Perception, Interactions, and Safety Aspects of Sidewalk Delivery Robots Using Sentiment Analysis
by: Du, Yuchen, et al.
Published: (2024)
by: Du, Yuchen, et al.
Published: (2024)
TalkWithMachines: Enhancing Human-Robot Interaction for Interpretable Industrial Robotics Through Large/Vision Language Models
by: Abbas, Ammar N., et al.
Published: (2024)
by: Abbas, Ammar N., et al.
Published: (2024)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
Prompt Selection and Augmentation for Few Examples Code Generation in Large Language Model and its Application in Robotics Control
by: Wu, On Tai, et al.
Published: (2024)
by: Wu, On Tai, et al.
Published: (2024)
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
by: Tavella, Federico, et al.
Published: (2025)
by: Tavella, Federico, et al.
Published: (2025)
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods
by: Cao, Yuji, et al.
Published: (2024)
by: Cao, Yuji, et al.
Published: (2024)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
by: Li, Haozhan, et al.
Published: (2025)
by: Li, Haozhan, et al.
Published: (2025)
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
by: Chow, Wei, et al.
Published: (2025)
by: Chow, Wei, et al.
Published: (2025)
WorldSpeech: A Multilingual Speech Corpus from Around the World
by: Asonitis, Antonis, et al.
Published: (2026)
by: Asonitis, Antonis, et al.
Published: (2026)
DiffuSpeech: Silent Thought, Spoken Answer via Unified Speech-Text Diffusion
by: Lou, Yuxuan, et al.
Published: (2026)
by: Lou, Yuxuan, et al.
Published: (2026)
Active Task Disambiguation with LLMs
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
by: Kobalczyk, Katarzyna, et al.
Published: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
by: Almohaimeed, Saad, et al.
Published: (2025)
by: Almohaimeed, Saad, et al.
Published: (2025)
Advancing Speech Understanding in Speech-Aware Language Models with GRPO
by: Elmakies, Avishai, et al.
Published: (2025)
by: Elmakies, Avishai, et al.
Published: (2025)
IROSA: Interactive Robot Skill Adaptation using Natural Language
by: Knauer, Markus, et al.
Published: (2026)
by: Knauer, Markus, et al.
Published: (2026)
ProVox: Personalization and Proactive Planning for Situated Human-Robot Collaboration
by: Grannen, Jennifer, et al.
Published: (2025)
by: Grannen, Jennifer, et al.
Published: (2025)
Vocal Sandbox: Continual Learning and Adaptation for Situated Human-Robot Collaboration
by: Grannen, Jennifer, et al.
Published: (2024)
by: Grannen, Jennifer, et al.
Published: (2024)
Generating Traffic Scenarios via In-Context Learning to Learn Better Motion Planner
by: Aiersilan, Aizierjiang
Published: (2024)
by: Aiersilan, Aizierjiang
Published: (2024)
Improving Direct Persian-English Speech-to-Speech Translation with Discrete Units and Synthetic Parallel Data
by: Rashidi, Sina, et al.
Published: (2025)
by: Rashidi, Sina, et al.
Published: (2025)
DRAGON: A Dialogue-Based Robot for Assistive Navigation with Visual Language Grounding
by: Liu, Shuijing, et al.
Published: (2023)
by: Liu, Shuijing, et al.
Published: (2023)
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
by: Huang, Yu Ti
Published: (2025)
by: Huang, Yu Ti
Published: (2025)
General Agentic Planning Through Simulative Reasoning with World Models
by: Deng, Mingkai, et al.
Published: (2025)
by: Deng, Mingkai, et al.
Published: (2025)
Flexible Agent Alignment with Goal Inference from Open-Ended Dialog
by: Ma, Rachel, et al.
Published: (2025)
by: Ma, Rachel, et al.
Published: (2025)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
by: Ivanova, Anastasiia, et al.
Published: (2025)
by: Ivanova, Anastasiia, et al.
Published: (2025)
AdaptBot: Combining LLM with Knowledge Graphs and Human Input for Generic-to-Specific Task Decomposition and Knowledge Refinement
by: Singh, Shivam, et al.
Published: (2025)
by: Singh, Shivam, et al.
Published: (2025)
Mental Modeling of Reinforcement Learning Agents by Language Models
by: Lu, Wenhao, et al.
Published: (2024)
by: Lu, Wenhao, et al.
Published: (2024)
Enabling robots to follow abstract instructions and complete complex dynamic tasks
by: Mon-Williams, Ruaridh, et al.
Published: (2024)
by: Mon-Williams, Ruaridh, et al.
Published: (2024)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
by: Zheng, Qinqing, et al.
Published: (2024)
by: Zheng, Qinqing, et al.
Published: (2024)
The RL/LLM Taxonomy Tree: Reviewing Synergies Between Reinforcement Learning and Large Language Models
by: Pternea, Moschoula, et al.
Published: (2024)
by: Pternea, Moschoula, et al.
Published: (2024)
The Essential Role of Causality in Foundation World Models for Embodied AI
by: Gupta, Tarun, et al.
Published: (2024)
by: Gupta, Tarun, et al.
Published: (2024)
MaxMin-RLHF: Alignment with Diverse Human Preferences
by: Chakraborty, Souradip, et al.
Published: (2024)
by: Chakraborty, Souradip, et al.
Published: (2024)
Similar Items
-
Exploiting Contextual Structure to Generate Useful Auxiliary Tasks
by: Quartey, Benedict, et al.
Published: (2023) -
LLMs for Robotic Object Disambiguation
by: Jiang, Connie, et al.
Published: (2024) -
Verifiably Following Complex Robot Instructions with Foundation Models
by: Quartey, Benedict, et al.
Published: (2024) -
ConformalNL2LTL: Translating Natural Language Instructions into Temporal Logic Formulas with Conformal Correctness Guarantees
by: Sundarsingh, David Smith, et al.
Published: (2025) -
λ: A Benchmark for Data-Efficiency in Long-Horizon Indoor Mobile Manipulation Robotics
by: Jaafar, Ahmed, et al.
Published: (2024)