Visual Cues Enhance Predictive Turn-Taking for Two-Party Human Interaction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Russell, Sam O'Connor, Harte, Naomi |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Visual Cues Support Robust Turn-taking Prediction in Noise
par: Russell, Sam O'Connor, et autres
Publié: (2025)
par: Russell, Sam O'Connor, et autres
Publié: (2025)
The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations
par: Russell, Sam OConnor, et autres
Publié: (2026)
par: Russell, Sam OConnor, et autres
Publié: (2026)
Applying General Turn-taking Models to Conversational Human-Robot Interaction
par: Skantze, Gabriel, et autres
Publié: (2025)
par: Skantze, Gabriel, et autres
Publié: (2025)
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
par: Inoue, Koji, et autres
Publié: (2025)
par: Inoue, Koji, et autres
Publié: (2025)
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
par: Zhu, Wang Bill, et autres
Publié: (2025)
par: Zhu, Wang Bill, et autres
Publié: (2025)
Utilization of Non-verbal Behaviour and Social Gaze in Classroom Human-Robot Interaction Communications
par: Shaghaghi, Sahand, et autres
Publié: (2023)
par: Shaghaghi, Sahand, et autres
Publié: (2023)
Leveraging Large Language Models in Human-Robot Interaction: A Critical Analysis of Potential and Pitfalls
par: Atuhurra, Jesse
Publié: (2024)
par: Atuhurra, Jesse
Publié: (2024)
Predicting Turn-Taking and Backchannel in Human-Machine Conversations Using Linguistic, Acoustic, and Visual Signals
par: Lin, Yuxin, et autres
Publié: (2025)
par: Lin, Yuxin, et autres
Publié: (2025)
NERsocial: Efficient Named Entity Recognition Dataset Construction for Human-Robot Interaction Utilizing RapidNER
par: Atuhurra, Jesse, et autres
Publié: (2024)
par: Atuhurra, Jesse, et autres
Publié: (2024)
Speak or Stay Silent: Context-Aware Turn-Taking in Multi-Party Dialogue
par: Bhagtani, Kratika, et autres
Publié: (2026)
par: Bhagtani, Kratika, et autres
Publié: (2026)
The Design of Informative Take-Over Requests for Semi-Autonomous Cyber-Physical Systems: Combining Spoken Language and Visual Icons in a Drone-Controller Setting
par: Gundappa, Ashwini, et autres
Publié: (2024)
par: Gundappa, Ashwini, et autres
Publié: (2024)
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
par: Kranti, Chalamalasetti, et autres
Publié: (2026)
par: Kranti, Chalamalasetti, et autres
Publié: (2026)
Child Speech Recognition in Human-Robot Interaction: Problem Solved?
par: Janssens, Ruben, et autres
Publié: (2024)
par: Janssens, Ruben, et autres
Publié: (2024)
A Framework for Adapting Human-Robot Interaction to Diverse User Groups
par: Rosin, Theresa Pekarek, et autres
Publié: (2024)
par: Rosin, Theresa Pekarek, et autres
Publié: (2024)
Large Language Models as Zero-Shot Human Models for Human-Robot Interaction
par: Zhang, Bowen, et autres
Publié: (2023)
par: Zhang, Bowen, et autres
Publié: (2023)
Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
par: Sasabuchi, Kazuhiro, et autres
Publié: (2025)
par: Sasabuchi, Kazuhiro, et autres
Publié: (2025)
Prompt-Guided Turn-Taking Prediction
par: Inoue, Koji, et autres
Publié: (2025)
par: Inoue, Koji, et autres
Publié: (2025)
Enhancing Human-Robot Interaction in Healthcare: A Study on Nonverbal Communication Cues and Trust Dynamics with NAO Robot Caregivers
par: Raju, S M Taslim Uddin
Publié: (2025)
par: Raju, S M Taslim Uddin
Publié: (2025)
OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
par: Budzianowski, Paweł, et autres
Publié: (2025)
par: Budzianowski, Paweł, et autres
Publié: (2025)
PROGrasp: Pragmatic Human-Robot Communication for Object Grasping
par: Kang, Gi-Cheon, et autres
Publié: (2023)
par: Kang, Gi-Cheon, et autres
Publié: (2023)
Instruct Large Language Models to Drive like Humans
par: Zhang, Ruijun, et autres
Publié: (2024)
par: Zhang, Ruijun, et autres
Publié: (2024)
In-Context Learning Enables Robot Action Prediction in LLMs
par: Yin, Yida, et autres
Publié: (2024)
par: Yin, Yida, et autres
Publié: (2024)
From Human Intention to Action Prediction: Intention-Driven End-to-End Autonomous Driving
par: Zheng, Huan, et autres
Publié: (2025)
par: Zheng, Huan, et autres
Publié: (2025)
TalkWithMachines: Enhancing Human-Robot Interaction for Interpretable Industrial Robotics Through Large/Vision Language Models
par: Abbas, Ammar N., et autres
Publié: (2024)
par: Abbas, Ammar N., et autres
Publié: (2024)
A Monte Carlo Language Model Pipeline for Zero-Shot Sociopolitical Event Extraction
par: Cai, Erica, et autres
Publié: (2023)
par: Cai, Erica, et autres
Publié: (2023)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
par: Wang, Tianyu, et autres
Publié: (2024)
par: Wang, Tianyu, et autres
Publié: (2024)
Towards Predicting Any Human Trajectory In Context
par: Fujii, Ryo, et autres
Publié: (2025)
par: Fujii, Ryo, et autres
Publié: (2025)
Can LLMs Translate Human Instructions into a Reinforcement Learning Agent's Internal Emergent Symbolic Representation?
par: Ma, Ziqi, et autres
Publié: (2025)
par: Ma, Ziqi, et autres
Publié: (2025)
MINT-RVAE: Multi-Cues Intention Prediction of Human-Robot Interaction using Human Pose and Emotion Information from RGB-only Camera Data
par: Mohsen, Farida, et autres
Publié: (2025)
par: Mohsen, Farida, et autres
Publié: (2025)
Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues
par: Bayrak, Ahmet Tuğrul, et autres
Publié: (2026)
par: Bayrak, Ahmet Tuğrul, et autres
Publié: (2026)
How Can Large Language Models Enable Better Socially Assistive Human-Robot Interaction: A Brief Survey
par: Shi, Zhonghao, et autres
Publié: (2024)
par: Shi, Zhonghao, et autres
Publié: (2024)
When2Speak: A Dataset for Temporal Participation and Turn-Taking in Multi-Party Conversations for Large Language Models
par: Nama, Vihaan, et autres
Publié: (2026)
par: Nama, Vihaan, et autres
Publié: (2026)
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
par: Wang, Yunfei, et autres
Publié: (2026)
par: Wang, Yunfei, et autres
Publié: (2026)
FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback
par: Qian, Kangan, et autres
Publié: (2025)
par: Qian, Kangan, et autres
Publié: (2025)
Coordinates from Context: Using LLMs to Ground Complex Location References
par: Masis, Tessa, et autres
Publié: (2025)
par: Masis, Tessa, et autres
Publié: (2025)
Where on Earth Do Users Say They Are?: Geo-Entity Linking for Noisy Multilingual User Input
par: Masis, Tessa, et autres
Publié: (2024)
par: Masis, Tessa, et autres
Publié: (2024)
Signs of Language: Embodied Sign Language Fingerspelling Acquisition from Demonstrations for Human-Robot Interaction
par: Tavella, Federico, et autres
Publié: (2022)
par: Tavella, Federico, et autres
Publié: (2022)
FrontierNet: Learning Visual Cues to Explore
par: Sun, Boyang, et autres
Publié: (2025)
par: Sun, Boyang, et autres
Publié: (2025)
Human Impression of Humanoid Robots Mirroring Social Cues
par: Fu, Di, et autres
Publié: (2024)
par: Fu, Di, et autres
Publié: (2024)
SCOPE for Hexapod Gait Generation
par: O'Connor, Jim, et autres
Publié: (2025)
par: O'Connor, Jim, et autres
Publié: (2025)
Documents similaires
-
Visual Cues Support Robust Turn-taking Prediction in Noise
par: Russell, Sam O'Connor, et autres
Publié: (2025) -
The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations
par: Russell, Sam OConnor, et autres
Publié: (2026) -
Applying General Turn-taking Models to Conversational Human-Robot Interaction
par: Skantze, Gabriel, et autres
Publié: (2025) -
A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
par: Inoue, Koji, et autres
Publié: (2025) -
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
par: Zhu, Wang Bill, et autres
Publié: (2025)