A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
Fuente:
arXiv
Saved in:
| Main Authors: | Inoue, Koji, Okafuji, Yuki, Baba, Jun, Ohira, Yoshiki, Hyodo, Katsuya, Kawahara, Tatsuya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt-Guided Turn-Taking Prediction
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
by: Li, Yuanchao, et al.
Published: (2023)
by: Li, Yuanchao, et al.
Published: (2023)
From Metrics to Meaning: Insights from a Mixed-Methods Field Experiment on Retail Robot Deployment
by: Song, Sichao, et al.
Published: (2026)
by: Song, Sichao, et al.
Published: (2026)
Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
Multilingual Turn-taking Prediction Using Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
What Drives You to Interact?: The Role of User Motivation for a Robot in the Wild
by: Koike, Amy, et al.
Published: (2025)
by: Koike, Amy, et al.
Published: (2025)
An LLM Benchmark for Addressee Recognition in Multi-modal Multi-party Dialogue
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
by: Shi, Hao, et al.
Published: (2024)
by: Shi, Hao, et al.
Published: (2024)
Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System
by: Kato, Kazushi, et al.
Published: (2025)
by: Kato, Kazushi, et al.
Published: (2025)
Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions?
by: Pang, Zi Haur, et al.
Published: (2025)
by: Pang, Zi Haur, et al.
Published: (2025)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
by: Miyoshi, Ryo, et al.
Published: (2025)
by: Miyoshi, Ryo, et al.
Published: (2025)
Paralinguistic Emotion-Aware Validation Timing Detection in Japanese Empathetic Spoken Dialogue
by: Pang, Zi Haur, et al.
Published: (2026)
by: Pang, Zi Haur, et al.
Published: (2026)
An Efficient GPU-based Implementation for Noise Robust Sound Source Localization
by: Lin, Zirui, et al.
Published: (2025)
by: Lin, Zirui, et al.
Published: (2025)
Whom to Respond To? A Transformer-Based Model for Multi-Party Social Robot Interaction
by: Zhu, He, et al.
Published: (2025)
by: Zhu, He, et al.
Published: (2025)
Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models
by: Riera, Pablo, et al.
Published: (2026)
by: Riera, Pablo, et al.
Published: (2026)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
by: Wei, Xiangyi, et al.
Published: (2025)
by: Wei, Xiangyi, et al.
Published: (2025)
Lend me an Ear: Speech Enhancement Using a Robotic Arm with a Microphone Array
by: Turcotte, Zachary, et al.
Published: (2026)
by: Turcotte, Zachary, et al.
Published: (2026)
Efficient and Robust Long-Form Speech Recognition with Hybrid H3-Conformer
by: Honda, Tomoki, et al.
Published: (2024)
by: Honda, Tomoki, et al.
Published: (2024)
Human-mimetic binaural ear design and sound source direction estimation for task realization of musculoskeletal humanoids
by: Omura, Yusuke, et al.
Published: (2024)
by: Omura, Yusuke, et al.
Published: (2024)
Triadic Multi-party Voice Activity Projection for Turn-taking in Spoken Dialogue Systems
by: Elmers, Mikey, et al.
Published: (2025)
by: Elmers, Mikey, et al.
Published: (2025)
From Turn-Taking to Synchronous Dialogue: A Survey of Full-Duplex Spoken Language Models
by: Chen, Yuxuan, et al.
Published: (2025)
by: Chen, Yuxuan, et al.
Published: (2025)
Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation
by: Nagahisa, Haruto, et al.
Published: (2026)
by: Nagahisa, Haruto, et al.
Published: (2026)
Long-Term, Store-Front Robotics: Interactive Music for Robotic Arm, Caxixi and Frame Drums
by: Savery, Richard, et al.
Published: (2024)
by: Savery, Richard, et al.
Published: (2024)
Multilingual and Continuous Backchannel Prediction: A Cross-lingual Study
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Beyond Voice Assistants: Exploring Advantages and Risks of an In-Car Social Robot in Real Driving Scenarios
by: Li, Yuanchao, et al.
Published: (2024)
by: Li, Yuanchao, et al.
Published: (2024)
Semi-Automatic Flute Robot and Its Acoustic Sensing
by: Kuriyama, Hikari, et al.
Published: (2026)
by: Kuriyama, Hikari, et al.
Published: (2026)
Two-Stage Acoustic Adaptation with Gated Cross-Attention Adapters for LLM-Based Multi-Talker Speech Recognition
by: Shi, Hao, et al.
Published: (2026)
by: Shi, Hao, et al.
Published: (2026)
Evaluating Speech-in-Speech Perception via a Humanoid Robot
by: Meyer, Luke, et al.
Published: (2023)
by: Meyer, Luke, et al.
Published: (2023)
Disentangled Acoustic Fields For Multimodal Physical Scene Understanding
by: Yin, Jie, et al.
Published: (2024)
by: Yin, Jie, et al.
Published: (2024)
Theoretical Framework for the Optimization of Microphone Array Configuration for Humanoid Robot Audition
by: Tourbabin, Vladimir, et al.
Published: (2024)
by: Tourbabin, Vladimir, et al.
Published: (2024)
Analysis and Detection of Differences in Spoken User Behaviors between Autonomous and Wizard-of-Oz Systems
by: Elmers, Mikey, et al.
Published: (2024)
by: Elmers, Mikey, et al.
Published: (2024)
Improved Extrinsic Calibration of Acoustic Cameras via Batch Optimization
by: Li, Zhi, et al.
Published: (2025)
by: Li, Zhi, et al.
Published: (2025)
Calibration of Multiple Asynchronous Microphone Arrays using Hybrid TDOA
by: Zhang, Chengjie, et al.
Published: (2025)
by: Zhang, Chengjie, et al.
Published: (2025)
SAMAY: System for Acoustic Measurement and Analysis
by: R, Adheep Arya G, et al.
Published: (2025)
by: R, Adheep Arya G, et al.
Published: (2025)
CoDeTT: A Context-Aware Decision Benchmark for Turn-Taking Evaluation
by: Shen, Huan, et al.
Published: (2026)
by: Shen, Huan, et al.
Published: (2026)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
by: Wang, Jiang, et al.
Published: (2025)
by: Wang, Jiang, et al.
Published: (2025)
Direction of Arrival Estimation Using Microphone Array Processing for Moving Humanoid Robots
by: Tourbabin, Vladimir, et al.
Published: (2024)
by: Tourbabin, Vladimir, et al.
Published: (2024)
Robot Confirmation Generation and Action Planning Using Long-context Q-Former Integrated with Multimodal LLM
by: Hori, Chiori, et al.
Published: (2025)
by: Hori, Chiori, et al.
Published: (2025)
A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods
by: Jalayer, Reza, et al.
Published: (2025)
by: Jalayer, Reza, et al.
Published: (2025)
Similar Items
-
Prompt-Guided Turn-Taking Prediction
by: Inoue, Koji, et al.
Published: (2025) -
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
by: Li, Yuanchao, et al.
Published: (2023) -
From Metrics to Meaning: Insights from a Mixed-Methods Field Experiment on Retail Robot Deployment
by: Song, Sichao, et al.
Published: (2026) -
Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024) -
Multilingual Turn-taking Prediction Using Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)