What Do Humans Hear When Interacting? Experiments on Selective Listening for Evaluating ASR of Spoken Dialogue Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Mori, Kiyotada, Kawano, Seiya, Liu, Chaoran, Ishi, Carlos Toshinori, Contreras, Angel Fernando Garcia, Yoshino, Koichiro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dialogue Response Prefetching Based on Semantic Similarity and Prediction Confidence of Language Model
by: Mori, Kiyotada, et al.
Published: (2025)
by: Mori, Kiyotada, et al.
Published: (2025)
Rapport-Driven Virtual Agent: Rapport Building Dialogue Strategy for Improving User Experience at First Meeting
by: Baihaqi, Muhammad Yeza, et al.
Published: (2024)
by: Baihaqi, Muhammad Yeza, et al.
Published: (2024)
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
by: Kawano, Seiya, et al.
Published: (2024)
by: Kawano, Seiya, et al.
Published: (2024)
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
by: Sato, Takuma, et al.
Published: (2025)
by: Sato, Takuma, et al.
Published: (2025)
Training Dialogue Systems by AI Feedback for Improving Overall Dialogue Impression
by: Yoshida, Kai, et al.
Published: (2025)
by: Yoshida, Kai, et al.
Published: (2025)
WarrantScore: Modeling Warrants between Claims and Evidence for Substantiation Evaluation in Peer Reviews
by: Mori, Kiyotada, et al.
Published: (2026)
by: Mori, Kiyotada, et al.
Published: (2026)
ASMR: Augmenting Life Scenario using Large Generative Models for Robotic Action Reflection
by: Tsai, Shang-Chi, et al.
Published: (2025)
by: Tsai, Shang-Chi, et al.
Published: (2025)
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
by: Inadumi, Shun, et al.
Published: (2024)
by: Inadumi, Shun, et al.
Published: (2024)
Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures
by: Inadumi, Shun, et al.
Published: (2025)
by: Inadumi, Shun, et al.
Published: (2025)
Who Spoke What When? Evaluating Spoken Language Models for Conversational ASR with Semantic and Overlap-Aware Metrics
by: Tawara, Naohiro, et al.
Published: (2026)
by: Tawara, Naohiro, et al.
Published: (2026)
Quadrupedal Spine Control Strategies: Exploring Correlations Between System Dynamic Responses and Human Perspectives
by: Hafner, Nicholas, et al.
Published: (2025)
by: Hafner, Nicholas, et al.
Published: (2025)
J-CRe3: A Japanese Conversation Dataset for Real-world Reference Resolution
by: Ueda, Nobuhiro, et al.
Published: (2024)
by: Ueda, Nobuhiro, et al.
Published: (2024)
Audio MultiChallenge: A Multi-Turn Evaluation of Spoken Dialogue Systems on Natural Human Interaction
by: Gosai, Advait, et al.
Published: (2025)
by: Gosai, Advait, et al.
Published: (2025)
Aligning Spoken Dialogue Models from User Interactions
by: Wu, Anne, et al.
Published: (2025)
by: Wu, Anne, et al.
Published: (2025)
I Know Your Feelings Before You Do: Predicting Future Affective Reactions in Human-Computer Dialogue
by: Li, Yuanchao, et al.
Published: (2023)
by: Li, Yuanchao, et al.
Published: (2023)
Wide frequency-range acceleration using second harmonic RF bucket in fixed field accelerators
by: Uesugi, Tomonori, et al.
Published: (2025)
by: Uesugi, Tomonori, et al.
Published: (2025)
How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue
by: Lu, Hui, et al.
Published: (2026)
by: Lu, Hui, et al.
Published: (2026)
Hear You in Silence: Designing for Active Listening in Human Interaction with Conversational Agents Using Context-Aware Pacing
by: Jiang, Zhihan, et al.
Published: (2026)
by: Jiang, Zhihan, et al.
Published: (2026)
Are LLMs Robust for Spoken Dialogues?
by: Mousavi, Seyed Mahed, et al.
Published: (2024)
by: Mousavi, Seyed Mahed, et al.
Published: (2024)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
by: Nakata, Wataru, et al.
Published: (2024)
by: Nakata, Wataru, et al.
Published: (2024)
Listen Up to Protect Your Hearing
Published: (2026)
Published: (2026)
An Analysis of User Behaviors for Objectively Evaluating Spoken Dialogue Systems
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
WavReward: Spoken Dialogue Models With Generalist Reward Evaluators
by: Ji, Shengpeng, et al.
Published: (2025)
by: Ji, Shengpeng, et al.
Published: (2025)
HearSay Benchmark: Do Audio LLMs Leak What They Hear?
by: Wang, Jin, et al.
Published: (2026)
by: Wang, Jin, et al.
Published: (2026)
Just ASR + LLM? A Study on Speech Large Language Models' Ability to Identify and Understand Speaker in Spoken Dialogue
by: Wu, Junkai, et al.
Published: (2024)
by: Wu, Junkai, et al.
Published: (2024)
The Body Speaks: What Do Nurses Hear?
by: Lee SmithBattle, et al.
Published: (2025)
by: Lee SmithBattle, et al.
Published: (2025)
SpokenUS: A Spoken User Simulator for Task-Oriented Dialogue
by: Lee, Jonggeun, et al.
Published: (2026)
by: Lee, Jonggeun, et al.
Published: (2026)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
by: Wu, Yihao, et al.
Published: (2025)
by: Wu, Yihao, et al.
Published: (2025)
The Oracle Has Spoken: A Multi-Aspect Evaluation of Dialogue in Pythia
by: Chen, Zixun, et al.
Published: (2025)
by: Chen, Zixun, et al.
Published: (2025)
Adapting Text-based Dialogue State Tracker for Spoken Dialogues
by: Yoon, Jaeseok, et al.
Published: (2023)
by: Yoon, Jaeseok, et al.
Published: (2023)
Towards Stable and Personalised Profiles for Lexical Alignment in Spoken Human-Agent Dialogue
by: Schaaij, Keara, et al.
Published: (2025)
by: Schaaij, Keara, et al.
Published: (2025)
Listen, Correct, and Feed Back: Spoken Pedagogical Feedback Generation
by: Liang, Junhong, et al.
Published: (2026)
by: Liang, Junhong, et al.
Published: (2026)
SHANKS: Simultaneous Hearing and Thinking for Spoken Language Models
by: Chiang, Cheng-Han, et al.
Published: (2025)
by: Chiang, Cheng-Han, et al.
Published: (2025)
Spoken DialogSum: An Emotion-Rich Conversational Dataset for Spoken Dialogue Summarization
by: Lu, Yen-Ju, et al.
Published: (2025)
by: Lu, Yen-Ju, et al.
Published: (2025)
Practical Causal Evaluation Metrics for Biological Networks
by: Sato, Noriaki, et al.
Published: (2025)
by: Sato, Noriaki, et al.
Published: (2025)
Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems
by: Peng, Yizhou, et al.
Published: (2026)
by: Peng, Yizhou, et al.
Published: (2026)
LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems
by: Zhang, Hao, et al.
Published: (2025)
by: Zhang, Hao, et al.
Published: (2025)
FormalASR: End-to-End Spoken Chinese to Formal Text
by: Ning, Wanyi, et al.
Published: (2026)
by: Ning, Wanyi, et al.
Published: (2026)
URO-Bench: Towards Comprehensive Evaluation for End-to-End Spoken Dialogue Models
by: Yan, Ruiqi, et al.
Published: (2025)
by: Yan, Ruiqi, et al.
Published: (2025)
MOSS-TTSD: Text to Spoken Dialogue Generation
by: Zhang, Yuqian, et al.
Published: (2026)
by: Zhang, Yuqian, et al.
Published: (2026)
Similar Items
-
Dialogue Response Prefetching Based on Semantic Similarity and Prediction Confidence of Language Model
by: Mori, Kiyotada, et al.
Published: (2025) -
Rapport-Driven Virtual Agent: Rapport Building Dialogue Strategy for Improving User Experience at First Meeting
by: Baihaqi, Muhammad Yeza, et al.
Published: (2024) -
ClaimBrush: A Novel Framework for Automated Patent Claim Refinement Based on Large Language Models
by: Kawano, Seiya, et al.
Published: (2024) -
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
by: Sato, Takuma, et al.
Published: (2025) -
Training Dialogue Systems by AI Feedback for Improving Overall Dialogue Impression
by: Yoshida, Kai, et al.
Published: (2025)