Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Yihao, Wang, Tianrui, Peng, Yizhou, Chao, Yi-Wen, Zhuang, Xuyi, Wang, Xinsheng, Yin, Shunshun, Ma, Ziyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems
di: Peng, Yizhou, et al.
Pubblicazione: (2026)
di: Peng, Yizhou, et al.
Pubblicazione: (2026)
FD-Bench: A Full-Duplex Benchmarking Pipeline Designed for Full Duplex Spoken Dialogue Systems
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
Are LLMs Robust for Spoken Dialogues?
di: Mousavi, Seyed Mahed, et al.
Pubblicazione: (2024)
di: Mousavi, Seyed Mahed, et al.
Pubblicazione: (2024)
URO-Bench: Towards Comprehensive Evaluation for End-to-End Spoken Dialogue Models
di: Yan, Ruiqi, et al.
Pubblicazione: (2025)
di: Yan, Ruiqi, et al.
Pubblicazione: (2025)
How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue
di: Lu, Hui, et al.
Pubblicazione: (2026)
di: Lu, Hui, et al.
Pubblicazione: (2026)
The ICASSP 2026 HumDial Challenge: Benchmarking Human-like Spoken Dialogue Systems in the LLM Era
di: Zhao, Zhixian, et al.
Pubblicazione: (2026)
di: Zhao, Zhixian, et al.
Pubblicazione: (2026)
SAC: Neural Speech Codec with Semantic-Acoustic Dual-Stream Quantization
di: Chen, Wenxi, et al.
Pubblicazione: (2025)
di: Chen, Wenxi, et al.
Pubblicazione: (2025)
Spoken DialogSum: An Emotion-Rich Conversational Dataset for Spoken Dialogue Summarization
di: Lu, Yen-Ju, et al.
Pubblicazione: (2025)
di: Lu, Yen-Ju, et al.
Pubblicazione: (2025)
OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue
di: Geng, Xuelong, et al.
Pubblicazione: (2025)
di: Geng, Xuelong, et al.
Pubblicazione: (2025)
UltraVoice: Scaling Fine-Grained Style-Controlled Speech Conversations for Spoken Dialogue Models
di: Tu, Wenming, et al.
Pubblicazione: (2025)
di: Tu, Wenming, et al.
Pubblicazione: (2025)
Analysis of Frequency Collisions in Parametrically Modulated Superconducting Circuits
di: Ma, Zhuang, et al.
Pubblicazione: (2025)
di: Ma, Zhuang, et al.
Pubblicazione: (2025)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
di: Nakata, Wataru, et al.
Pubblicazione: (2024)
di: Nakata, Wataru, et al.
Pubblicazione: (2024)
HPSU: A Benchmark for Human-Level Perception in Real-World Spoken Speech Understanding
di: Li, Chen, et al.
Pubblicazione: (2025)
di: Li, Chen, et al.
Pubblicazione: (2025)
An Analysis of User Behaviors for Objectively Evaluating Spoken Dialogue Systems
di: Inoue, Koji, et al.
Pubblicazione: (2024)
di: Inoue, Koji, et al.
Pubblicazione: (2024)
WavReward: Spoken Dialogue Models With Generalist Reward Evaluators
di: Ji, Shengpeng, et al.
Pubblicazione: (2025)
di: Ji, Shengpeng, et al.
Pubblicazione: (2025)
Hard vs. Noise: Resolving Hard-Noisy Sample Confusion in Recommender Systems via Large Language Models
di: Song, Tianrui, et al.
Pubblicazione: (2025)
di: Song, Tianrui, et al.
Pubblicazione: (2025)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
di: Si, Shuzheng, et al.
Pubblicazione: (2023)
di: Si, Shuzheng, et al.
Pubblicazione: (2023)
Aligning Spoken Dialogue Models from User Interactions
di: Wu, Anne, et al.
Pubblicazione: (2025)
di: Wu, Anne, et al.
Pubblicazione: (2025)
Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
di: Lin, Guan-Ting, et al.
Pubblicazione: (2023)
di: Lin, Guan-Ting, et al.
Pubblicazione: (2023)
PATCHEVAL: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
di: Wei, Zichao, et al.
Pubblicazione: (2025)
di: Wei, Zichao, et al.
Pubblicazione: (2025)
Bongard-OpenWorld: Few-Shot Reasoning for Free-form Visual Concepts in the Real World
di: Wu, Rujie, et al.
Pubblicazione: (2023)
di: Wu, Rujie, et al.
Pubblicazione: (2023)
SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness
di: Lu, Jingyu, et al.
Pubblicazione: (2026)
di: Lu, Jingyu, et al.
Pubblicazione: (2026)
Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
di: Park, Se Jin, et al.
Pubblicazione: (2024)
di: Park, Se Jin, et al.
Pubblicazione: (2024)
SpokenUS: A Spoken User Simulator for Task-Oriented Dialogue
di: Lee, Jonggeun, et al.
Pubblicazione: (2026)
di: Lee, Jonggeun, et al.
Pubblicazione: (2026)
NTU Speechlab LLM-Based Multilingual ASR System for Interspeech MLC-SLM Challenge 2025
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
di: Peng, Yizhou, et al.
Pubblicazione: (2025)
SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words
di: Ao, Junyi, et al.
Pubblicazione: (2024)
di: Ao, Junyi, et al.
Pubblicazione: (2024)
The Oracle Has Spoken: A Multi-Aspect Evaluation of Dialogue in Pythia
di: Chen, Zixun, et al.
Pubblicazione: (2025)
di: Chen, Zixun, et al.
Pubblicazione: (2025)
Adapting Text-based Dialogue State Tracker for Spoken Dialogues
di: Yoon, Jaeseok, et al.
Pubblicazione: (2023)
di: Yoon, Jaeseok, et al.
Pubblicazione: (2023)
WavChat: A Survey of Spoken Dialogue Models
di: Ji, Shengpeng, et al.
Pubblicazione: (2024)
di: Ji, Shengpeng, et al.
Pubblicazione: (2024)
Spoken Stereoset: On Evaluating Social Bias Toward Speaker in Speech Large Language Models
di: Lin, Yi-Cheng, et al.
Pubblicazione: (2024)
di: Lin, Yi-Cheng, et al.
Pubblicazione: (2024)
UniSS: Unified Expressive Speech-to-Speech Translation with Your Voice
di: Cheng, Sitong, et al.
Pubblicazione: (2025)
di: Cheng, Sitong, et al.
Pubblicazione: (2025)
PSLM: Parallel Generation of Text and Speech with LLMs for Low-Latency Spoken Dialogue Systems
di: Mitsui, Kentaro, et al.
Pubblicazione: (2024)
di: Mitsui, Kentaro, et al.
Pubblicazione: (2024)
The Use of Real‐World Evidence for Regulatory Decisions in China
di: Jiayue Xu, et al.
Pubblicazione: (2024)
di: Jiayue Xu, et al.
Pubblicazione: (2024)
Discourse-Aware Dual-Track Streaming Response for Low-Latency Spoken Dialogue Systems
di: Liu, Siyuan, et al.
Pubblicazione: (2026)
di: Liu, Siyuan, et al.
Pubblicazione: (2026)
LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems
di: Zhang, Hao, et al.
Pubblicazione: (2025)
di: Zhang, Hao, et al.
Pubblicazione: (2025)
Leveraging Chain of Thought towards Empathetic Spoken Dialogue without Corresponding Question-Answering Data
di: Xie, Jingran, et al.
Pubblicazione: (2025)
di: Xie, Jingran, et al.
Pubblicazione: (2025)
Weak-Mamba-UNet: Visual Mamba Makes CNN and ViT Work Better for Scribble-based Medical Image Segmentation
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
Semi-Mamba-UNet: Pixel-Level Contrastive and Pixel-Level Cross-Supervised Visual Mamba-based UNet for Semi-Supervised Medical Image Segmentation
di: Ma, Chao, et al.
Pubblicazione: (2024)
di: Ma, Chao, et al.
Pubblicazione: (2024)
Evaluating the Bias in LLMs for Surveying Opinion and Decision Making in Healthcare
di: Khaokaew, Yonchanok, et al.
Pubblicazione: (2025)
di: Khaokaew, Yonchanok, et al.
Pubblicazione: (2025)
Bring the Power of Diffusion Model to Defect Detection
di: Yu, Xuyi
Pubblicazione: (2024)
di: Yu, Xuyi
Pubblicazione: (2024)
Documenti analoghi
-
Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems
di: Peng, Yizhou, et al.
Pubblicazione: (2026) -
FD-Bench: A Full-Duplex Benchmarking Pipeline Designed for Full Duplex Spoken Dialogue Systems
di: Peng, Yizhou, et al.
Pubblicazione: (2025) -
Are LLMs Robust for Spoken Dialogues?
di: Mousavi, Seyed Mahed, et al.
Pubblicazione: (2024) -
URO-Bench: Towards Comprehensive Evaluation for End-to-End Spoken Dialogue Models
di: Yan, Ruiqi, et al.
Pubblicazione: (2025) -
How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue
di: Lu, Hui, et al.
Pubblicazione: (2026)