StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Haishu, Hao, Aokai, Ge, Yuan, Hong, Zhenqiang, Xiao, Tong, Zhu, Jingbo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StyleBench: Evaluating thinking styles in Large Language Models
von: Guo, Junyu, et al.
Veröffentlicht: (2025)
von: Guo, Junyu, et al.
Veröffentlicht: (2025)
On the Emotion Understanding of Synthesized Speech
von: Ge, Yuan, et al.
Veröffentlicht: (2026)
von: Ge, Yuan, et al.
Veröffentlicht: (2026)
Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models
von: Lin, Yu-Xiang, et al.
Veröffentlicht: (2025)
von: Lin, Yu-Xiang, et al.
Veröffentlicht: (2025)
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
von: Kang, Wonjune, et al.
Veröffentlicht: (2025)
Style-Talker: Finetuning Audio Language Model and Style-Based Text-to-Speech Model for Fast Spoken Dialogue Generation
von: Li, Yinghao Aaron, et al.
Veröffentlicht: (2024)
von: Li, Yinghao Aaron, et al.
Veröffentlicht: (2024)
SpeechCaps: Advancing Instruction-Based Universal Speech Models with Multi-Talker Speaking Style Captioning
von: Huang, Chien-yu, et al.
Veröffentlicht: (2024)
von: Huang, Chien-yu, et al.
Veröffentlicht: (2024)
Advancing Large Language Models to Capture Varied Speaking Styles and Respond Properly in Spoken Conversations
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
von: Kang, Jaehoon, et al.
Veröffentlicht: (2026)
von: Kang, Jaehoon, et al.
Veröffentlicht: (2026)
RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
Audio-Aware Large Language Models as Judges for Speaking Styles
von: Chiang, Cheng-Han, et al.
Veröffentlicht: (2025)
von: Chiang, Cheng-Han, et al.
Veröffentlicht: (2025)
mStyleDistance: Multilingual Style Embeddings and their Evaluation
von: Qiu, Justin, et al.
Veröffentlicht: (2025)
von: Qiu, Justin, et al.
Veröffentlicht: (2025)
LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning
von: Kawamura, Masaya, et al.
Veröffentlicht: (2024)
von: Kawamura, Masaya, et al.
Veröffentlicht: (2024)
Foundations of Large Language Models
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
UltraVoice: Scaling Fine-Grained Style-Controlled Speech Conversations for Spoken Dialogue Models
von: Tu, Wenming, et al.
Veröffentlicht: (2025)
von: Tu, Wenming, et al.
Veröffentlicht: (2025)
Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
von: Woszczyk, Dominika, et al.
Veröffentlicht: (2025)
Factor-Conditioned Speaking-Style Captioning
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
von: Ando, Atsushi, et al.
Veröffentlicht: (2024)
FLEXI: Benchmarking Full-duplex Human-LLM Speech Interaction
von: Ge, Yuan, et al.
Veröffentlicht: (2025)
von: Ge, Yuan, et al.
Veröffentlicht: (2025)
Lombard Speech Synthesis for Any Voice with Controllable Style Embeddings
von: Akti, Seymanur, et al.
Veröffentlicht: (2026)
von: Akti, Seymanur, et al.
Veröffentlicht: (2026)
DiffStyleTTS: Diffusion-based Hierarchical Prosody Modeling for Text-to-Speech with Diverse and Controllable Styles
von: Liu, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Liu, Jiaxuan, et al.
Veröffentlicht: (2024)
Substance over Style: Evaluating Proactive Conversational Coaching Agents
von: Srinivas, Vidya, et al.
Veröffentlicht: (2025)
von: Srinivas, Vidya, et al.
Veröffentlicht: (2025)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
von: Brown, Andrew, et al.
Veröffentlicht: (2024)
Languages in Whisper-Style Speech Encoders Align Both Phonetically and Semantically
von: Shim, Ryan Soh-Eun, et al.
Veröffentlicht: (2025)
von: Shim, Ryan Soh-Eun, et al.
Veröffentlicht: (2025)
Mind the Style Gap: Meta-Evaluation of Style and Attribute Transfer Metrics
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2025)
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2025)
PSST: A Benchmark for Evaluation-driven Text Public-Speaking Style Transfer
von: Sun, Huashan, et al.
Veröffentlicht: (2023)
von: Sun, Huashan, et al.
Veröffentlicht: (2023)
Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models
von: Li, Weiqin, et al.
Veröffentlicht: (2024)
von: Li, Weiqin, et al.
Veröffentlicht: (2024)
MTP-S2UT: Enhancing Speech-to-Speech Translation Quality with Multi-token Prediction
von: Wang, Jianjin, et al.
Veröffentlicht: (2025)
von: Wang, Jianjin, et al.
Veröffentlicht: (2025)
ISCA: A Framework for Interview-Style Conversational Agents
von: Welch, Charles, et al.
Veröffentlicht: (2025)
von: Welch, Charles, et al.
Veröffentlicht: (2025)
RankPrompt: Step-by-Step Comparisons Make Language Models Better Reasoners
von: Hu, Chi, et al.
Veröffentlicht: (2024)
von: Hu, Chi, et al.
Veröffentlicht: (2024)
Style-Compress: An LLM-Based Prompt Compression Framework Considering Task-Specific Styles
von: Pu, Xiao, et al.
Veröffentlicht: (2024)
von: Pu, Xiao, et al.
Veröffentlicht: (2024)
POWSM: A Phonetic Open Whisper-Style Speech Foundation Model
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
Large Language Models on Wikipedia-Style Survey Generation: an Evaluation in NLP Concepts
von: Gao, Fan, et al.
Veröffentlicht: (2023)
von: Gao, Fan, et al.
Veröffentlicht: (2023)
Replacing Language Model for Style Transfer
von: Cheng, Pengyu, et al.
Veröffentlicht: (2022)
von: Cheng, Pengyu, et al.
Veröffentlicht: (2022)
Teaching Language Models to Self-Improve by Learning from Language Feedback
von: Hu, Chi, et al.
Veröffentlicht: (2024)
von: Hu, Chi, et al.
Veröffentlicht: (2024)
A Concise Agent is Less Expert: Revealing Side Effects of Using Style Features on Conversational Agents
von: Cho, Young-Min, et al.
Veröffentlicht: (2026)
von: Cho, Young-Min, et al.
Veröffentlicht: (2026)
Controlling Chat Style in Language Models via Single-Direction Editing
von: Xu, Zhenyu, et al.
Veröffentlicht: (2026)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2026)
CAT-LLM: Style-enhanced Large Language Models with Text Style Definition for Chinese Article-style Transfer
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
Attention2Probability: Attention-Driven Terminology Probability Estimation for Robust Speech-to-Text System
von: Du, Yanfan, et al.
Veröffentlicht: (2025)
von: Du, Yanfan, et al.
Veröffentlicht: (2025)
SC2: Towards Enhancing Content Preservation and Style Consistency in Long Text Style Transfer
von: Zhao, Jie, et al.
Veröffentlicht: (2024)
von: Zhao, Jie, et al.
Veröffentlicht: (2024)
Style Vectors for Steering Generative Large Language Model
von: Konen, Kai, et al.
Veröffentlicht: (2024)
von: Konen, Kai, et al.
Veröffentlicht: (2024)
Capturing Human Cognitive Styles with Language: Towards an Experimental Evaluation Paradigm
von: Varadarajan, Vasudha, et al.
Veröffentlicht: (2025)
von: Varadarajan, Vasudha, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
StyleBench: Evaluating thinking styles in Large Language Models
von: Guo, Junyu, et al.
Veröffentlicht: (2025) -
On the Emotion Understanding of Synthesized Speech
von: Ge, Yuan, et al.
Veröffentlicht: (2026) -
Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models
von: Lin, Yu-Xiang, et al.
Veröffentlicht: (2025) -
Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style
von: Kang, Wonjune, et al.
Veröffentlicht: (2025) -
Style-Talker: Finetuning Audio Language Model and Style-Based Text-to-Speech Model for Fast Spoken Dialogue Generation
von: Li, Yinghao Aaron, et al.
Veröffentlicht: (2024)