Semantic Differentiation in Speech Emotion Recognition: Insights from Descriptive and Expressive Speech Roles
Fuente:
arXiv
Salvato in:
| Autori principali: | Guo, Rongchen, Francoeur, Vincent, Nejadgholi, Isar, Gagnon, Sylvain, Bolic, Miodrag |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models
di: Ghanadian, Hamideh, et al.
Pubblicazione: (2024)
di: Ghanadian, Hamideh, et al.
Pubblicazione: (2024)
A Taxonomy for Design and Evaluation of Prompt-Based Natural Language Explanations
di: Nejadgholi, Isar, et al.
Pubblicazione: (2025)
di: Nejadgholi, Isar, et al.
Pubblicazione: (2025)
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
di: Ma, Ziyang, et al.
Pubblicazione: (2023)
di: Ma, Ziyang, et al.
Pubblicazione: (2023)
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
di: Trinh, Viet Anh, et al.
Pubblicazione: (2025)
di: Trinh, Viet Anh, et al.
Pubblicazione: (2025)
SeamlessExpressiveLM: Speech Language Model for Expressive Speech-to-Speech Translation with Chain-of-Thought
di: Gong, Hongyu, et al.
Pubblicazione: (2024)
di: Gong, Hongyu, et al.
Pubblicazione: (2024)
SER Evals: In-domain and Out-of-domain Benchmarking for Speech Emotion Recognition
di: Osman, Mohamed, et al.
Pubblicazione: (2024)
di: Osman, Mohamed, et al.
Pubblicazione: (2024)
Adaptable Moral Stances of Large Language Models on Sexist Content: Implications for Society and Gender Discourse
di: Guo, Rongchen, et al.
Pubblicazione: (2024)
di: Guo, Rongchen, et al.
Pubblicazione: (2024)
Hybrid CNN-Transformer Architecture for Arabic Speech Emotion Recognition
di: Gheffari, Youcef Soufiane, et al.
Pubblicazione: (2026)
di: Gheffari, Youcef Soufiane, et al.
Pubblicazione: (2026)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
di: Zhang, Hezhao, et al.
Pubblicazione: (2026)
di: Zhang, Hezhao, et al.
Pubblicazione: (2026)
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
di: Shome, Debaditya, et al.
Pubblicazione: (2023)
di: Shome, Debaditya, et al.
Pubblicazione: (2023)
SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition
di: Dkhissi, Youness, et al.
Pubblicazione: (2026)
di: Dkhissi, Youness, et al.
Pubblicazione: (2026)
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers
di: Serai, Prashant, et al.
Pubblicazione: (2024)
di: Serai, Prashant, et al.
Pubblicazione: (2024)
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
di: Attaluri, Kaushal, et al.
Pubblicazione: (2024)
di: Attaluri, Kaushal, et al.
Pubblicazione: (2024)
How do Hyenas deal with Human Speech? Speech Recognition and Translation with ConfHyena
di: Gaido, Marco, et al.
Pubblicazione: (2024)
di: Gaido, Marco, et al.
Pubblicazione: (2024)
Bimodal Connection Attention Fusion for Speech Emotion Recognition
di: Luo, Jiachen, et al.
Pubblicazione: (2025)
di: Luo, Jiachen, et al.
Pubblicazione: (2025)
Adapting Foundation Speech Recognition Models to Impaired Speech: A Semantic Re-chaining Approach for Personalization of German Speech
di: Pokel, Niclas, et al.
Pubblicazione: (2025)
di: Pokel, Niclas, et al.
Pubblicazione: (2025)
Semantically Corrected Amharic Automatic Speech Recognition
di: Adnew, Samuael, et al.
Pubblicazione: (2024)
di: Adnew, Samuael, et al.
Pubblicazione: (2024)
Automatic Speech Recognition for Sanskrit with Transfer Learning
di: Sadhukhan, Bidit, et al.
Pubblicazione: (2025)
di: Sadhukhan, Bidit, et al.
Pubblicazione: (2025)
Automatic Speech Recognition for Greek Medical Dictation
di: Georgilas, Vardis, et al.
Pubblicazione: (2025)
di: Georgilas, Vardis, et al.
Pubblicazione: (2025)
Towards Unsupervised Speech Recognition at the Syllable-Level
di: Wang, Liming, et al.
Pubblicazione: (2025)
di: Wang, Liming, et al.
Pubblicazione: (2025)
Amplifying Emotional Signals: Data-Efficient Deep Learning for Robust Speech Emotion Recognition
di: Vu, Tai
Pubblicazione: (2025)
di: Vu, Tai
Pubblicazione: (2025)
Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition
di: Wang, Peng, et al.
Pubblicazione: (2026)
di: Wang, Peng, et al.
Pubblicazione: (2026)
In-context Language Learning for Endangered Languages in Speech Recognition
di: Li, Zhaolin, et al.
Pubblicazione: (2025)
di: Li, Zhaolin, et al.
Pubblicazione: (2025)
Augmenting Automatic Speech Recognition Models with Disfluency Detection
di: Amann, Robin, et al.
Pubblicazione: (2024)
di: Amann, Robin, et al.
Pubblicazione: (2024)
Improved Contextual Recognition In Automatic Speech Recognition Systems By Semantic Lattice Rescoring
di: Sudarshan, Ankitha, et al.
Pubblicazione: (2023)
di: Sudarshan, Ankitha, et al.
Pubblicazione: (2023)
Beyond Global Emotion: Fine-Grained Emotional Speech Synthesis with Dynamic Word-Level Modulation
di: Wang, Sirui, et al.
Pubblicazione: (2025)
di: Wang, Sirui, et al.
Pubblicazione: (2025)
DepFlow: Disentangled Speech Generation to Mitigate Semantic Bias in Depression Detection
di: Li, Yuxin, et al.
Pubblicazione: (2026)
di: Li, Yuxin, et al.
Pubblicazione: (2026)
WMT24 Test Suite: Gender Resolution in Speaker-Listener Dialogue Roles
di: Dawkins, Hillary, et al.
Pubblicazione: (2024)
di: Dawkins, Hillary, et al.
Pubblicazione: (2024)
Speech Emotion Recognition Leveraging OpenAI's Whisper Representations and Attentive Pooling Methods
di: Shendabadi, Ali, et al.
Pubblicazione: (2026)
di: Shendabadi, Ali, et al.
Pubblicazione: (2026)
Decoding Emotion: Speech Perception Patterns in Individuals with Self-reported Depression
di: Vats, Guneesh, et al.
Pubblicazione: (2024)
di: Vats, Guneesh, et al.
Pubblicazione: (2024)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
di: Min, Do June, et al.
Pubblicazione: (2024)
di: Min, Do June, et al.
Pubblicazione: (2024)
ViSpeR: Multilingual Audio-Visual Speech Recognition
di: Narayan, Sanath, et al.
Pubblicazione: (2024)
di: Narayan, Sanath, et al.
Pubblicazione: (2024)
Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
di: Lee, Wonjun, et al.
Pubblicazione: (2026)
di: Lee, Wonjun, et al.
Pubblicazione: (2026)
Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information
di: Taguchi, Chihiro, et al.
Pubblicazione: (2024)
di: Taguchi, Chihiro, et al.
Pubblicazione: (2024)
IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech
di: Zhou, Siyi, et al.
Pubblicazione: (2025)
di: Zhou, Siyi, et al.
Pubblicazione: (2025)
UDDETTS: Unifying Discrete and Dimensional Emotions for Controllable Emotional Text-to-Speech
di: Liu, Jiaxuan, et al.
Pubblicazione: (2025)
di: Liu, Jiaxuan, et al.
Pubblicazione: (2025)
Whispering Context: Distilling Syntax and Semantics for Long Speech Transcripts
di: Altinok, Duygu
Pubblicazione: (2025)
di: Altinok, Duygu
Pubblicazione: (2025)
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
di: Chen, Jianan, et al.
Pubblicazione: (2026)
di: Chen, Jianan, et al.
Pubblicazione: (2026)
Understanding Emotion in Discourse: Recognition Insights and Linguistic Patterns for Generation
di: Jeong, Cheonkam, et al.
Pubblicazione: (2026)
di: Jeong, Cheonkam, et al.
Pubblicazione: (2026)
Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias
di: Ogunnubi, Tomisin, et al.
Pubblicazione: (2026)
di: Ogunnubi, Tomisin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models
di: Ghanadian, Hamideh, et al.
Pubblicazione: (2024) -
A Taxonomy for Design and Evaluation of Prompt-Based Natural Language Explanations
di: Nejadgholi, Isar, et al.
Pubblicazione: (2025) -
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
di: Ma, Ziyang, et al.
Pubblicazione: (2023) -
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
di: Trinh, Viet Anh, et al.
Pubblicazione: (2025) -
SeamlessExpressiveLM: Speech Language Model for Expressive Speech-to-Speech Translation with Chain-of-Thought
di: Gong, Hongyu, et al.
Pubblicazione: (2024)