Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ferreira, Alef Iury Siqueira, Gris, Lucas Rafael, Filho, Alexandre Ferro, Ólives, Lucas, Ribeiro, Daniel, Fernando, Luiz, Lustosa, Fernanda, Tanaka, Rodrigo, de Oliveira, Frederico Santos, Filho, Arlindo Galvão |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
von: Shome, Debaditya, et al.
Veröffentlicht: (2023)
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
MEDUSA: A Multimodal Deep Fusion Multi-Stage Training Framework for Speech Emotion Recognition in Naturalistic Conditions
von: Chatzichristodoulou, Georgios, et al.
Veröffentlicht: (2025)
von: Chatzichristodoulou, Georgios, et al.
Veröffentlicht: (2025)
Developing a High-performance Framework for Speech Emotion Recognition in Naturalistic Conditions Challenge for Emotional Attribute Prediction
von: Lertpetchpun, Thanathai, et al.
Veröffentlicht: (2025)
von: Lertpetchpun, Thanathai, et al.
Veröffentlicht: (2025)
Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech
von: Corrêa, Pedro, et al.
Veröffentlicht: (2025)
von: Corrêa, Pedro, et al.
Veröffentlicht: (2025)
FreeSVC: Towards Zero-shot Multilingual Singing Voice Conversion
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
von: Ferreira, Alef Iury Siqueira, et al.
Veröffentlicht: (2025)
Crab: Multi Layer Contrastive Supervision to Improve Speech Emotion Recognition Under Both Acted and Natural Speech Condition
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
von: Ueda, Lucas H., et al.
Veröffentlicht: (2026)
Bimodal Connection Attention Fusion for Speech Emotion Recognition
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
Tagarela - A Portuguese speech dataset from podcasts
von: de Oliveira, Frederico Santos, et al.
Veröffentlicht: (2026)
von: de Oliveira, Frederico Santos, et al.
Veröffentlicht: (2026)
MSP-Conversation: A Corpus for Naturalistic, Time-Continuous Emotion Recognition
von: Martinez-Lucas, Luz, et al.
Veröffentlicht: (2026)
von: Martinez-Lucas, Luz, et al.
Veröffentlicht: (2026)
EMO-SUPERB: An In-depth Look at Speech Emotion Recognition
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Color-based Emotion Representation for Speech Emotion Recognition
von: Nagase, Ryotaro, et al.
Veröffentlicht: (2026)
von: Nagase, Ryotaro, et al.
Veröffentlicht: (2026)
Leveraging Speech PTM, Text LLM, and Emotional TTS for Speech Emotion Recognition
von: Ma, Ziyang, et al.
Veröffentlicht: (2023)
von: Ma, Ziyang, et al.
Veröffentlicht: (2023)
Speech Emotion Recognition with ASR Integration
von: Li, Yuanchao
Veröffentlicht: (2026)
von: Li, Yuanchao
Veröffentlicht: (2026)
Explainable Transformer-CNN Fusion for Noise-Robust Speech Emotion Recognition
von: Chakrabarty, Sudip, et al.
Veröffentlicht: (2025)
von: Chakrabarty, Sudip, et al.
Veröffentlicht: (2025)
Emotion Neural Transducer for Fine-Grained Speech Emotion Recognition
von: Shen, Siyuan, et al.
Veröffentlicht: (2024)
von: Shen, Siyuan, et al.
Veröffentlicht: (2024)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
von: Zhang, Hezhao, et al.
Veröffentlicht: (2026)
von: Zhang, Hezhao, et al.
Veröffentlicht: (2026)
Improving Speech Emotion Recognition Through Cross Modal Attention Alignment and Balanced Stacking Model
von: Ueda, Lucas, et al.
Veröffentlicht: (2025)
von: Ueda, Lucas, et al.
Veröffentlicht: (2025)
Test-Time Adaptation for Speech Emotion Recognition
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
On the Contribution of Lexical Features to Speech Emotion Recognition
von: Combei, David
Veröffentlicht: (2025)
von: Combei, David
Veröffentlicht: (2025)
Adapting WavLM for Speech Emotion Recognition
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
Enhancing Speech Emotion Recognition with Multi-Task Learning and Dynamic Feature Fusion
von: Wang, Honghong, et al.
Veröffentlicht: (2025)
von: Wang, Honghong, et al.
Veröffentlicht: (2025)
Multi-Channel Speech Enhancement for Cocktail Party Speech Emotion Recognition
von: Chen, Youjun, et al.
Veröffentlicht: (2026)
von: Chen, Youjun, et al.
Veröffentlicht: (2026)
E‐Speech: Development of a Dataset for Speech Emotion Recognition and Analysis
von: Wenjin Liu, et al.
Veröffentlicht: (2024)
von: Wenjin Liu, et al.
Veröffentlicht: (2024)
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
von: Gao, Yuan, et al.
Veröffentlicht: (2025)
Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model
von: Ihori, Mana, et al.
Veröffentlicht: (2025)
von: Ihori, Mana, et al.
Veröffentlicht: (2025)
Syntactic and Prosodic Phrasal Alignment in Naturalistic Language
von: Julie Bannon, et al.
Veröffentlicht: (2026)
von: Julie Bannon, et al.
Veröffentlicht: (2026)
Are Paralinguistic Representations all that is needed for Speech Emotion Recognition?
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
Hardware-Aware Federated Learning for Speech Emotion Recognition
von: Yuksel, Beyazit Bestami, et al.
Veröffentlicht: (2026)
von: Yuksel, Beyazit Bestami, et al.
Veröffentlicht: (2026)
Iterative Feature Boosting for Explainable Speech Emotion Recognition
von: Nfissi, Alaa, et al.
Veröffentlicht: (2024)
von: Nfissi, Alaa, et al.
Veröffentlicht: (2024)
Multi-Scale Temporal Transformer For Speech Emotion Recognition
von: Li, Zhipeng, et al.
Veröffentlicht: (2024)
von: Li, Zhipeng, et al.
Veröffentlicht: (2024)
Pre-Finetuning for Few-Shot Emotional Speech Recognition
von: Chen, Maximillian, et al.
Veröffentlicht: (2023)
von: Chen, Maximillian, et al.
Veröffentlicht: (2023)
Human Feedback Driven Dynamic Speech Emotion Recognition
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Dataset-Distillation Generative Model for Speech Emotion Recognition
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2024)
von: Ritter-Gutierrez, Fabian, et al.
Veröffentlicht: (2024)
THAI Speech Emotion Recognition (THAI-SER) corpus
von: Wongpithayadisai, Jilamika, et al.
Veröffentlicht: (2025)
von: Wongpithayadisai, Jilamika, et al.
Veröffentlicht: (2025)
Investigating the Impact of Word Informativeness on Speech Emotion Recognition
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
Speech Emotion Recognition Software System for Forensic Analysis
von: GABRIEL ELÍAS CHANCHÍ-GOLONDRINO
Veröffentlicht: (2024)
von: GABRIEL ELÍAS CHANCHÍ-GOLONDRINO
Veröffentlicht: (2024)
Persian Speech Emotion Recognition by Fine-Tuning Transformers
von: Shayaninasab, Minoo, et al.
Veröffentlicht: (2024)
von: Shayaninasab, Minoo, et al.
Veröffentlicht: (2024)
Iterative Prototype Refinement for Ambiguous Speech Emotion Recognition
von: Sun, Haoqin, et al.
Veröffentlicht: (2024)
von: Sun, Haoqin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Speech Emotion Recognition with Distilled Prosodic and Linguistic Affect Representations
von: Shome, Debaditya, et al.
Veröffentlicht: (2023) -
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
von: Dutta, Soumya, et al.
Veröffentlicht: (2025) -
MEDUSA: A Multimodal Deep Fusion Multi-Stage Training Framework for Speech Emotion Recognition in Naturalistic Conditions
von: Chatzichristodoulou, Georgios, et al.
Veröffentlicht: (2025) -
Developing a High-performance Framework for Speech Emotion Recognition in Naturalistic Conditions Challenge for Emotional Attribute Prediction
von: Lertpetchpun, Thanathai, et al.
Veröffentlicht: (2025) -
Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech
von: Corrêa, Pedro, et al.
Veröffentlicht: (2025)