A Mixture-of-Experts Model for Multimodal Emotion Recognition in Conversations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dutta, Soumya, Balaji, Smruthi, Ganapathy, Sriram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
LLM supervised Pre-training for Multimodal Emotion Recognition in Conversations
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
Textless and Non-Parallel Speech-to-Speech Emotion Style Transfer
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
Spoken Language Understanding on Unseen Tasks With In-Context Learning
von: Agrawal, Neeraj, et al.
Veröffentlicht: (2025)
von: Agrawal, Neeraj, et al.
Veröffentlicht: (2025)
GatedxLSTM: A Multimodal Affective Computing Approach for Emotion Recognition in Conversations
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
Acoustic and Semantic Modeling of Emotion in Spoken Language
von: Dutta, Soumya
Veröffentlicht: (2026)
von: Dutta, Soumya
Veröffentlicht: (2026)
Distribution-based Emotion Recognition in Conversation
von: Wu, Wen, et al.
Veröffentlicht: (2022)
von: Wu, Wen, et al.
Veröffentlicht: (2022)
Emotion-Anchored Contrastive Learning Framework for Emotion Recognition in Conversation
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
Are Paralinguistic Representations all that is needed for Speech Emotion Recognition?
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
GSDNet: Revisiting Incomplete Multimodal-Diffusion from Graph Spectrum Perspective for Conversation Emotion Recognition
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
Enhancing Code-Switching Speech Recognition with LID-Based Collaborative Mixture of Experts Model
von: Huang, Hukai, et al.
Veröffentlicht: (2024)
von: Huang, Hukai, et al.
Veröffentlicht: (2024)
Evaluating Emotion Recognition in Spoken Language Models on Emotionally Incongruent Speech
von: Corrêa, Pedro, et al.
Veröffentlicht: (2025)
von: Corrêa, Pedro, et al.
Veröffentlicht: (2025)
Visual-Aware Speech Recognition for Noisy Scenarios
von: Balaji, Lakshmipathi, et al.
Veröffentlicht: (2025)
von: Balaji, Lakshmipathi, et al.
Veröffentlicht: (2025)
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
von: Cha, Jun-Hyeok, et al.
Veröffentlicht: (2025)
von: Cha, Jun-Hyeok, et al.
Veröffentlicht: (2025)
TelME: Teacher-leading Multimodal Fusion Network for Emotion Recognition in Conversation
von: Yun, Taeyang, et al.
Veröffentlicht: (2024)
von: Yun, Taeyang, et al.
Veröffentlicht: (2024)
Towards the Next Frontier in Speech Representation Learning Using Disentanglement
von: Krishna, Varun, et al.
Veröffentlicht: (2024)
von: Krishna, Varun, et al.
Veröffentlicht: (2024)
AIMDiT: Modality Augmentation and Interaction via Multimodal Dimension Transformation for Emotion Recognition in Conversations
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
von: Wu, Sheng, et al.
Veröffentlicht: (2024)
Enhancing Multimodal Emotion Recognition through Multi-Granularity Cross-Modal Alignment
von: Wang, Xuechen, et al.
Veröffentlicht: (2024)
von: Wang, Xuechen, et al.
Veröffentlicht: (2024)
Context and System Fusion in Post-ASR Emotion Recognition with Large Language Models
von: Stepachev, Pavel, et al.
Veröffentlicht: (2024)
von: Stepachev, Pavel, et al.
Veröffentlicht: (2024)
MoHAVE: Mixture of Hierarchical Audio-Visual Experts for Robust Speech Recognition
von: Kim, Sungnyun, et al.
Veröffentlicht: (2025)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2025)
Investigating the Impact of Word Informativeness on Speech Emotion Recognition
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
von: Kakouros, Sofoklis
Veröffentlicht: (2025)
Robust Audiovisual Speech Recognition Models with Mixture-of-Experts
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
von: Wu, Yihan, et al.
Veröffentlicht: (2024)
Lamer-SSL: Layer-aware Mixture of LoRA Experts for Continual Multilingual Expansion of Self-supervised Models without Forgetting
von: Xu, Jing, et al.
Veröffentlicht: (2026)
von: Xu, Jing, et al.
Veröffentlicht: (2026)
Vesper: A Compact and Effective Pretrained Model for Speech Emotion Recognition
von: Chen, Weidong, et al.
Veröffentlicht: (2023)
von: Chen, Weidong, et al.
Veröffentlicht: (2023)
CO-VADA: A Confidence-Oriented Voice Augmentation Debiasing Approach for Fair Speech Emotion Recognition
von: Tsai, Yun-Shao, et al.
Veröffentlicht: (2025)
von: Tsai, Yun-Shao, et al.
Veröffentlicht: (2025)
EMO-Debias: Benchmarking Gender Debiasing Techniques in Multi-Label Speech Emotion Recognition
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2025)
von: Lin, Yi-Cheng, et al.
Veröffentlicht: (2025)
On the Contribution of Lexical Features to Speech Emotion Recognition
von: Combei, David
Veröffentlicht: (2025)
von: Combei, David
Veröffentlicht: (2025)
SHNU Multilingual Conversational Speech Recognition System for INTERSPEECH 2025 MLC-SLM Challenge
von: Mei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Mei, Yuxiang, et al.
Veröffentlicht: (2025)
Benchmarking and Confidence Evaluation of LALMs For Temporal Reasoning
von: Bhattacharya, Debarpan, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Debarpan, et al.
Veröffentlicht: (2025)
STAB: Speech Tokenizer Assessment Benchmark
von: Vashishth, Shikhar, et al.
Veröffentlicht: (2024)
von: Vashishth, Shikhar, et al.
Veröffentlicht: (2024)
Generating Data with Text-to-Speech and Large-Language Models for Conversational Speech Recognition
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
von: Cornell, Samuele, et al.
Veröffentlicht: (2024)
Meta-PerSER: Few-Shot Listener Personalized Speech Emotion Recognition via Meta-learning
von: Shen, Liang-Yeh, et al.
Veröffentlicht: (2025)
von: Shen, Liang-Yeh, et al.
Veröffentlicht: (2025)
BERSting at the Screams: A Benchmark for Distanced, Emotional and Shouted Speech Recognition
von: Tuttösí, Paige, et al.
Veröffentlicht: (2025)
von: Tuttösí, Paige, et al.
Veröffentlicht: (2025)
An Effective Mixture-Of-Experts Approach For Code-Switching Speech Recognition Leveraging Encoder Disentanglement
von: Yang, Tzu-Ting, et al.
Veröffentlicht: (2024)
von: Yang, Tzu-Ting, et al.
Veröffentlicht: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
Speaker-Aware Simulation Improves Conversational Speech Recognition
von: Gedeon, Máté, et al.
Veröffentlicht: (2026)
von: Gedeon, Máté, et al.
Veröffentlicht: (2026)
Steering Language Model to Stable Speech Emotion Recognition via Contextual Perception and Chain of Thought
von: Zhao, Zhixian, et al.
Veröffentlicht: (2025)
von: Zhao, Zhixian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2023) -
LLM supervised Pre-training for Multimodal Emotion Recognition in Conversations
von: Dutta, Soumya, et al.
Veröffentlicht: (2025) -
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
von: Dutta, Soumya, et al.
Veröffentlicht: (2025) -
Leveraging Content and Acoustic Representations for Speech Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2024) -
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)