M4SER: Multimodal, Multirepresentation, Multitask, and Multistrategy Learning for Speech Emotion Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | He, Jiajun, Shi, Xiaohan, Hu, Cheng-Hung, Mi, Jinyi, Li, Xingfeng, Toda, Tomoki |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MF-AED-AEC: Speech Emotion Recognition by Leveraging Multimodal Fusion, Asr Error Detection, and Asr Error Correction
di: He, Jiajun, et al.
Pubblicazione: (2024)
di: He, Jiajun, et al.
Pubblicazione: (2024)
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
di: Shi, Xiaohan, et al.
Pubblicazione: (2023)
di: Shi, Xiaohan, et al.
Pubblicazione: (2023)
GIA-MIC: Multimodal Emotion Recognition with Gated Interactive Attention and Modality-Invariant Learning Constraints
di: He, Jiajun, et al.
Pubblicazione: (2025)
di: He, Jiajun, et al.
Pubblicazione: (2025)
MuMTAffect: A Multimodal Multitask Affective Framework for Personality and Emotion Recognition from Physiological Signals
di: Seikavandi, Meisam Jamshidi, et al.
Pubblicazione: (2025)
di: Seikavandi, Meisam Jamshidi, et al.
Pubblicazione: (2025)
Two-stage Framework for Robust Speech Emotion Recognition Using Target Speaker Extraction in Human Speech Noise Conditions
di: Mi, Jinyi, et al.
Pubblicazione: (2024)
di: Mi, Jinyi, et al.
Pubblicazione: (2024)
Explainable Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2023)
di: Lian, Zheng, et al.
Pubblicazione: (2023)
AffectGPT-R1: Leveraging Reinforcement Learning for Open-Vocabulary Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2025)
di: Lian, Zheng, et al.
Pubblicazione: (2025)
Hardware-Aware Federated Learning for Speech Emotion Recognition
di: Yuksel, Beyazit Bestami, et al.
Pubblicazione: (2026)
di: Yuksel, Beyazit Bestami, et al.
Pubblicazione: (2026)
OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2024)
di: Lian, Zheng, et al.
Pubblicazione: (2024)
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2024)
di: Lian, Zheng, et al.
Pubblicazione: (2024)
MERBench: A Unified Evaluation Benchmark for Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2024)
di: Lian, Zheng, et al.
Pubblicazione: (2024)
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
Speech Command + Speech Emotion: Exploring Emotional Speech Commands as a Compound and Playful Modality
di: Aslan, Ilhan, et al.
Pubblicazione: (2025)
di: Aslan, Ilhan, et al.
Pubblicazione: (2025)
Hierarchical MoE: Continuous Multimodal Emotion Recognition with Incomplete and Asynchronous Inputs
di: Zhu, Yitong, et al.
Pubblicazione: (2025)
di: Zhu, Yitong, et al.
Pubblicazione: (2025)
Automatic design optimization of preference-based subjective evaluation with online learning in crowdsourcing environment
di: Yasuda, Yusuke, et al.
Pubblicazione: (2024)
di: Yasuda, Yusuke, et al.
Pubblicazione: (2024)
FIRMED: A Peak-Centered Multimodal Dataset with Fine-Grained Annotation for Emotion Recognition
di: Tang, Hao, et al.
Pubblicazione: (2025)
di: Tang, Hao, et al.
Pubblicazione: (2025)
MER 2026: From Discriminative Emotion Recognition to Generative Emotion Understanding
di: Lian, Zheng, et al.
Pubblicazione: (2026)
di: Lian, Zheng, et al.
Pubblicazione: (2026)
Pioneering Multimodal Emotion Recognition in the Era of Large Models: From Closed Sets to Open Vocabularies
di: Han, Jing, et al.
Pubblicazione: (2025)
di: Han, Jing, et al.
Pubblicazione: (2025)
AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2026)
di: Lian, Zheng, et al.
Pubblicazione: (2026)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2024)
di: Lian, Zheng, et al.
Pubblicazione: (2024)
Hypergraph Multi-Modal Learning for EEG-based Emotion Recognition in Conversation
di: Kang, Zijian, et al.
Pubblicazione: (2025)
di: Kang, Zijian, et al.
Pubblicazione: (2025)
Follow the Clues, Frame the Truth: Hybrid-evidential Deductive Reasoning in Open-Vocabulary Multimodal Emotion Recognition
di: Liu, Yu, et al.
Pubblicazione: (2026)
di: Liu, Yu, et al.
Pubblicazione: (2026)
"How to Explore Biases in Speech Emotion AI with Users?" A Speech-Emotion-Acting Study Exploring Age and Language Biases
di: Borre, Josephine Beatrice Skovbo, et al.
Pubblicazione: (2025)
di: Borre, Josephine Beatrice Skovbo, et al.
Pubblicazione: (2025)
Multimodal Functional Maximum Correlation for Emotion Recognition
di: Zheng, Deyang, et al.
Pubblicazione: (2025)
di: Zheng, Deyang, et al.
Pubblicazione: (2025)
Towards LLM-Empowered Fine-Grained Speech Descriptors for Explainable Emotion Recognition
di: Chen, Youjun, et al.
Pubblicazione: (2025)
di: Chen, Youjun, et al.
Pubblicazione: (2025)
UMind: A Unified Multitask Network for Zero-Shot M/EEG Visual Decoding
di: Xu, Chengjian, et al.
Pubblicazione: (2025)
di: Xu, Chengjian, et al.
Pubblicazione: (2025)
MIND-EEG: Multi-granularity Integration Network with Discrete Codebook for EEG-based Emotion Recognition
di: Zhang, Yuzhe, et al.
Pubblicazione: (2025)
di: Zhang, Yuzhe, et al.
Pubblicazione: (2025)
Exploring the Impact of Emotional Voice Integration in Sign-to-Speech Translators for Deaf-to-Hearing Communication
di: Lim, Hyunchul, et al.
Pubblicazione: (2024)
di: Lim, Hyunchul, et al.
Pubblicazione: (2024)
EEG-based Multimodal Representation Learning for Emotion Recognition
di: Yin, Kang, et al.
Pubblicazione: (2024)
di: Yin, Kang, et al.
Pubblicazione: (2024)
EVA-MED: An Enhanced Valence-Arousal Multimodal Emotion Dataset for Emotion Recognition
di: Huang, Xin, et al.
Pubblicazione: (2025)
di: Huang, Xin, et al.
Pubblicazione: (2025)
Relational Co-Adaptation in Emotionally Supportive AI: Tensions in Authentic Emotional Interaction
di: Shi, Mengqi
Pubblicazione: (2026)
di: Shi, Mengqi
Pubblicazione: (2026)
EmotionCarrier: A Multimodality 'Mindfulness-Training' Tool for Positive Emotional Value
di: Wang, Yi
Pubblicazione: (2025)
di: Wang, Yi
Pubblicazione: (2025)
Feel my Speech: Automatic Speech Emotion Conversion for Tangible, Haptic, or Proxemic Interaction Design
di: Aslan, Ilhan
Pubblicazione: (2024)
di: Aslan, Ilhan
Pubblicazione: (2024)
Improving Inclusivity for Emotion Recognition Based on Face Tracking
di: Ellenberg, Mats Ole, et al.
Pubblicazione: (2025)
di: Ellenberg, Mats Ole, et al.
Pubblicazione: (2025)
Human-AI Alignment of Multimodal Large Language Models with Speech-Language Pathologists in Parent-Child Interactions
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
FEEL: Quantifying Heterogeneity in Physiological Signals for Generalizable Emotion Recognition
di: Singh, Pragya, et al.
Pubblicazione: (2026)
di: Singh, Pragya, et al.
Pubblicazione: (2026)
Knowledge-based Emotion Recognition using Large Language Models
di: Han, Bin, et al.
Pubblicazione: (2024)
di: Han, Bin, et al.
Pubblicazione: (2024)
Challenges in Automatic Speech Recognition for Adults with Cognitive Impairment
di: Cohn, Michelle, et al.
Pubblicazione: (2026)
di: Cohn, Michelle, et al.
Pubblicazione: (2026)
Gaze-Hand Steering for Travel and Multitasking in Virtual Environments
di: Zavichi, Mona, et al.
Pubblicazione: (2025)
di: Zavichi, Mona, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MF-AED-AEC: Speech Emotion Recognition by Leveraging Multimodal Fusion, Asr Error Detection, and Asr Error Correction
di: He, Jiajun, et al.
Pubblicazione: (2024) -
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
di: Shi, Xiaohan, et al.
Pubblicazione: (2023) -
GIA-MIC: Multimodal Emotion Recognition with Gated Interactive Attention and Modality-Invariant Learning Constraints
di: He, Jiajun, et al.
Pubblicazione: (2025) -
MuMTAffect: A Multimodal Multitask Affective Framework for Personality and Emotion Recognition from Physiological Signals
di: Seikavandi, Meisam Jamshidi, et al.
Pubblicazione: (2025) -
Two-stage Framework for Robust Speech Emotion Recognition Using Target Speaker Extraction in Human Speech Noise Conditions
di: Mi, Jinyi, et al.
Pubblicazione: (2024)