Improving Speech Emotion Recognition with Mutual Information Regularized Generative Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Ahn, Chung-Soo, Rana, Rajib, Sivadas, Sunil, Busso, Carlos, Rajapakse, Jagath C. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2022)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2022)
Mouth Articulation-Based Anchoring for Improved Cross-Corpus Speech Emotion Recognition
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2025)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2025)
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
di: Ulgen, Ismail Rasim, et al.
Pubblicazione: (2024)
di: Ulgen, Ismail Rasim, et al.
Pubblicazione: (2024)
A Layer-Anchoring Strategy for Enhancing Cross-Lingual Speech Emotion Recognition
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
Recovering Performance in Speech Emotion Recognition from Discrete Tokens via Multi-Layer Fusion and Paralinguistic Feature Integration
di: Sun, Esther, et al.
Pubblicazione: (2026)
di: Sun, Esther, et al.
Pubblicazione: (2026)
Describe Where You Are: Improving Noise-Robustness for Speech Emotion Recognition with Text Description of the Environment
di: Leem, Seong-Gyun, et al.
Pubblicazione: (2024)
di: Leem, Seong-Gyun, et al.
Pubblicazione: (2024)
Speech Emotion Recognition with Phonation Excitation Information and Articulatory Kinematics
di: Zhang, Ziqian, et al.
Pubblicazione: (2025)
di: Zhang, Ziqian, et al.
Pubblicazione: (2025)
EmoHRNet: High-Resolution Neural Network Based Speech Emotion Recognition
di: Muppidi, Akshay, et al.
Pubblicazione: (2025)
di: Muppidi, Akshay, et al.
Pubblicazione: (2025)
Emotion-Disentangled Embedding Alignment for Noise-Robust and Cross-Corpus Speech Emotion Recognition
di: Tiwari, Upasana, et al.
Pubblicazione: (2025)
di: Tiwari, Upasana, et al.
Pubblicazione: (2025)
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 2025
di: Ferreira, Alef Iury Siqueira, et al.
Pubblicazione: (2025)
di: Ferreira, Alef Iury Siqueira, et al.
Pubblicazione: (2025)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
di: Lu, Cheng, et al.
Pubblicazione: (2024)
di: Lu, Cheng, et al.
Pubblicazione: (2024)
Parameter Efficient Finetuning for Speech Emotion Recognition and Domain Adaptation
di: Lashkarashvili, Nineli, et al.
Pubblicazione: (2024)
di: Lashkarashvili, Nineli, et al.
Pubblicazione: (2024)
Test-Time Adaptation for Speech Emotion Recognition
di: Dong, Jiaheng, et al.
Pubblicazione: (2026)
di: Dong, Jiaheng, et al.
Pubblicazione: (2026)
Adapting WavLM for Speech Emotion Recognition
di: Diatlova, Daria, et al.
Pubblicazione: (2024)
di: Diatlova, Daria, et al.
Pubblicazione: (2024)
Can Emotion Fool Anti-spoofing?
di: Mahapatra, Aurosweta, et al.
Pubblicazione: (2025)
di: Mahapatra, Aurosweta, et al.
Pubblicazione: (2025)
SpectroFusion-ViT: A Lightweight Transformer for Speech Emotion Recognition Using Harmonic Mel-Chroma Fusion
di: Ahmed, Faria, et al.
Pubblicazione: (2026)
di: Ahmed, Faria, et al.
Pubblicazione: (2026)
More Similar than Dissimilar: Modeling Annotators for Cross-Corpus Speech Emotion Recognition
di: Tavernor, James, et al.
Pubblicazione: (2025)
di: Tavernor, James, et al.
Pubblicazione: (2025)
Koopman Regularized Deep Speech Disentanglement for Speaker Verification
di: Chazaridis, Nikos, et al.
Pubblicazione: (2026)
di: Chazaridis, Nikos, et al.
Pubblicazione: (2026)
Pharmacophore-guided de novo drug design with diffusion bridge
di: Wang, Conghao, et al.
Pubblicazione: (2024)
di: Wang, Conghao, et al.
Pubblicazione: (2024)
TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition
di: Chen, Chengxin, et al.
Pubblicazione: (2024)
di: Chen, Chengxin, et al.
Pubblicazione: (2024)
Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
di: Hizabri, Marouane El, et al.
Pubblicazione: (2026)
di: Hizabri, Marouane El, et al.
Pubblicazione: (2026)
Large Language Model Data Generation for Enhanced Intent Recognition in German Speech
di: Rosin, Theresa Pekarek, et al.
Pubblicazione: (2025)
di: Rosin, Theresa Pekarek, et al.
Pubblicazione: (2025)
A Study on the Data Distribution Gap in Music Emotion Recognition
di: Ching, Joann, et al.
Pubblicazione: (2025)
di: Ching, Joann, et al.
Pubblicazione: (2025)
Regularizing Learnable Feature Extraction for Automatic Speech Recognition
di: Vieting, Peter, et al.
Pubblicazione: (2025)
di: Vieting, Peter, et al.
Pubblicazione: (2025)
Pre-Finetuning for Few-Shot Emotional Speech Recognition
di: Chen, Maximillian, et al.
Pubblicazione: (2023)
di: Chen, Maximillian, et al.
Pubblicazione: (2023)
Task Vector in TTS: Toward Emotionally Expressive Dialectal Speech Synthesis
di: Feng, Pengchao, et al.
Pubblicazione: (2025)
di: Feng, Pengchao, et al.
Pubblicazione: (2025)
A Systematic Evaluation of Adversarial Attacks against Speech Emotion Recognition Models
di: Facchinetti, Nicolas, et al.
Pubblicazione: (2024)
di: Facchinetti, Nicolas, et al.
Pubblicazione: (2024)
Behind the Scenes: Mechanistic Interpretability of LoRA-adapted Whisper for Speech Emotion Recognition
di: Ma, Yujian, et al.
Pubblicazione: (2025)
di: Ma, Yujian, et al.
Pubblicazione: (2025)
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets
di: Jia, Zijun, et al.
Pubblicazione: (2025)
di: Jia, Zijun, et al.
Pubblicazione: (2025)
Scaling Ambiguity: Augmenting Human Annotation in Speech Emotion Recognition with Audio-Language Models
di: Zhang, Wenda, et al.
Pubblicazione: (2026)
di: Zhang, Wenda, et al.
Pubblicazione: (2026)
Multimodal Attention Merging for Improved Speech Recognition and Audio Event Classification
di: Sundar, Anirudh S., et al.
Pubblicazione: (2023)
di: Sundar, Anirudh S., et al.
Pubblicazione: (2023)
Human Feedback Driven Dynamic Speech Emotion Recognition
di: Fedorov, Ilya, et al.
Pubblicazione: (2025)
di: Fedorov, Ilya, et al.
Pubblicazione: (2025)
NaturalVoices: A Large-Scale, Spontaneous and Emotional Podcast Dataset for Voice Conversion
di: Du, Zongyang, et al.
Pubblicazione: (2025)
di: Du, Zongyang, et al.
Pubblicazione: (2025)
EmoAugNet: A Signal-Augmented Hybrid CNN-LSTM Framework for Speech Emotion Recognition
di: Paul, Durjoy Chandra, et al.
Pubblicazione: (2025)
di: Paul, Durjoy Chandra, et al.
Pubblicazione: (2025)
Speech Emotion Recognition Leveraging OpenAI's Whisper Representations and Attentive Pooling Methods
di: Shendabadi, Ali, et al.
Pubblicazione: (2026)
di: Shendabadi, Ali, et al.
Pubblicazione: (2026)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
di: Xu, Ji, et al.
Pubblicazione: (2023)
di: Xu, Ji, et al.
Pubblicazione: (2023)
Can Synthetic Audio From Generative Foundation Models Assist Audio Recognition and Speech Modeling?
di: Feng, Tiantian, et al.
Pubblicazione: (2024)
di: Feng, Tiantian, et al.
Pubblicazione: (2024)
Documenti analoghi
-
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024) -
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023) -
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2022) -
Mouth Articulation-Based Anchoring for Improved Cross-Corpus Speech Emotion Recognition
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024) -
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2025)