Better Spanish Emotion Recognition In-the-wild: Bringing Attention to Deep Spectrum Voice Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ortega-Beltrán, Elena, Cabacas-Maso, Josep, Benito-Altamirano, Ismael, Ventura, Carles |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Facial Expression Recognition through Dual-Direction Attention Mixed Feature Networks: Application to 7th ABAW Challenge
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2024)
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2024)
Enhancing Facial Expression Recognition through Dual-Direction Attention Mixed Feature Networks and CLIP: Application to 8th ABAW Challenge
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2025)
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2025)
HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
von: Dutta, Soumya, et al.
Veröffentlicht: (2023)
GSDNet: Revisiting Incomplete Multimodal-Diffusion from Graph Spectrum Perspective for Conversation Emotion Recognition
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
von: Shou, Yuntao, et al.
Veröffentlicht: (2025)
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)
Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2025)
von: Zhao, Ruoyu, et al.
Veröffentlicht: (2025)
Emotion-Anchored Contrastive Learning Framework for Emotion Recognition in Conversation
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
DeepDialogue: A Multi-Turn Emotionally-Rich Spoken Dialogue Dataset
von: Koudounas, Alkis, et al.
Veröffentlicht: (2025)
von: Koudounas, Alkis, et al.
Veröffentlicht: (2025)
Distribution-based Emotion Recognition in Conversation
von: Wu, Wen, et al.
Veröffentlicht: (2022)
von: Wu, Wen, et al.
Veröffentlicht: (2022)
Emotional Dimension Control in Language Model-Based Text-to-Speech: Spanning a Broad Spectrum of Human Emotions
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
von: Zhou, Kun, et al.
Veröffentlicht: (2024)
On the Contribution of Lexical Features to Speech Emotion Recognition
von: Combei, David
Veröffentlicht: (2025)
von: Combei, David
Veröffentlicht: (2025)
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
Are Paralinguistic Representations all that is needed for Speech Emotion Recognition?
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2024)
ArabEmoNet: A Lightweight Hybrid 2D CNN-BiLSTM Model with Attention for Robust Arabic Speech Emotion Recognition
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
von: Abouzeid, Ali, et al.
Veröffentlicht: (2025)
Attention Is Not Always the Answer: Optimizing Voice Activity Detection with Simple Feature Fusion
von: Tripathi, Kumud, et al.
Veröffentlicht: (2025)
von: Tripathi, Kumud, et al.
Veröffentlicht: (2025)
Re-Parameterization of Lightweight Transformer for On-Device Speech Emotion Recognition
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
von: Zhang, Zixing, et al.
Veröffentlicht: (2024)
Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India
von: Bhogale, Kaushal, et al.
Veröffentlicht: (2026)
von: Bhogale, Kaushal, et al.
Veröffentlicht: (2026)
SEF-VC: Speaker Embedding Free Zero-Shot Voice Conversion with Cross Attention
von: Li, Junjie, et al.
Veröffentlicht: (2023)
von: Li, Junjie, et al.
Veröffentlicht: (2023)
Vesper: A Compact and Effective Pretrained Model for Speech Emotion Recognition
von: Chen, Weidong, et al.
Veröffentlicht: (2023)
von: Chen, Weidong, et al.
Veröffentlicht: (2023)
Improving Speech-based Emotion Recognition with Contextual Utterance Analysis and LLMs
von: Zhang, Enshi, et al.
Veröffentlicht: (2024)
von: Zhang, Enshi, et al.
Veröffentlicht: (2024)
BERSting at the Screams: A Benchmark for Distanced, Emotional and Shouted Speech Recognition
von: Tuttösí, Paige, et al.
Veröffentlicht: (2025)
von: Tuttösí, Paige, et al.
Veröffentlicht: (2025)
Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
von: Li, Chin-Jou, et al.
Veröffentlicht: (2025)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
von: Ma, Yong, et al.
Veröffentlicht: (2025)
von: Ma, Yong, et al.
Veröffentlicht: (2025)
Amplifying Emotional Signals: Data-Efficient Deep Learning for Robust Speech Emotion Recognition
von: Vu, Tai
Veröffentlicht: (2025)
von: Vu, Tai
Veröffentlicht: (2025)
GatedxLSTM: A Multimodal Affective Computing Approach for Emotion Recognition in Conversations
von: Li, Yupei, et al.
Veröffentlicht: (2025)
von: Li, Yupei, et al.
Veröffentlicht: (2025)
Multi-Teacher Language-Aware Knowledge Distillation for Multilingual Speech Emotion Recognition
von: Bijoy, Mehedi Hasan, et al.
Veröffentlicht: (2025)
von: Bijoy, Mehedi Hasan, et al.
Veröffentlicht: (2025)
Enhancing Multimodal Emotion Recognition through Multi-Granularity Cross-Modal Alignment
von: Wang, Xuechen, et al.
Veröffentlicht: (2024)
von: Wang, Xuechen, et al.
Veröffentlicht: (2024)
JELLY: Joint Emotion Recognition and Context Reasoning with LLMs for Conversational Speech Synthesis
von: Cha, Jun-Hyeok, et al.
Veröffentlicht: (2025)
von: Cha, Jun-Hyeok, et al.
Veröffentlicht: (2025)
Recursive Joint Cross-Modal Attention for Multimodal Fusion in Dimensional Emotion Recognition
von: Praveen, R. Gnana, et al.
Veröffentlicht: (2024)
von: Praveen, R. Gnana, et al.
Veröffentlicht: (2024)
Bimodal Connection Attention Fusion for Speech Emotion Recognition
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
von: Luo, Jiachen, et al.
Veröffentlicht: (2025)
A Cross-Corpus Speech Emotion Recognition Method Based on Supervised Contrastive Learning
von: minjie, Xiang
Veröffentlicht: (2024)
von: minjie, Xiang
Veröffentlicht: (2024)
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
von: Shi, Xiaohan, et al.
Veröffentlicht: (2023)
von: Shi, Xiaohan, et al.
Veröffentlicht: (2023)
Estimating the Uncertainty in Emotion Attributes using Deep Evidential Regression
von: Wu, Wen, et al.
Veröffentlicht: (2023)
von: Wu, Wen, et al.
Veröffentlicht: (2023)
Addressing Emotion Bias in Music Emotion Recognition and Generation with Frechet Audio Distance
von: Li, Yuanchao, et al.
Veröffentlicht: (2024)
von: Li, Yuanchao, et al.
Veröffentlicht: (2024)
MFLA: Monotonic Finite Look-ahead Attention for Streaming Speech Recognition
von: Xia, Yinfeng, et al.
Veröffentlicht: (2025)
von: Xia, Yinfeng, et al.
Veröffentlicht: (2025)
Steering Language Model to Stable Speech Emotion Recognition via Contextual Perception and Chain of Thought
von: Zhao, Zhixian, et al.
Veröffentlicht: (2025)
von: Zhao, Zhixian, et al.
Veröffentlicht: (2025)
MFSN: Multi-perspective Fusion Search Network For Pre-training Knowledge in Speech Emotion Recognition
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
von: Sun, Haiyang, et al.
Veröffentlicht: (2023)
MultiMed: Multilingual Medical Speech Recognition via Attention Encoder Decoder
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
Towards Energy-Efficient and Low-Latency Voice-Controlled Smart Homes: A Proposal for Offline Speech Recognition and IoT Integration
von: Huang, Peng, et al.
Veröffentlicht: (2025)
von: Huang, Peng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Enhancing Facial Expression Recognition through Dual-Direction Attention Mixed Feature Networks: Application to 7th ABAW Challenge
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2024) -
Enhancing Facial Expression Recognition through Dual-Direction Attention Mixed Feature Networks and CLIP: Application to 8th ABAW Challenge
von: Cabacas-Maso, Josep, et al.
Veröffentlicht: (2025) -
HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition
von: Dutta, Soumya, et al.
Veröffentlicht: (2023) -
GSDNet: Revisiting Incomplete Multimodal-Diffusion from Graph Spectrum Perspective for Conversation Emotion Recognition
von: Shou, Yuntao, et al.
Veröffentlicht: (2025) -
The Third VoicePrivacy Challenge: Preserving Emotional Expressiveness and Linguistic Content in Voice Anonymization
von: Tomashenko, Natalia, et al.
Veröffentlicht: (2026)