Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hizabri, Marouane El, Bezzaz, Abdelfattah, Hayoukane, Ismail, Taki, Youssef |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2024)
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2024)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023)
A Layer-Anchoring Strategy for Enhancing Cross-Lingual Speech Emotion Recognition
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
Test-Time Adaptation for Speech Emotion Recognition
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026)
Adapting WavLM for Speech Emotion Recognition
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
von: Lu, Cheng, et al.
Veröffentlicht: (2024)
Parameter Efficient Finetuning for Speech Emotion Recognition and Domain Adaptation
von: Lashkarashvili, Nineli, et al.
Veröffentlicht: (2024)
von: Lashkarashvili, Nineli, et al.
Veröffentlicht: (2024)
Recovering Performance in Speech Emotion Recognition from Discrete Tokens via Multi-Layer Fusion and Paralinguistic Feature Integration
von: Sun, Esther, et al.
Veröffentlicht: (2026)
von: Sun, Esther, et al.
Veröffentlicht: (2026)
HuLA: Prosody-Aware Anti-Spoofing with Multi-Task Learning for Expressive and Emotional Synthetic Speech
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2025)
TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition
von: Chen, Chengxin, et al.
Veröffentlicht: (2024)
von: Chen, Chengxin, et al.
Veröffentlicht: (2024)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2022)
Mouth Articulation-Based Anchoring for Improved Cross-Corpus Speech Emotion Recognition
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024)
Behind the Scenes: Mechanistic Interpretability of LoRA-adapted Whisper for Speech Emotion Recognition
von: Ma, Yujian, et al.
Veröffentlicht: (2025)
von: Ma, Yujian, et al.
Veröffentlicht: (2025)
More Similar than Dissimilar: Modeling Annotators for Cross-Corpus Speech Emotion Recognition
von: Tavernor, James, et al.
Veröffentlicht: (2025)
von: Tavernor, James, et al.
Veröffentlicht: (2025)
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
von: Jia, Zijun, et al.
Veröffentlicht: (2025)
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2024)
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2024)
Human Feedback Driven Dynamic Speech Emotion Recognition
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
von: Fedorov, Ilya, et al.
Veröffentlicht: (2025)
Multi-blank Transducers for Speech Recognition
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
von: Xu, Hainan, et al.
Veröffentlicht: (2022)
Comparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
von: Agbo, Idoko, et al.
Veröffentlicht: (2024)
Can Emotion Fool Anti-spoofing?
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
von: Mahapatra, Aurosweta, et al.
Veröffentlicht: (2025)
Pre-Finetuning for Few-Shot Emotional Speech Recognition
von: Chen, Maximillian, et al.
Veröffentlicht: (2023)
von: Chen, Maximillian, et al.
Veröffentlicht: (2023)
Speech Slytherin: Examining the Performance and Efficiency of Mamba for Speech Separation, Recognition, and Synthesis
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
von: Jiang, Xilin, et al.
Veröffentlicht: (2024)
Drax: Speech Recognition with Discrete Flow Matching
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
Keyword-Guided Adaptation of Automatic Speech Recognition
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
von: Romana, Amrit, et al.
Veröffentlicht: (2025)
von: Romana, Amrit, et al.
Veröffentlicht: (2025)
Speech Emotion Detection Based on MFCC and CNN-LSTM Architecture
von: Ouyang, Qianhe
Veröffentlicht: (2025)
von: Ouyang, Qianhe
Veröffentlicht: (2025)
Multi-Microphone and Multi-Modal Emotion Recognition in Reverberant Environment
von: Cohen, Ohad, et al.
Veröffentlicht: (2024)
von: Cohen, Ohad, et al.
Veröffentlicht: (2024)
Enhanced Speech Emotion Recognition with Efficient Channel Attention Guided Deep CNN-BiLSTM Framework
von: Kundu, Niloy Kumar, et al.
Veröffentlicht: (2024)
von: Kundu, Niloy Kumar, et al.
Veröffentlicht: (2024)
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
von: Nanmalar, M., et al.
Veröffentlicht: (2024)
von: Nanmalar, M., et al.
Veröffentlicht: (2024)
Memory-guided Prototypical Co-occurrence Learning for Mixed Emotion Recognition
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
A Novel Hybrid Deep Learning Technique for Speech Emotion Detection using Feature Engineering
von: Chowdhury, Shahana Yasmin, et al.
Veröffentlicht: (2025)
von: Chowdhury, Shahana Yasmin, et al.
Veröffentlicht: (2025)
SELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios
von: Bukhari, Hazim, et al.
Veröffentlicht: (2024)
von: Bukhari, Hazim, et al.
Veröffentlicht: (2024)
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
von: Ravenscroft, William, et al.
Veröffentlicht: (2024)
Regularizing Learnable Feature Extraction for Automatic Speech Recognition
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
von: Vieting, Peter, et al.
Veröffentlicht: (2025)
Underwater Acoustic Signal Recognition Based on Salient Feature
von: Chen, Minghao
Veröffentlicht: (2023)
von: Chen, Minghao
Veröffentlicht: (2023)
Training Universal Vocoders with Feature Smoothing-Based Augmentation Methods for High-Quality TTS Systems
von: Liu, Jeongmin, et al.
Veröffentlicht: (2024)
von: Liu, Jeongmin, et al.
Veröffentlicht: (2024)
Multimodal Attention Merging for Improved Speech Recognition and Audio Event Classification
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
von: Sundar, Anirudh S., et al.
Veröffentlicht: (2023)
Adapting Automatic Speech Recognition for Accented Air Traffic Control Communications
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
von: Wee, Marcus Yu Zhe, et al.
Veröffentlicht: (2025)
Chunked Attention-based Encoder-Decoder Model for Streaming Speech Recognition
von: Zeineldeen, Mohammad, et al.
Veröffentlicht: (2023)
von: Zeineldeen, Mohammad, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
von: Ulgen, Ismail Rasim, et al.
Veröffentlicht: (2024) -
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
von: Rajapakshe, Thejan, et al.
Veröffentlicht: (2023) -
A Layer-Anchoring Strategy for Enhancing Cross-Lingual Speech Emotion Recognition
von: Upadhyay, Shreya G., et al.
Veröffentlicht: (2024) -
Test-Time Adaptation for Speech Emotion Recognition
von: Dong, Jiaheng, et al.
Veröffentlicht: (2026) -
Adapting WavLM for Speech Emotion Recognition
von: Diatlova, Daria, et al.
Veröffentlicht: (2024)