Speech Emotion Detection Based on MFCC and CNN-LSTM Architecture
Fuente:
arXiv
Salvato in:
| Autore principale: | Ouyang, Qianhe |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model
di: Oluwademilade, Adelekun, et al.
Pubblicazione: (2026)
di: Oluwademilade, Adelekun, et al.
Pubblicazione: (2026)
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024)
Optimising MFCC parameters for the automatic detection of respiratory diseases
di: Yan, Yuyang, et al.
Pubblicazione: (2024)
di: Yan, Yuyang, et al.
Pubblicazione: (2024)
Enhanced Speech Emotion Recognition with Efficient Channel Attention Guided Deep CNN-BiLSTM Framework
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
di: Manukalpa, J. M. Chan Sri, et al.
Pubblicazione: (2025)
Mouth Articulation-Based Anchoring for Improved Cross-Corpus Speech Emotion Recognition
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
Test-Time Adaptation for Speech Emotion Recognition
di: Dong, Jiaheng, et al.
Pubblicazione: (2026)
di: Dong, Jiaheng, et al.
Pubblicazione: (2026)
Adapting WavLM for Speech Emotion Recognition
di: Diatlova, Daria, et al.
Pubblicazione: (2024)
di: Diatlova, Daria, et al.
Pubblicazione: (2024)
Speech Emotion Recognition Using CNN and Its Use Case in Digital Healthcare
di: Nigar, Nishargo
Pubblicazione: (2024)
di: Nigar, Nishargo
Pubblicazione: (2024)
Comparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech
di: Agbo, Idoko, et al.
Pubblicazione: (2024)
di: Agbo, Idoko, et al.
Pubblicazione: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
Revealing Emotional Clusters in Speaker Embeddings: A Contrastive Learning Strategy for Speech Emotion Recognition
di: Ulgen, Ismail Rasim, et al.
Pubblicazione: (2024)
di: Ulgen, Ismail Rasim, et al.
Pubblicazione: (2024)
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
di: Romana, Amrit, et al.
Pubblicazione: (2025)
di: Romana, Amrit, et al.
Pubblicazione: (2025)
Parameter Efficient Finetuning for Speech Emotion Recognition and Domain Adaptation
di: Lashkarashvili, Nineli, et al.
Pubblicazione: (2024)
di: Lashkarashvili, Nineli, et al.
Pubblicazione: (2024)
TRNet: Two-level Refinement Network leveraging Speech Enhancement for Noise Robust Speech Emotion Recognition
di: Chen, Chengxin, et al.
Pubblicazione: (2024)
di: Chen, Chengxin, et al.
Pubblicazione: (2024)
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
di: Nanmalar, M., et al.
Pubblicazione: (2024)
di: Nanmalar, M., et al.
Pubblicazione: (2024)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2025)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2025)
Noise-Robust Contrastive Learning with an MFCC-Conformer For Coronary Artery Disease Detection
di: Marocchi, Milan, et al.
Pubblicazione: (2026)
di: Marocchi, Milan, et al.
Pubblicazione: (2026)
An LSTM-Based Chord Generation System Using Chroma Histogram Representations
di: Hardwick, Jack
Pubblicazione: (2024)
di: Hardwick, Jack
Pubblicazione: (2024)
Impact of Speech Mode in Automatic Pathological Speech Detection
di: Sheikh, Shakeel A., et al.
Pubblicazione: (2024)
di: Sheikh, Shakeel A., et al.
Pubblicazione: (2024)
Enhancing Speech Emotion Recognition using Dynamic Spectral Features and Kalman Smoothing
di: Hizabri, Marouane El, et al.
Pubblicazione: (2026)
di: Hizabri, Marouane El, et al.
Pubblicazione: (2026)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2022)
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2022)
A Layer-Anchoring Strategy for Enhancing Cross-Lingual Speech Emotion Recognition
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
di: Upadhyay, Shreya G., et al.
Pubblicazione: (2024)
Behind the Scenes: Mechanistic Interpretability of LoRA-adapted Whisper for Speech Emotion Recognition
di: Ma, Yujian, et al.
Pubblicazione: (2025)
di: Ma, Yujian, et al.
Pubblicazione: (2025)
More Similar than Dissimilar: Modeling Annotators for Cross-Corpus Speech Emotion Recognition
di: Tavernor, James, et al.
Pubblicazione: (2025)
di: Tavernor, James, et al.
Pubblicazione: (2025)
Coverage-Guaranteed Speech Emotion Recognition via Calibrated Uncertainty-Adaptive Prediction Sets
di: Jia, Zijun, et al.
Pubblicazione: (2025)
di: Jia, Zijun, et al.
Pubblicazione: (2025)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
di: Lu, Cheng, et al.
Pubblicazione: (2024)
di: Lu, Cheng, et al.
Pubblicazione: (2024)
Reverse-Speech-Finder: A Neural Network Backtracking Architecture for Generating Alzheimer's Disease Speech Samples and Improving Diagnosis Performance
di: Li, Victor OK, et al.
Pubblicazione: (2025)
di: Li, Victor OK, et al.
Pubblicazione: (2025)
Recurrence-Based Nonlinear Vocal Dynamics as Digital Biomarkers for Depression Detection from Conversational Speech
di: Samanta, Himadri S
Pubblicazione: (2026)
di: Samanta, Himadri S
Pubblicazione: (2026)
Tracking Articulatory Dynamics in Speech with a Fixed-Weight BiLSTM-CNN Architecture
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
HuLA: Prosody-Aware Anti-Spoofing with Multi-Task Learning for Expressive and Emotional Synthetic Speech
di: Mahapatra, Aurosweta, et al.
Pubblicazione: (2025)
di: Mahapatra, Aurosweta, et al.
Pubblicazione: (2025)
Recovering Performance in Speech Emotion Recognition from Discrete Tokens via Multi-Layer Fusion and Paralinguistic Feature Integration
di: Sun, Esther, et al.
Pubblicazione: (2026)
di: Sun, Esther, et al.
Pubblicazione: (2026)
Few-Shot Speech Deepfake Detection Adaptation with Gaussian Processes
di: Glazer, Neta, et al.
Pubblicazione: (2025)
di: Glazer, Neta, et al.
Pubblicazione: (2025)
Multiview Canonical Correlation Analysis for Automatic Pathological Speech Detection
di: Kaloga, Yacouba, et al.
Pubblicazione: (2024)
di: Kaloga, Yacouba, et al.
Pubblicazione: (2024)
Multi-Channel Replay Speech Detection using Acoustic Maps
di: Neri, Michael, et al.
Pubblicazione: (2026)
di: Neri, Michael, et al.
Pubblicazione: (2026)
RawTFNet: A Lightweight CNN Architecture for Speech Anti-spoofing
di: Xiao, Yang, et al.
Pubblicazione: (2025)
di: Xiao, Yang, et al.
Pubblicazione: (2025)
Detection and Forecasting of Parkinson Disease Progression from Speech Signal Features Using MultiLayer Perceptron and LSTM
di: Ali, Majid, et al.
Pubblicazione: (2024)
di: Ali, Majid, et al.
Pubblicazione: (2024)
Exploring WavLM Back-ends for Speech Spoofing and Deepfake Detection
di: Stourbe, Theophile, et al.
Pubblicazione: (2024)
di: Stourbe, Theophile, et al.
Pubblicazione: (2024)
Phoneme-Level Deepfake Detection Across Emotional Conditions Using Self-Supervised Embeddings
di: Nallaguntla, Vamshi, et al.
Pubblicazione: (2026)
di: Nallaguntla, Vamshi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Speech Emotion Recognition Using MFCC Features and LSTM-Based Deep Learning Model
di: Oluwademilade, Adelekun, et al.
Pubblicazione: (2026) -
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2024) -
Optimising MFCC parameters for the automatic detection of respiratory diseases
di: Yan, Yuyang, et al.
Pubblicazione: (2024) -
Enhanced Speech Emotion Recognition with Efficient Channel Attention Guided Deep CNN-BiLSTM Framework
di: Kundu, Niloy Kumar, et al.
Pubblicazione: (2024) -
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
di: Rajapakshe, Thejan, et al.
Pubblicazione: (2023)