Are you sure? Analysing Uncertainty Quantification Approaches for Real-world Speech Emotion Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Schrüfer, Oliver, Milling, Manuel, Burkhardt, Felix, Eyben, Florian, Schuller, Björn |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023)
por: Derington, Anna, et al.
Publicado: (2023)
Using voice analysis as an early indicator of risk for depression in young adults
por: Scherer, Klaus R., et al.
Publicado: (2024)
por: Scherer, Klaus R., et al.
Publicado: (2024)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
por: Rajapakshe, Thejan, et al.
Publicado: (2022)
por: Rajapakshe, Thejan, et al.
Publicado: (2022)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
An automatic analysis of ultrasound vocalisations for the prediction of interaction context in captive Egyptian fruit bats
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
por: Zhao, Yan, et al.
Publicado: (2024)
por: Zhao, Yan, et al.
Publicado: (2024)
Enhancing Speech Emotion Recognition Through Differentiable Architecture Search
por: Rajapakshe, Thejan, et al.
Publicado: (2023)
por: Rajapakshe, Thejan, et al.
Publicado: (2023)
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
por: Qi, Tianhua, et al.
Publicado: (2026)
por: Qi, Tianhua, et al.
Publicado: (2026)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
por: Rajapakshe, Thejan, et al.
Publicado: (2025)
por: Rajapakshe, Thejan, et al.
Publicado: (2025)
Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers
por: Latif, Siddique, et al.
Publicado: (2023)
por: Latif, Siddique, et al.
Publicado: (2023)
Audio Enhancement for Computer Audition -- An Iterative Training Paradigm Using Sample Importance
por: Milling, Manuel, et al.
Publicado: (2024)
por: Milling, Manuel, et al.
Publicado: (2024)
Improving Speaker-independent Speech Emotion Recognition Using Dynamic Joint Distribution Adaptation
por: Lu, Cheng, et al.
Publicado: (2024)
por: Lu, Cheng, et al.
Publicado: (2024)
Breaking Resource Barriers in Speech Emotion Recognition via Data Distillation
por: Chang, Yi, et al.
Publicado: (2024)
por: Chang, Yi, et al.
Publicado: (2024)
autrainer: A Modular and Extensible Deep Learning Toolkit for Computer Audition Tasks
por: Rampp, Simon, et al.
Publicado: (2024)
por: Rampp, Simon, et al.
Publicado: (2024)
emoDARTS: Joint Optimisation of CNN & Sequential Neural Network Architectures for Superior Speech Emotion Recognition
por: Rajapakshe, Thejan, et al.
Publicado: (2024)
por: Rajapakshe, Thejan, et al.
Publicado: (2024)
Quantifying Dimensional Independence in Speech: An Information-Theoretic Framework for Disentangled Representation Learning
por: Kashyap, Bipasha, et al.
Publicado: (2026)
por: Kashyap, Bipasha, et al.
Publicado: (2026)
Bringing the Discussion of Minima Sharpness to the Audio Domain: a Filter-Normalised Evaluation for Acoustic Scene Classification
por: Milling, Manuel, et al.
Publicado: (2023)
por: Milling, Manuel, et al.
Publicado: (2023)
GatedxLSTM: A Multimodal Affective Computing Approach for Emotion Recognition in Conversations
por: Li, Yupei, et al.
Publicado: (2025)
por: Li, Yupei, et al.
Publicado: (2025)
Cross-Dialect Bird Species Recognition with Dialect-Calibrated Augmentation
por: Ding, Jiani, et al.
Publicado: (2025)
por: Ding, Jiani, et al.
Publicado: (2025)
Intelligent Cardiac Auscultation for Murmur Detection via Parallel-Attentive Models with Uncertainty Estimation
por: Zhang, Zixing, et al.
Publicado: (2024)
por: Zhang, Zixing, et al.
Publicado: (2024)
Abusive Speech Detection in Indic Languages Using Acoustic Features
por: Spiesberger, Anika A., et al.
Publicado: (2024)
por: Spiesberger, Anika A., et al.
Publicado: (2024)
Speech Emotion Recognition with ASR Integration
por: Li, Yuanchao
Publicado: (2026)
por: Li, Yuanchao
Publicado: (2026)
Emotion Neural Transducer for Fine-Grained Speech Emotion Recognition
por: Shen, Siyuan, et al.
Publicado: (2024)
por: Shen, Siyuan, et al.
Publicado: (2024)
EMO-SUPERB: An In-depth Look at Speech Emotion Recognition
por: Wu, Haibin, et al.
Publicado: (2024)
por: Wu, Haibin, et al.
Publicado: (2024)
Dataset-Distillation Generative Model for Speech Emotion Recognition
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2024)
por: Ritter-Gutierrez, Fabian, et al.
Publicado: (2024)
THAI Speech Emotion Recognition (THAI-SER) corpus
por: Wongpithayadisai, Jilamika, et al.
Publicado: (2025)
por: Wongpithayadisai, Jilamika, et al.
Publicado: (2025)
Iterative Prototype Refinement for Ambiguous Speech Emotion Recognition
por: Sun, Haoqin, et al.
Publicado: (2024)
por: Sun, Haoqin, et al.
Publicado: (2024)
Can you Remove the Downstream Model for Speaker Recognition with Self-Supervised Speech Features?
por: Aldeneh, Zakaria, et al.
Publicado: (2024)
por: Aldeneh, Zakaria, et al.
Publicado: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
PEFT-SER: On the Use of Parameter Efficient Transfer Learning Approaches For Speech Emotion Recognition Using Pre-trained Speech Models
por: Feng, Tiantian, et al.
Publicado: (2023)
por: Feng, Tiantian, et al.
Publicado: (2023)
PCQ: Emotion Recognition in Speech via Progressive Channel Querying
por: Wang, Xincheng, et al.
Publicado: (2024)
por: Wang, Xincheng, et al.
Publicado: (2024)
SELM: Enhancing Speech Emotion Recognition for Out-of-Domain Scenarios
por: Bukhari, Hazim, et al.
Publicado: (2024)
por: Bukhari, Hazim, et al.
Publicado: (2024)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
Machine Unlearning in Speech Emotion Recognition via Forget Set Alone
por: Ren, Zhao, et al.
Publicado: (2025)
por: Ren, Zhao, et al.
Publicado: (2025)
Semantic-Emotional Resonance Embedding: A Semi-Supervised Paradigm for Cross-Lingual Speech Emotion Recognition
por: Zhao, Ya, et al.
Publicado: (2026)
por: Zhao, Ya, et al.
Publicado: (2026)
EmoQ: Speech Emotion Recognition via Speech-Aware Q-Former and Large Language Model
por: Yang, Yiqing, et al.
Publicado: (2025)
por: Yang, Yiqing, et al.
Publicado: (2025)
Bridging Speech Emotion Recognition and Personality: Dataset and Temporal Interaction Condition Network
por: Gao, Yuan, et al.
Publicado: (2025)
por: Gao, Yuan, et al.
Publicado: (2025)
Temporal-Frequency State Space Duality: An Efficient Paradigm for Speech Emotion Recognition
por: Zhao, Jiaqi, et al.
Publicado: (2024)
por: Zhao, Jiaqi, et al.
Publicado: (2024)
Ejemplares similares
-
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024) -
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023) -
Using voice analysis as an early indicator of risk for depression in young adults
por: Scherer, Klaus R., et al.
Publicado: (2024) -
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
por: Rajapakshe, Thejan, et al.
Publicado: (2022) -
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
por: Li, Yupei, et al.
Publicado: (2024)