Uncertainty-Based Ensemble Learning For Speech Classification
Fuente:
arXiv
Guardado en:
| Autores principales: | Atmaja, Bagus Tris, Burkhardt, Felix |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Toward Natural Emotional Text-To-Speech System with Fine-Grained Non-Verbal Expression Control
por: Zhou, Wangzixi, et al.
Publicado: (2026)
por: Zhou, Wangzixi, et al.
Publicado: (2026)
Are you sure? Analysing Uncertainty Quantification Approaches for Real-world Speech Emotion Recognition
por: Schrüfer, Oliver, et al.
Publicado: (2024)
por: Schrüfer, Oliver, et al.
Publicado: (2024)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023)
por: Derington, Anna, et al.
Publicado: (2023)
Automatic Detection of Depression in Speech Using Ensemble Convolutional Neural Networks
por: Vázquez-Romero, Adrián, et al.
Publicado: (2024)
por: Vázquez-Romero, Adrián, et al.
Publicado: (2024)
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
por: Kuhlmann, Michael, et al.
Publicado: (2026)
por: Kuhlmann, Michael, et al.
Publicado: (2026)
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
SpeechMLC: Speech Multi-label Classification
por: Kim, Miseul, et al.
Publicado: (2025)
por: Kim, Miseul, et al.
Publicado: (2025)
Neck-Learn: Attention-Based Multiple Instance Learning and Ensemble Framework for Ecological Momentary Assessment
por: Cheema, Ahsan Jamal
Publicado: (2026)
por: Cheema, Ahsan Jamal
Publicado: (2026)
Speech Quality-Based Localization of Low-Quality Speech and Text-to-Speech Synthesis Artefacts
por: Kuhlmann, Michael, et al.
Publicado: (2026)
por: Kuhlmann, Michael, et al.
Publicado: (2026)
On Calibration of Speech Classification Models: Insights from Energy-Based Model Investigations
por: Hao, Yaqian, et al.
Publicado: (2024)
por: Hao, Yaqian, et al.
Publicado: (2024)
Leveraging Cascaded Binary Classification and Multimodal Fusion for Dementia Detection through Spontaneous Speech
por: Liu, Yin-Long, et al.
Publicado: (2025)
por: Liu, Yin-Long, et al.
Publicado: (2025)
HPP-Voice: A Large-Scale Evaluation of Speech Embeddings for Multi-Phenotypic Classification
por: Krongauz, David, et al.
Publicado: (2025)
por: Krongauz, David, et al.
Publicado: (2025)
AmbiDrop: Array-Agnostic Speech Enhancement Using Ambisonics Encoding and Dropout-Based Learning
por: Tatarjitzky, Michael, et al.
Publicado: (2025)
por: Tatarjitzky, Michael, et al.
Publicado: (2025)
TripleC Learning and Lightweight Speech Enhancement for Multi-Condition Target Speech Extraction
por: Huang, Ziling
Publicado: (2025)
por: Huang, Ziling
Publicado: (2025)
Enhancement of Dysarthric Speech Reconstruction by Contrastive Learning
por: Fatemeh, Keshvari, et al.
Publicado: (2024)
por: Fatemeh, Keshvari, et al.
Publicado: (2024)
AdaMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
por: Eom, SooHwan, et al.
Publicado: (2024)
por: Eom, SooHwan, et al.
Publicado: (2024)
Classification of Autistic and Non-Autistic Children's Speech: A Cross-Linguistic Study in Finnish, French, and Slovak
por: Kakouros, Sofoklis, et al.
Publicado: (2026)
por: Kakouros, Sofoklis, et al.
Publicado: (2026)
Speech Intelligibility Assessment with Uncertainty-Aware Whisper Embeddings and sLSTM
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
por: Zezario, Ryandhimas E., et al.
Publicado: (2025)
A Multilingual Framework for Dysarthria: Detection, Severity Classification, Speech-to-Text, and Clean Speech Generation
por: Raghu, Ananya, et al.
Publicado: (2025)
por: Raghu, Ananya, et al.
Publicado: (2025)
An Exploration of Length Generalization in Transformer-Based Speech Enhancement
por: Zhang, Qiquan, et al.
Publicado: (2024)
por: Zhang, Qiquan, et al.
Publicado: (2024)
GMM-ResNet2: Ensemble of Group ResNet Networks for Synthetic Speech Detection
por: Lei, Zhenchun, et al.
Publicado: (2024)
por: Lei, Zhenchun, et al.
Publicado: (2024)
Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model
por: Ihori, Mana, et al.
Publicado: (2025)
por: Ihori, Mana, et al.
Publicado: (2025)
Egonoise Resilient Source Localization and Speech Enhancement for Drones Using a Hybrid Model and Learning-Based Approach
por: Wu, Yihsuan, et al.
Publicado: (2025)
por: Wu, Yihsuan, et al.
Publicado: (2025)
Using voice analysis as an early indicator of risk for depression in young adults
por: Scherer, Klaus R., et al.
Publicado: (2024)
por: Scherer, Klaus R., et al.
Publicado: (2024)
Adaptive Speech Emotion Representation Learning Based On Dynamic Graph
por: Gao, Yingxue, et al.
Publicado: (2024)
por: Gao, Yingxue, et al.
Publicado: (2024)
Improving Active Learning for Melody Estimation by Disentangling Uncertainties
por: Jaiswal, Aayush, et al.
Publicado: (2025)
por: Jaiswal, Aayush, et al.
Publicado: (2025)
Unsupervised Online Continual Learning for Automatic Speech Recognition
por: Eeckt, Steven Vander, et al.
Publicado: (2024)
por: Eeckt, Steven Vander, et al.
Publicado: (2024)
Multitask Learning with Capsule Networks for Speech-to-Intent Applications
por: Poncelet, Jakob, et al.
Publicado: (2020)
por: Poncelet, Jakob, et al.
Publicado: (2020)
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
por: Bhattacharjee, Susmita, et al.
Publicado: (2025)
por: Bhattacharjee, Susmita, et al.
Publicado: (2025)
Speech Understanding on Tiny Devices with A Learning Cache
por: Benazir, Afsara, et al.
Publicado: (2023)
por: Benazir, Afsara, et al.
Publicado: (2023)
Speech-Based Estimation of Schizophrenia Severity Using Feature Fusion
por: Premananth, Gowtham, et al.
Publicado: (2024)
por: Premananth, Gowtham, et al.
Publicado: (2024)
Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens
por: Zhao, Jinzheng, et al.
Publicado: (2024)
por: Zhao, Jinzheng, et al.
Publicado: (2024)
AS-Speech: Adaptive Style For Speech Synthesis
por: Li, Zhipeng, et al.
Publicado: (2024)
por: Li, Zhipeng, et al.
Publicado: (2024)
Domain-Incremental Learning for Audio Classification
por: Mulimani, Manjunath, et al.
Publicado: (2024)
por: Mulimani, Manjunath, et al.
Publicado: (2024)
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition
por: de Groot, Dimme, et al.
Publicado: (2026)
por: de Groot, Dimme, et al.
Publicado: (2026)
LMU-Based Sequential Learning and Posterior Ensemble Fusion for Cross-Domain Infant Cry Classification
por: Jazaeri, Niloofar, et al.
Publicado: (2026)
por: Jazaeri, Niloofar, et al.
Publicado: (2026)
Learning Time-Graph Frequency Representation for Monaural Speech Enhancement
por: Wang, Tingting, et al.
Publicado: (2025)
por: Wang, Tingting, et al.
Publicado: (2025)
ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals
por: E, Ameenudeen P, et al.
Publicado: (2026)
por: E, Ameenudeen P, et al.
Publicado: (2026)
EffortNet: A Deep Learning Framework for Objective Assessment of Speech Enhancement Technologies Using EEG-Based Alpha Oscillations
por: Sung, Ching-Chih, et al.
Publicado: (2025)
por: Sung, Ching-Chih, et al.
Publicado: (2025)
Reverse Attention for Lightweight Speech Enhancement on Edge Devices
por: Ojha, Shuubham, et al.
Publicado: (2025)
por: Ojha, Shuubham, et al.
Publicado: (2025)
Ejemplares similares
-
Toward Natural Emotional Text-To-Speech System with Fine-Grained Non-Verbal Expression Control
por: Zhou, Wangzixi, et al.
Publicado: (2026) -
Are you sure? Analysing Uncertainty Quantification Approaches for Real-world Speech Emotion Recognition
por: Schrüfer, Oliver, et al.
Publicado: (2024) -
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023) -
Automatic Detection of Depression in Speech Using Ensemble Convolutional Neural Networks
por: Vázquez-Romero, Adrián, et al.
Publicado: (2024) -
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
por: Kuhlmann, Michael, et al.
Publicado: (2026)