Discrimination loss vs. SRT: A model-based approach towards harmonizing speech test interpretations
Fuente:
arXiv
Saved in:
| Main Authors: | Buhl, Mareike, Kludt, Eugen, Schell-Majoor, Lena, Avan, Paul, Campi, Marta |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Objective comparison of auditory profiles using manifold learning and intrinsic measures
by: Xu, Chen, et al.
Published: (2026)
by: Xu, Chen, et al.
Published: (2026)
Data Standards in Audiology: A Mixed-Methods Exploration of Community Perspectives and Implementation Considerations
by: Vercammen, Charlotte, et al.
Published: (2025)
by: Vercammen, Charlotte, et al.
Published: (2025)
Integrating audiological datasets via federated merging of Auditory Profiles
by: Saak, Samira, et al.
Published: (2024)
by: Saak, Samira, et al.
Published: (2024)
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
by: Deng, Qingkun, et al.
Published: (2024)
by: Deng, Qingkun, et al.
Published: (2024)
The UmboMic: A PVDF Cantilever Microphone
by: Yeiser, Aaron J., et al.
Published: (2023)
by: Yeiser, Aaron J., et al.
Published: (2023)
An Implantable Piezofilm Middle Ear Microphone: Performance in Human Cadaveric Temporal Bones
by: Zhang, John Z., et al.
Published: (2023)
by: Zhang, John Z., et al.
Published: (2023)
On the relevance of acoustic measurements for creating realistic virtual acoustic environments
by: Gündert, Siegfried, et al.
Published: (2023)
by: Gündert, Siegfried, et al.
Published: (2023)
Standard audiogram classification from loudness scaling data using unsupervised, supervised, and explainable machine learning techniques
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Disentangling peripheral hearing loss from central and cognitive effects on speech intelligibility in older adults
by: Irino, Toshio, et al.
Published: (2025)
by: Irino, Toshio, et al.
Published: (2025)
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
by: Kumar, Anurag, et al.
Published: (2024)
by: Kumar, Anurag, et al.
Published: (2024)
Semantic enrichment towards efficient speech representations
by: Laperrière, Gaëlle, et al.
Published: (2023)
by: Laperrière, Gaëlle, et al.
Published: (2023)
An adaptive filter bank based neural network approach for time delay estimation and speech enhancement
by: Ma, Lu
Published: (2025)
by: Ma, Lu
Published: (2025)
Single-channel speech enhancement using learnable loss mixup
by: Chang, Oscar, et al.
Published: (2023)
by: Chang, Oscar, et al.
Published: (2023)
A two-step approach for speech enhancement in low-SNR scenarios using cyclostationary beamforming and DNNs
by: Bologni, Giovanni, et al.
Published: (2026)
by: Bologni, Giovanni, et al.
Published: (2026)
Towards noise-robust speech inversion through multi-task learning with speech enhancement
by: Tabatabaee, Saba, et al.
Published: (2026)
by: Tabatabaee, Saba, et al.
Published: (2026)
On the relationship between speech and hearing
by: Umesh, Srinivasan, et al.
Published: (2024)
by: Umesh, Srinivasan, et al.
Published: (2024)
Robust fine-tuning of speech recognition models via model merging: application to disordered speech
by: Ducorroy, Alexandre, et al.
Published: (2025)
by: Ducorroy, Alexandre, et al.
Published: (2025)
learning discriminative features from spectrograms using center loss for speech emotion recognition
by: Dai, Dongyang, et al.
Published: (2025)
by: Dai, Dongyang, et al.
Published: (2025)
Probing mental health information in speech foundation models
by: de Gennes, Marc, et al.
Published: (2024)
by: de Gennes, Marc, et al.
Published: (2024)
On the effectiveness of enrollment speech augmentation for Target Speaker Extraction
by: Li, Junjie, et al.
Published: (2024)
by: Li, Junjie, et al.
Published: (2024)
Distilling a speech and music encoder with task arithmetic
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
by: Ritter-Gutierrez, Fabian, et al.
Published: (2025)
WhisperFlow: speech foundation models in real time
by: Wang, Rongxiang, et al.
Published: (2024)
by: Wang, Rongxiang, et al.
Published: (2024)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
by: Maisonneuve, Malo, et al.
Published: (2024)
by: Maisonneuve, Malo, et al.
Published: (2024)
Omni-directional attention mechanism based on Mamba for speech separation
by: Xue, Ke, et al.
Published: (2026)
by: Xue, Ke, et al.
Published: (2026)
Unsupervised speech enhancement with spectral kurtosis and double deep priors
by: Ohnaka, Hien, et al.
Published: (2024)
by: Ohnaka, Hien, et al.
Published: (2024)
BFA: Real-time Multilingual Text-to-speech Forced Alignment
by: Rehman, Abdul, et al.
Published: (2025)
by: Rehman, Abdul, et al.
Published: (2025)
SPGM: Prioritizing Local Features for enhanced speech separation performance
by: Yip, Jia Qi, et al.
Published: (2023)
by: Yip, Jia Qi, et al.
Published: (2023)
Inter-channel Conv-TasNet for multichannel speech enhancement
by: Lee, Dongheon, et al.
Published: (2021)
by: Lee, Dongheon, et al.
Published: (2021)
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
by: Yanir, Efrayim, et al.
Published: (2025)
by: Yanir, Efrayim, et al.
Published: (2025)
Graph-based multi-Feature fusion method for speech emotion recognition
by: Liu, Xueyu, et al.
Published: (2024)
by: Liu, Xueyu, et al.
Published: (2024)
Expressive paragraph text-to-speech synthesis with multi-step variational autoencoder
by: Li, Xuyuan, et al.
Published: (2023)
by: Li, Xuyuan, et al.
Published: (2023)
Towards generalisable and calibrated synthetic speech detection with self-supervised representations
by: Pascu, Octavian, et al.
Published: (2023)
by: Pascu, Octavian, et al.
Published: (2023)
A lightweight and robust method for blind wideband-to-fullband extension of speech
by: Büthe, Jan, et al.
Published: (2024)
by: Büthe, Jan, et al.
Published: (2024)
Monaural speech enhancement on drone via Adapter based transfer learning
by: Chen, Xingyu, et al.
Published: (2024)
by: Chen, Xingyu, et al.
Published: (2024)
FreeCodec: A disentangled neural speech codec with fewer tokens
by: Zheng, Youqiang, et al.
Published: (2024)
by: Zheng, Youqiang, et al.
Published: (2024)
Prosodic Parameter Manipulation in TTS generated speech for Controlled Speech Generation
by: Chary, Podakanti Satyajith
Published: (2024)
by: Chary, Podakanti Satyajith
Published: (2024)
Adversarial speech for voice privacy protection from Personalized Speech generation
by: Chen, Shihao, et al.
Published: (2024)
by: Chen, Shihao, et al.
Published: (2024)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
by: Das, Sneha, et al.
Published: (2020)
by: Das, Sneha, et al.
Published: (2020)
An efficient text augmentation approach for contextualized Mandarin speech recognition
by: Zheng, Naijun, et al.
Published: (2024)
by: Zheng, Naijun, et al.
Published: (2024)
Towards a generalized monaural and binaural auditory model for psychoacoustics and speech intelligibility
by: Biberger, Thomas, et al.
Published: (2021)
by: Biberger, Thomas, et al.
Published: (2021)
Similar Items
-
Objective comparison of auditory profiles using manifold learning and intrinsic measures
by: Xu, Chen, et al.
Published: (2026) -
Data Standards in Audiology: A Mixed-Methods Exploration of Community Perspectives and Implementation Considerations
by: Vercammen, Charlotte, et al.
Published: (2025) -
Integrating audiological datasets via federated merging of Auditory Profiles
by: Saak, Samira, et al.
Published: (2024) -
An interpretable speech foundation model for depression detection by revealing prediction-relevant acoustic features from long speech
by: Deng, Qingkun, et al.
Published: (2024) -
The UmboMic: A PVDF Cantilever Microphone
by: Yeiser, Aaron J., et al.
Published: (2023)