Beyond Word Error Rate: Auditing the Diversity Tax in Speech Recognition through Dataset Cartography
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Ting-Hui, Clemmensen, Line H., Das, Sneha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
by: Paraschou, Eva, et al.
Published: (2026)
by: Paraschou, Eva, et al.
Published: (2026)
Pantypes: Diverse Representatives for Self-Explainable Models
by: Kjærsgaard, Rune, et al.
Published: (2024)
by: Kjærsgaard, Rune, et al.
Published: (2024)
Exploring Local Interpretable Model-Agnostic Explanations for Speech Emotion Recognition with Distribution-Shift
by: Hjuler, Maja J., et al.
Published: (2025)
by: Hjuler, Maja J., et al.
Published: (2025)
Post-hoc Self-explanation of CNNs
by: Boubekki, Ahcène, et al.
Published: (2026)
by: Boubekki, Ahcène, et al.
Published: (2026)
Evaluation of Stress Detection as Time Series Events -- A Novel Window-Based F1-Metric
by: Skat-Rørdam, Harald Vilhelm, et al.
Published: (2025)
by: Skat-Rørdam, Harald Vilhelm, et al.
Published: (2025)
EmoTale: An Enacted Speech-emotion Dataset in Danish
by: Hjuler, Maja J., et al.
Published: (2025)
by: Hjuler, Maja J., et al.
Published: (2025)
The Verification Tax: Fundamental Limits of AI Auditing in the Rare-Error Regime
by: Wang, Jason Z
Published: (2026)
by: Wang, Jason Z
Published: (2026)
ArEEG_Words: Dataset for Envisioned Speech Recognition using EEG for Arabic Words
by: Darwish, Hazem, et al.
Published: (2024)
by: Darwish, Hazem, et al.
Published: (2024)
Classification of Tennis Actions Using Deep Learning
by: Hovad, Emil, et al.
Published: (2024)
by: Hovad, Emil, et al.
Published: (2024)
A Self-Organizing Clustering System for Unsupervised Distribution Shift Detection
by: Basterrech, Sebastián, et al.
Published: (2024)
by: Basterrech, Sebastián, et al.
Published: (2024)
Continuously Learning New Words in Automatic Speech Recognition
by: Huber, Christian, et al.
Published: (2024)
by: Huber, Christian, et al.
Published: (2024)
Beyond the Labels: Unveiling Text-Dependency in Paralinguistic Speech Recognition Datasets
by: Pešán, Jan, et al.
Published: (2024)
by: Pešán, Jan, et al.
Published: (2024)
Exploratory Evaluation of Speech Content Masking
by: Williams, Jennifer, et al.
Published: (2024)
by: Williams, Jennifer, et al.
Published: (2024)
MSNER: A Multilingual Speech Dataset for Named Entity Recognition
by: Meeus, Quentin, et al.
Published: (2024)
by: Meeus, Quentin, et al.
Published: (2024)
Vision-based Deep Learning Analysis of Unordered Biomedical Tabular Datasets via Optimal Spatial Cartography
by: Mostafa, Sakib, et al.
Published: (2026)
by: Mostafa, Sakib, et al.
Published: (2026)
Privacy Auditing Synthetic Data Release through Local Likelihood Attacks
by: Ward, Joshua, et al.
Published: (2025)
by: Ward, Joshua, et al.
Published: (2025)
Not All Errors Are Equal: Investigation of Speech Recognition Errors in Alzheimer's Disease Detection
by: Kang, Jiawen, et al.
Published: (2024)
by: Kang, Jiawen, et al.
Published: (2024)
WhisperD: Dementia Speech Recognition and Filler Word Detection with Whisper
by: Akinrintoyo, Emmanuel, et al.
Published: (2025)
by: Akinrintoyo, Emmanuel, et al.
Published: (2025)
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models
by: Waghela, Hetvi, et al.
Published: (2024)
by: Waghela, Hetvi, et al.
Published: (2024)
Test-Time Adaptation for Speech Emotion Recognition
by: Dong, Jiaheng, et al.
Published: (2026)
by: Dong, Jiaheng, et al.
Published: (2026)
SoK: Dataset Copyright Auditing in Machine Learning Systems
by: Du, Linkang, et al.
Published: (2024)
by: Du, Linkang, et al.
Published: (2024)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
by: Ranjbar, Hossein, et al.
Published: (2025)
by: Ranjbar, Hossein, et al.
Published: (2025)
Accelerating Regularized Attention Kernel Regression for Spectrum Cartography
by: Tao, Liping, et al.
Published: (2026)
by: Tao, Liping, et al.
Published: (2026)
Supplementary Resources and Analysis for Automatic Speech Recognition Systems Trained on the Loquacious Dataset
by: Rossenbach, Nick, et al.
Published: (2025)
by: Rossenbach, Nick, et al.
Published: (2025)
Beyond Confusion: A Fine-grained Dialectical Examination of Human Activity Recognition Benchmark Datasets
by: Geissler, Daniel, et al.
Published: (2024)
by: Geissler, Daniel, et al.
Published: (2024)
Fast Rate Information-theoretic Bounds on Generalization Errors
by: Wu, Xuetong, et al.
Published: (2023)
by: Wu, Xuetong, et al.
Published: (2023)
KoTaP: A Panel Dataset for Corporate Tax Avoidance, Performance, and Governance in Korea
by: Na, Hyungjong, et al.
Published: (2025)
by: Na, Hyungjong, et al.
Published: (2025)
VariFace: Fair and Diverse Synthetic Dataset Generation for Face Recognition
by: Yeung, Michael, et al.
Published: (2024)
by: Yeung, Michael, et al.
Published: (2024)
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
by: Krishna, Kundan, et al.
Published: (2024)
by: Krishna, Kundan, et al.
Published: (2024)
Is it the model or the metric -- On robustness measures of deeplearning models
by: Lyu, Zhijin, et al.
Published: (2024)
by: Lyu, Zhijin, et al.
Published: (2024)
Optimal Multiclass U-Calibration Error and Beyond
by: Luo, Haipeng, et al.
Published: (2024)
by: Luo, Haipeng, et al.
Published: (2024)
Leveraging Error Diversity in Group Rollouts for Reinforcement Learning
by: Liu, Wenpu, et al.
Published: (2026)
by: Liu, Wenpu, et al.
Published: (2026)
Bayes Error Rate Estimation in Difficult Situations
by: Wheat, Lesley, et al.
Published: (2025)
by: Wheat, Lesley, et al.
Published: (2025)
Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation
by: Fu, Yu, et al.
Published: (2026)
by: Fu, Yu, et al.
Published: (2026)
Error Diversity Matters: An Error-Resistant Ensemble Method for Unsupervised Dependency Parsing
by: Shayegh, Behzad, et al.
Published: (2024)
by: Shayegh, Behzad, et al.
Published: (2024)
Modelling Emotions is an Elusive Pursuit in Affective Computing
by: Larsen, Anders Rolighed, et al.
Published: (2026)
by: Larsen, Anders Rolighed, et al.
Published: (2026)
Mitigating the Alignment Tax of RLHF
by: Lin, Yong, et al.
Published: (2023)
by: Lin, Yong, et al.
Published: (2023)
Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
by: Yang, Chao-Han Huck, et al.
Published: (2024)
by: Yang, Chao-Han Huck, et al.
Published: (2024)
Arabic Little STT: Arabic Children Speech Recognition Dataset
by: Alkadri, Mouhand, et al.
Published: (2025)
by: Alkadri, Mouhand, et al.
Published: (2025)
Similar Items
-
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
by: Paraschou, Eva, et al.
Published: (2026) -
Pantypes: Diverse Representatives for Self-Explainable Models
by: Kjærsgaard, Rune, et al.
Published: (2024) -
Exploring Local Interpretable Model-Agnostic Explanations for Speech Emotion Recognition with Distribution-Shift
by: Hjuler, Maja J., et al.
Published: (2025) -
Post-hoc Self-explanation of CNNs
by: Boubekki, Ahcène, et al.
Published: (2026) -
Evaluation of Stress Detection as Time Series Events -- A Novel Window-Based F1-Metric
by: Skat-Rørdam, Harald Vilhelm, et al.
Published: (2025)