Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
Fuente:
arXiv
Guardado en:
| Autores principales: | Mujtaba, Dena, Mahapatra, Nihar R., Arney, Megan, Yaruss, J. Scott, Gerlach-Houck, Hope, Herring, Caryn, Bin, Jia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Inclusive ASR for Disfluent Speech: Cascaded Large-Scale Self-Supervised Learning with Targeted Fine-Tuning and Data Augmentation
por: Mujtaba, Dena, et al.
Publicado: (2024)
por: Mujtaba, Dena, et al.
Publicado: (2024)
Fine-Tuning ASR for Stuttered Speech: Personalized vs. Generalized Approaches
por: Mujtaba, Dena, et al.
Publicado: (2025)
por: Mujtaba, Dena, et al.
Publicado: (2025)
Behind the Screens: Uncovering Bias in AI-Driven Video Interview Assessments Using Counterfactuals
por: Mujtaba, Dena F., et al.
Publicado: (2025)
por: Mujtaba, Dena F., et al.
Publicado: (2025)
Fairness in AI-Driven Recruitment: Challenges, Metrics, Methods, and Future Directions
por: Mujtaba, Dena F., et al.
Publicado: (2024)
por: Mujtaba, Dena F., et al.
Publicado: (2024)
Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content Evaluations
por: Kadoma, Kowe, et al.
Publicado: (2026)
por: Kadoma, Kowe, et al.
Publicado: (2026)
Automatic Speech Recognition Biases in Newcastle English: an Error Analysis
por: Serditova, Dana, et al.
Publicado: (2025)
por: Serditova, Dana, et al.
Publicado: (2025)
Quantifying the Role of Textual Predictability in Automatic Speech Recognition
por: Robertson, Sean, et al.
Publicado: (2024)
por: Robertson, Sean, et al.
Publicado: (2024)
Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems
por: Iakovenko, Olga, et al.
Publicado: (2024)
por: Iakovenko, Olga, et al.
Publicado: (2024)
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
por: Huber, Christian, et al.
Publicado: (2025)
por: Huber, Christian, et al.
Publicado: (2025)
Measuring the Accuracy of Automatic Speech Recognition Solutions
por: Kuhn, Korbinian, et al.
Publicado: (2024)
por: Kuhn, Korbinian, et al.
Publicado: (2024)
Talking to...uh...um...Machines: The Impact of Disfluent Speech Agents on Partner Models and Perspective Taking
por: Jacka, Rhys, et al.
Publicado: (2025)
por: Jacka, Rhys, et al.
Publicado: (2025)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
por: Shakeel, Muhammad, et al.
Publicado: (2024)
por: Shakeel, Muhammad, et al.
Publicado: (2024)
OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary
por: Sudo, Yui, et al.
Publicado: (2025)
por: Sudo, Yui, et al.
Publicado: (2025)
The Impact of Automatic Speech Transcription on Speaker Attribution
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
por: Aggazzotti, Cristina, et al.
Publicado: (2025)
Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models
por: Atwany, Hanin, et al.
Publicado: (2025)
por: Atwany, Hanin, et al.
Publicado: (2025)
Automatic Speech Recognition for Non-Native English: Accuracy and Disfluency Handling
por: McGuire, Michael
Publicado: (2025)
por: McGuire, Michael
Publicado: (2025)
A Paradigm for Interpreting Metrics and Identifying Critical Errors in Automatic Speech Recognition
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Automatic Speech Recognition for the Ika Language
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
por: Nzenwata, Uchenna, et al.
Publicado: (2024)
Hallucinations in Neural Automatic Speech Recognition: Identifying Errors and Hallucinatory Models
por: Frieske, Rita, et al.
Publicado: (2024)
por: Frieske, Rita, et al.
Publicado: (2024)
Automatic Speech Recognition for Hindi
por: Saha, Anish, et al.
Publicado: (2024)
por: Saha, Anish, et al.
Publicado: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
por: Luu, Nam, et al.
Publicado: (2025)
por: Luu, Nam, et al.
Publicado: (2025)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
por: Kulkarni, Ajinkya, et al.
Publicado: (2025)
por: Kulkarni, Ajinkya, et al.
Publicado: (2025)
Responsible Benchmarking of Fairness for Automatic Speech Recognition
por: Herron, Felix, et al.
Publicado: (2026)
por: Herron, Felix, et al.
Publicado: (2026)
Vietnamese Automatic Speech Recognition: A Revisit
por: Vu, Thi, et al.
Publicado: (2026)
por: Vu, Thi, et al.
Publicado: (2026)
Speech-Aware Long Context Pruning and Integration for Contextualized Automatic Speech Recognition
por: Rong, Yiming, et al.
Publicado: (2025)
por: Rong, Yiming, et al.
Publicado: (2025)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
por: Min, Do June, et al.
Publicado: (2024)
por: Min, Do June, et al.
Publicado: (2024)
Automatic Screening for Children with Speech Disorder using Automatic Speech Recognition: Opportunities and Challenges
por: Liu, Dancheng, et al.
Publicado: (2024)
por: Liu, Dancheng, et al.
Publicado: (2024)
Automatic Speech Recognition for Sanskrit with Transfer Learning
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
por: Sadhukhan, Bidit, et al.
Publicado: (2025)
Automatic Speech Recognition for Greek Medical Dictation
por: Georgilas, Vardis, et al.
Publicado: (2025)
por: Georgilas, Vardis, et al.
Publicado: (2025)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
por: Hu, Jiliang, et al.
Publicado: (2025)
por: Hu, Jiliang, et al.
Publicado: (2025)
WST: Weakly Supervised Transducer for Automatic Speech Recognition
por: Gao, Dongji, et al.
Publicado: (2025)
por: Gao, Dongji, et al.
Publicado: (2025)
Stuttering-Aware Automatic Speech Recognition for Indonesian Language
por: Muhammad, Fadhil, et al.
Publicado: (2026)
por: Muhammad, Fadhil, et al.
Publicado: (2026)
Automatic Speech Recognition Advancements for Indigenous Languages of the Americas
por: Romero, Monica, et al.
Publicado: (2024)
por: Romero, Monica, et al.
Publicado: (2024)
Where Are We At with Automatic Speech Recognition for the Bambara Language?
por: Diallo, Seydou, et al.
Publicado: (2026)
por: Diallo, Seydou, et al.
Publicado: (2026)
Syllabic-Structure Decoder for Automatic Speech Recognition in Vietnamese
por: Nguyen, Nghia Hieu, et al.
Publicado: (2026)
por: Nguyen, Nghia Hieu, et al.
Publicado: (2026)
Moonshine: Speech Recognition for Live Transcription and Voice Commands
por: Jeffries, Nat, et al.
Publicado: (2024)
por: Jeffries, Nat, et al.
Publicado: (2024)
PhoWhisper: Automatic Speech Recognition for Vietnamese
por: Le, Thanh-Thien, et al.
Publicado: (2024)
por: Le, Thanh-Thien, et al.
Publicado: (2024)
SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
por: Du, Jiayu, et al.
Publicado: (2024)
por: Du, Jiayu, et al.
Publicado: (2024)
Challenges in Automatic Speech Recognition for Adults with Cognitive Impairment
por: Cohn, Michelle, et al.
Publicado: (2026)
por: Cohn, Michelle, et al.
Publicado: (2026)
Ejemplares similares
-
Inclusive ASR for Disfluent Speech: Cascaded Large-Scale Self-Supervised Learning with Targeted Fine-Tuning and Data Augmentation
por: Mujtaba, Dena, et al.
Publicado: (2024) -
Fine-Tuning ASR for Stuttered Speech: Personalized vs. Generalized Approaches
por: Mujtaba, Dena, et al.
Publicado: (2025) -
Behind the Screens: Uncovering Bias in AI-Driven Video Interview Assessments Using Counterfactuals
por: Mujtaba, Dena F., et al.
Publicado: (2025) -
Fairness in AI-Driven Recruitment: Challenges, Metrics, Methods, and Future Directions
por: Mujtaba, Dena F., et al.
Publicado: (2024) -
Lost in Transcription: Subtitle Errors in Automatic Speech Recognition Reduce Speaker and Content Evaluations
por: Kadoma, Kowe, et al.
Publicado: (2026)