Quantifying the Role of Textual Predictability in Automatic Speech Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Robertson, Sean, Penn, Gerald, Dunbar, Ewan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bigger is not Always Better: The Effect of Context Size on Speech Pre-Training
von: Robertson, Sean, et al.
Veröffentlicht: (2023)
von: Robertson, Sean, et al.
Veröffentlicht: (2023)
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
von: Ong, Michael, et al.
Veröffentlicht: (2024)
von: Ong, Michael, et al.
Veröffentlicht: (2024)
Automatic Textual Normalization for Hate Speech Detection
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2023)
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2023)
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
von: Huo, Robin, et al.
Veröffentlicht: (2025)
von: Huo, Robin, et al.
Veröffentlicht: (2025)
Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
von: Mujtaba, Dena, et al.
Veröffentlicht: (2024)
von: Mujtaba, Dena, et al.
Veröffentlicht: (2024)
LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect
von: Naouara, Hedi, et al.
Veröffentlicht: (2025)
von: Naouara, Hedi, et al.
Veröffentlicht: (2025)
Reexamining Racial Disparities in Automatic Speech Recognition Performance: The Role of Confounding by Provenance
von: Li, Changye, et al.
Veröffentlicht: (2024)
von: Li, Changye, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition for the Ika Language
von: Nzenwata, Uchenna, et al.
Veröffentlicht: (2024)
von: Nzenwata, Uchenna, et al.
Veröffentlicht: (2024)
Responsible Benchmarking of Fairness for Automatic Speech Recognition
von: Herron, Felix, et al.
Veröffentlicht: (2026)
von: Herron, Felix, et al.
Veröffentlicht: (2026)
Vietnamese Automatic Speech Recognition: A Revisit
von: Vu, Thi, et al.
Veröffentlicht: (2026)
von: Vu, Thi, et al.
Veröffentlicht: (2026)
Investigating Transcription Normalization in the Faetar ASR Benchmark
von: Peckham, Leo, et al.
Veröffentlicht: (2025)
von: Peckham, Leo, et al.
Veröffentlicht: (2025)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Hindi
von: Saha, Anish, et al.
Veröffentlicht: (2024)
von: Saha, Anish, et al.
Veröffentlicht: (2024)
Zero Resource Code-switched Speech Benchmark Using Speech Utterance Pairs For Multiple Spoken Languages
von: Huang, Kuan-Po, et al.
Veröffentlicht: (2023)
von: Huang, Kuan-Po, et al.
Veröffentlicht: (2023)
Automatic Speech Recognition Advancements for Indigenous Languages of the Americas
von: Romero, Monica, et al.
Veröffentlicht: (2024)
von: Romero, Monica, et al.
Veröffentlicht: (2024)
WST: Weakly Supervised Transducer for Automatic Speech Recognition
von: Gao, Dongji, et al.
Veröffentlicht: (2025)
von: Gao, Dongji, et al.
Veröffentlicht: (2025)
Stuttering-Aware Automatic Speech Recognition for Indonesian Language
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2026)
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2026)
Where Are We At with Automatic Speech Recognition for the Bambara Language?
von: Diallo, Seydou, et al.
Veröffentlicht: (2026)
von: Diallo, Seydou, et al.
Veröffentlicht: (2026)
Syllabic-Structure Decoder for Automatic Speech Recognition in Vietnamese
von: Nguyen, Nghia Hieu, et al.
Veröffentlicht: (2026)
von: Nguyen, Nghia Hieu, et al.
Veröffentlicht: (2026)
Speech-Aware Long Context Pruning and Integration for Contextualized Automatic Speech Recognition
von: Rong, Yiming, et al.
Veröffentlicht: (2025)
von: Rong, Yiming, et al.
Veröffentlicht: (2025)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Khoa Anh, et al.
Veröffentlicht: (2026)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
von: Luu, Nam, et al.
Veröffentlicht: (2025)
von: Luu, Nam, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Sanskrit with Transfer Learning
von: Sadhukhan, Bidit, et al.
Veröffentlicht: (2025)
von: Sadhukhan, Bidit, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Greek Medical Dictation
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units
von: Poli, Maxime, et al.
Veröffentlicht: (2026)
von: Poli, Maxime, et al.
Veröffentlicht: (2026)
What does the Knowledge Neuron Thesis Have to do with Knowledge?
von: Niu, Jingcheng, et al.
Veröffentlicht: (2024)
von: Niu, Jingcheng, et al.
Veröffentlicht: (2024)
Bilevel Joint Unsupervised and Supervised Training for Automatic Speech Recognition
von: Cui, Xiaodong, et al.
Veröffentlicht: (2024)
von: Cui, Xiaodong, et al.
Veröffentlicht: (2024)
Efficient infusion of self-supervised representations in Automatic Speech Recognition
von: Prabhu, Darshan, et al.
Veröffentlicht: (2024)
von: Prabhu, Darshan, et al.
Veröffentlicht: (2024)
Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
PhoWhisper: Automatic Speech Recognition for Vietnamese
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
von: Le, Thanh-Thien, et al.
Veröffentlicht: (2024)
Augmenting Automatic Speech Recognition Models with Disfluency Detection
von: Amann, Robin, et al.
Veröffentlicht: (2024)
von: Amann, Robin, et al.
Veröffentlicht: (2024)
Continuously Learning New Words in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems
von: Iakovenko, Olga, et al.
Veröffentlicht: (2024)
von: Iakovenko, Olga, et al.
Veröffentlicht: (2024)
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
von: Bañeras-Roux, Thibault, et al.
Veröffentlicht: (2026)
Federated Heterogeneous Language Model Optimization for Hybrid Automatic Speech Recognition
von: Hong, Mengze, et al.
Veröffentlicht: (2026)
von: Hong, Mengze, et al.
Veröffentlicht: (2026)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
Automatic Screening for Children with Speech Disorder using Automatic Speech Recognition: Opportunities and Challenges
von: Liu, Dancheng, et al.
Veröffentlicht: (2024)
von: Liu, Dancheng, et al.
Veröffentlicht: (2024)
Joint vs Sequential Speaker-Role Detection and Automatic Speech Recognition for Air-traffic Control
von: Blatt, Alexander, et al.
Veröffentlicht: (2024)
von: Blatt, Alexander, et al.
Veröffentlicht: (2024)
Huntington Disease Automatic Speech Recognition with Biomarker Supervision
von: Wang, Charles L., et al.
Veröffentlicht: (2026)
von: Wang, Charles L., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Bigger is not Always Better: The Effect of Context Size on Speech Pre-Training
von: Robertson, Sean, et al.
Veröffentlicht: (2023) -
The Faetar Benchmark: Speech Recognition in a Very Under-Resourced Language
von: Ong, Michael, et al.
Veröffentlicht: (2024) -
Automatic Textual Normalization for Hate Speech Detection
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2023) -
Iterative refinement, not training objective, makes HuBERT behave differently from wav2vec 2.0
von: Huo, Robin, et al.
Veröffentlicht: (2025) -
Lost in Transcription: Identifying and Quantifying the Accuracy Biases of Automatic Speech Recognition Systems Against Disfluent Speech
von: Mujtaba, Dena, et al.
Veröffentlicht: (2024)