How much speech data is necessary for ASR in African languages? An evaluation of data scaling in Kinyarwanda and Kikuyu
Fuente:
arXiv
Salvato in:
| Autori principali: | Akera, Benjamin, Nafula, Evelyn, Walukagga, Patrick, Yiga, Gilbert, Quinn, John, Mwebaze, Ernest |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sunflower: A New Approach To Expanding Coverage of African Languages in Large Language Models
di: Akera, Benjamin, et al.
Pubblicazione: (2025)
di: Akera, Benjamin, et al.
Pubblicazione: (2025)
How much do language models memorize?
di: Morris, John X., et al.
Pubblicazione: (2025)
di: Morris, John X., et al.
Pubblicazione: (2025)
`It's not just linguistically, there's much more going on’: The experiences and practices of bilingual paediatric speech and language therapists in the UK
di: Mélanie Gréaux, et al.
Pubblicazione: (2024)
di: Mélanie Gréaux, et al.
Pubblicazione: (2024)
AfriHuBERT: A self-supervised speech representation model for African languages
di: Alabi, Jesujoba O., et al.
Pubblicazione: (2024)
di: Alabi, Jesujoba O., et al.
Pubblicazione: (2024)
Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data
di: Kashiwagi, Yosuke, et al.
Pubblicazione: (2025)
di: Kashiwagi, Yosuke, et al.
Pubblicazione: (2025)
Evaluating ASR robustness to spontaneous speech errors: A study of WhisperX using a Speech Error Database
di: Alderete, John, et al.
Pubblicazione: (2025)
di: Alderete, John, et al.
Pubblicazione: (2025)
Ensemble of pre-trained language models and data augmentation for hate speech detection from Arabic tweets
di: Daouadi, Kheir Eddine, et al.
Pubblicazione: (2024)
di: Daouadi, Kheir Eddine, et al.
Pubblicazione: (2024)
How speech and language therapists and parents work together in the therapeutic process for children with speech sound disorder: A scoping review
di: Katherine Pritchard, et al.
Pubblicazione: (2024)
di: Katherine Pritchard, et al.
Pubblicazione: (2024)
Linguistically Informed Evaluation of Multilingual ASR for African Languages
di: Chen, Fei-Yueh, et al.
Pubblicazione: (2026)
di: Chen, Fei-Yueh, et al.
Pubblicazione: (2026)
The Greek podcast corpus: Competitive speech models for low-resourced languages with weakly supervised data
di: Paraskevopoulos, Georgios, et al.
Pubblicazione: (2024)
di: Paraskevopoulos, Georgios, et al.
Pubblicazione: (2024)
Balti-Tamko 02: Continuous speech dataset for Balti ASR.
di: Sharif, Muhammad
Pubblicazione: (2025)
di: Sharif, Muhammad
Pubblicazione: (2025)
WER is Unaware: Assessing How ASR Errors Distort Clinical Understanding in Patient Facing Dialogue
di: Ellis, Zachary, et al.
Pubblicazione: (2025)
di: Ellis, Zachary, et al.
Pubblicazione: (2025)
InkubaLM: A small language model for low-resource African languages
di: Tonja, Atnafu Lambebo, et al.
Pubblicazione: (2024)
di: Tonja, Atnafu Lambebo, et al.
Pubblicazione: (2024)
Natural language processing for African languages
di: Adelani, David Ifeoluwa
Pubblicazione: (2025)
di: Adelani, David Ifeoluwa
Pubblicazione: (2025)
Improving endpoint detection in end-to-end streaming ASR for conversational speech
di: C, Anandh, et al.
Pubblicazione: (2025)
di: C, Anandh, et al.
Pubblicazione: (2025)
How much do contextualized representations encode long-range context?
di: Sun, Simeng, et al.
Pubblicazione: (2024)
di: Sun, Simeng, et al.
Pubblicazione: (2024)
Building English ASR model with regional language support
di: Agrawal, Purvi, et al.
Pubblicazione: (2025)
di: Agrawal, Purvi, et al.
Pubblicazione: (2025)
Measures of time to speech and language therapy for children with autism spectrum disorder
di: Ana Carina Tamanaha
Pubblicazione: (2014)
di: Ana Carina Tamanaha
Pubblicazione: (2014)
KinSPEAK: Improving speech recognition for Kinyarwanda via semi-supervised learning methods
di: Nzeyimana, Antoine
Pubblicazione: (2023)
di: Nzeyimana, Antoine
Pubblicazione: (2023)
Fine-tuning Whisper for Pashto ASR: strategies and scale
di: Rahman, Hanif
Pubblicazione: (2026)
di: Rahman, Hanif
Pubblicazione: (2026)
Style-agnostic evaluation of ASR using multiple reference transcripts
di: McNamara, Quinten, et al.
Pubblicazione: (2024)
di: McNamara, Quinten, et al.
Pubblicazione: (2024)
Quantification of stylistic differences in human- and ASR-produced transcripts of African American English
di: Heuser, Annika, et al.
Pubblicazione: (2024)
di: Heuser, Annika, et al.
Pubblicazione: (2024)
The agreement of phonetic transcriptions between paediatric speech and language therapists transcribing a disordered speech sample
di: Laura Jane Mallaband
Pubblicazione: (2024)
di: Laura Jane Mallaband
Pubblicazione: (2024)
Are LLM-generated plain language summaries truly understandable? A large-scale crowdsourced evaluation
di: Guo, Yue, et al.
Pubblicazione: (2025)
di: Guo, Yue, et al.
Pubblicazione: (2025)
Strategies for improving low resource speech to text translation relying on pre-trained ASR models
di: Kesiraju, Santosh, et al.
Pubblicazione: (2023)
di: Kesiraju, Santosh, et al.
Pubblicazione: (2023)
Covertly improving intelligibility with data-driven adaptations of speech timing
di: Tuttösí, Paige, et al.
Pubblicazione: (2026)
di: Tuttösí, Paige, et al.
Pubblicazione: (2026)
How much reliable is ChatGPT's prediction on Information Extraction under Input Perturbations?
di: Mondal, Ishani, et al.
Pubblicazione: (2024)
di: Mondal, Ishani, et al.
Pubblicazione: (2024)
Can large audio language models understand child stuttering speech? speech summarization, and source separation
di: Okocha, Chibuzor, et al.
Pubblicazione: (2025)
di: Okocha, Chibuzor, et al.
Pubblicazione: (2025)
negativas: a prototype for searching and classifying sentential negation in speech data
di: de Gois, Túlio Sousa, et al.
Pubblicazione: (2025)
di: de Gois, Túlio Sousa, et al.
Pubblicazione: (2025)
How language models extrapolate outside the training data: A case study in Textualized Gridworld
di: Kim, Doyoung, et al.
Pubblicazione: (2024)
di: Kim, Doyoung, et al.
Pubblicazione: (2024)
Positive effects of speech and language therapy group interventions in primary progressive aphasia: A systematic review
di: Miyuki Watanabe, et al.
Pubblicazione: (2024)
di: Miyuki Watanabe, et al.
Pubblicazione: (2024)
Can we train ASR systems on Code-switch without real code-switch data? Case study for Singapore's languages
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
WAXAL-NET: Finetuned Edge ASR Across 19 African Languages
di: Olufemi, Victor Tolulope, et al.
Pubblicazione: (2026)
di: Olufemi, Victor Tolulope, et al.
Pubblicazione: (2026)
Transferable speech-to-text large language model alignment module
di: Wu, Boyong, et al.
Pubblicazione: (2024)
di: Wu, Boyong, et al.
Pubblicazione: (2024)
Classifying populist language in American presidential and governor speeches using automatic text analysis
di: van der Veen, Olaf, et al.
Pubblicazione: (2024)
di: van der Veen, Olaf, et al.
Pubblicazione: (2024)
Language-agnostic, automated assessment of listeners' speech recall using large language models
di: Herrmann, Björn
Pubblicazione: (2025)
di: Herrmann, Björn
Pubblicazione: (2025)
Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature Review
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
Are ASR foundation models generalized enough to capture features of regional dialects for low-resource languages?
di: Dipto, Tawsif Tashwar, et al.
Pubblicazione: (2025)
di: Dipto, Tawsif Tashwar, et al.
Pubblicazione: (2025)
Collaborative approaches with stakeholders in speech‐language pathology: Narrative literature review
di: Jessica Hassett, et al.
Pubblicazione: (2024)
di: Jessica Hassett, et al.
Pubblicazione: (2024)
Anthropocentric bias in language model evaluation
di: Millière, Raphaël, et al.
Pubblicazione: (2024)
di: Millière, Raphaël, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Sunflower: A New Approach To Expanding Coverage of African Languages in Large Language Models
di: Akera, Benjamin, et al.
Pubblicazione: (2025) -
How much do language models memorize?
di: Morris, John X., et al.
Pubblicazione: (2025) -
`It's not just linguistically, there's much more going on’: The experiences and practices of bilingual paediatric speech and language therapists in the UK
di: Mélanie Gréaux, et al.
Pubblicazione: (2024) -
AfriHuBERT: A self-supervised speech representation model for African languages
di: Alabi, Jesujoba O., et al.
Pubblicazione: (2024) -
Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data
di: Kashiwagi, Yosuke, et al.
Pubblicazione: (2025)