Multistage Fine-tuning Strategies for Automatic Speech Recognition in Low-resource Languages
Fuente:
arXiv
Salvato in:
| Autori principali: | Pillai, Leena G, Manohar, Kavya, Raju, Basil K, Sherly, Elizabeth |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tracking Articulatory Dynamics in Speech with a Fixed-Weight BiLSTM-CNN Architecture
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
Acoustic to Articulatory Inversion of Speech; Data Driven Approaches, Challenges, Applications, and Future Scope
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
di: Pillai, Leena G, et al.
Pubblicazione: (2025)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
di: Hsu, Ming-Hao, et al.
Pubblicazione: (2024)
di: Hsu, Ming-Hao, et al.
Pubblicazione: (2024)
Automatic Speech Recognition for Hindi
di: Saha, Anish, et al.
Pubblicazione: (2024)
di: Saha, Anish, et al.
Pubblicazione: (2024)
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)
Improving the Inclusivity of Dutch Speech Recognition by Fine-tuning Whisper on the JASMIN-CGN Corpus
di: Shekoufandeh, Golshid, et al.
Pubblicazione: (2025)
di: Shekoufandeh, Golshid, et al.
Pubblicazione: (2025)
Benchmarking Automatic Speech Recognition Models for African Languages
di: Nahabwe, Alvin, et al.
Pubblicazione: (2025)
di: Nahabwe, Alvin, et al.
Pubblicazione: (2025)
Automatic Speech Recognition for Biomedical Data in Bengali Language
di: Kabir, Shariar, et al.
Pubblicazione: (2024)
di: Kabir, Shariar, et al.
Pubblicazione: (2024)
Scalable Offline ASR for Command-Style Dictation in Courtrooms
di: Nethil, Kumarmanas, et al.
Pubblicazione: (2025)
di: Nethil, Kumarmanas, et al.
Pubblicazione: (2025)
Supporting SENCOTEN Language Documentation Efforts with Automatic Speech Recognition
di: Geng, Mengzhe, et al.
Pubblicazione: (2025)
di: Geng, Mengzhe, et al.
Pubblicazione: (2025)
Fine-Tuning Automatic Speech Recognition for People with Parkinson's: An Effective Strategy for Enhancing Speech Technology Accessibility
di: Zheng, Xiuwen, et al.
Pubblicazione: (2024)
di: Zheng, Xiuwen, et al.
Pubblicazione: (2024)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
di: Abu, Turi, et al.
Pubblicazione: (2025)
di: Abu, Turi, et al.
Pubblicazione: (2025)
A Deep Learning Automatic Speech Recognition Model for Shona Language
di: Sirora, Leslie Wellington, et al.
Pubblicazione: (2025)
di: Sirora, Leslie Wellington, et al.
Pubblicazione: (2025)
Rasa: Building Expressive Speech Synthesis Systems for Indian Languages in Low-resource Settings
di: Varadhan, Praveen Srinivasa, et al.
Pubblicazione: (2024)
di: Varadhan, Praveen Srinivasa, et al.
Pubblicazione: (2024)
Dynamic Data Pruning for Automatic Speech Recognition
di: Xiao, Qiao, et al.
Pubblicazione: (2024)
di: Xiao, Qiao, et al.
Pubblicazione: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
di: Sudo, Yui, et al.
Pubblicazione: (2024)
di: Sudo, Yui, et al.
Pubblicazione: (2024)
LI-TTA: Language Informed Test-Time Adaptation for Automatic Speech Recognition
di: Yoon, Eunseop, et al.
Pubblicazione: (2024)
di: Yoon, Eunseop, et al.
Pubblicazione: (2024)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
di: Hu, Jiliang, et al.
Pubblicazione: (2025)
di: Hu, Jiliang, et al.
Pubblicazione: (2025)
Position-invariant Fine-tuning of Speech Enhancement Models with Self-supervised Speech Representations
di: Meghanani, Amit, et al.
Pubblicazione: (2026)
di: Meghanani, Amit, et al.
Pubblicazione: (2026)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
di: Shi, Hao, et al.
Pubblicazione: (2024)
di: Shi, Hao, et al.
Pubblicazione: (2024)
Exploring Gender Disparities in Automatic Speech Recognition Technology
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
Exploring the Integration of Large Language Models into Automatic Speech Recognition Systems: An Empirical Study
di: Min, Zeping, et al.
Pubblicazione: (2023)
di: Min, Zeping, et al.
Pubblicazione: (2023)
Automatic Speech Recognition of Non-Native Child Speech for Language Learning Applications
di: Wills, Simone, et al.
Pubblicazione: (2023)
di: Wills, Simone, et al.
Pubblicazione: (2023)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
di: Wang, Yujin, et al.
Pubblicazione: (2022)
di: Wang, Yujin, et al.
Pubblicazione: (2022)
USM-Lite: Quantization and Sparsity Aware Fine-tuning for Speech Recognition with Universal Speech Models
di: Ding, Shaojin, et al.
Pubblicazione: (2023)
di: Ding, Shaojin, et al.
Pubblicazione: (2023)
Methods to Increase the Amount of Data for Speech Recognition for Low Resource Languages
di: Ayrapetyan, Alexan, et al.
Pubblicazione: (2025)
di: Ayrapetyan, Alexan, et al.
Pubblicazione: (2025)
Weighted Cross-entropy for Low-Resource Languages in Multilingual Speech Recognition
di: Piñeiro-Martín, Andrés, et al.
Pubblicazione: (2024)
di: Piñeiro-Martín, Andrés, et al.
Pubblicazione: (2024)
Exploring Spoken Language Identification Strategies for Automatic Transcription of Multilingual Broadcast and Institutional Speech
di: Valente, Martina, et al.
Pubblicazione: (2024)
di: Valente, Martina, et al.
Pubblicazione: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
di: Park, ChaeHun, et al.
Pubblicazione: (2024)
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
di: Iakovenko, Olga, et al.
Pubblicazione: (2024)
di: Iakovenko, Olga, et al.
Pubblicazione: (2024)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
di: Sasu, David, et al.
Pubblicazione: (2025)
di: Sasu, David, et al.
Pubblicazione: (2025)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
di: Zhu, Han, et al.
Pubblicazione: (2024)
di: Zhu, Han, et al.
Pubblicazione: (2024)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
di: Hu, Ke, et al.
Pubblicazione: (2025)
di: Hu, Ke, et al.
Pubblicazione: (2025)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
di: Lin, Zhennan, et al.
Pubblicazione: (2025)
di: Lin, Zhennan, et al.
Pubblicazione: (2025)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
di: Guo, Jiaxin, et al.
Pubblicazione: (2024)
di: Guo, Jiaxin, et al.
Pubblicazione: (2024)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
di: Adila, Aulia, et al.
Pubblicazione: (2024)
di: Adila, Aulia, et al.
Pubblicazione: (2024)
Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages
di: Liang, Siyu, et al.
Pubblicazione: (2025)
di: Liang, Siyu, et al.
Pubblicazione: (2025)
Automatic Speech Recognition System-Independent Word Error Rate Estimation
di: Park, Chanho, et al.
Pubblicazione: (2024)
di: Park, Chanho, et al.
Pubblicazione: (2024)
Hallucinations in Neural Automatic Speech Recognition: Identifying Errors and Hallucinatory Models
di: Frieske, Rita, et al.
Pubblicazione: (2024)
di: Frieske, Rita, et al.
Pubblicazione: (2024)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
di: Shakeel, Muhammad, et al.
Pubblicazione: (2024)
di: Shakeel, Muhammad, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Tracking Articulatory Dynamics in Speech with a Fixed-Weight BiLSTM-CNN Architecture
di: Pillai, Leena G, et al.
Pubblicazione: (2025) -
Acoustic to Articulatory Inversion of Speech; Data Driven Approaches, Challenges, Applications, and Future Scope
di: Pillai, Leena G, et al.
Pubblicazione: (2025) -
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
di: Hsu, Ming-Hao, et al.
Pubblicazione: (2024) -
Automatic Speech Recognition for Hindi
di: Saha, Anish, et al.
Pubblicazione: (2024) -
Automatic Speech Recognition for African Low-Resource Languages: Challenges and Future Directions
di: Imam, Sukairaj Hafiz, et al.
Pubblicazione: (2025)