Automatic Speech Recognition for Documenting Endangered Languages: Case Study of Ikema Miyakoan
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Taguchi, Chihiro, Takubo, Yukinori, Chiang, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Complexity and Speech Recognition Accuracy: Orthographic Complexity Hurts, Phonological Complexity Doesn't
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024)
Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024)
In-context Language Learning for Endangered Languages in Speech Recognition
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
von: Li, Zhaolin, et al.
Veröffentlicht: (2025)
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
NoLoR: An ASR-Based Framework for Expedited Endangered Language Documentation with Neo-Aramaic as a Case Study
von: Nazari, Matthew
Veröffentlicht: (2024)
von: Nazari, Matthew
Veröffentlicht: (2024)
Error-preserving Automatic Speech Recognition of Young English Learners' Language
von: Michot, Janick, et al.
Veröffentlicht: (2024)
von: Michot, Janick, et al.
Veröffentlicht: (2024)
Efficient Context Selection for Long-Context QA: No Tuning, No Iteration, Just Adaptive-$k$
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Sanskrit with Transfer Learning
von: Sadhukhan, Bidit, et al.
Veröffentlicht: (2025)
von: Sadhukhan, Bidit, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Greek Medical Dictation
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
von: Georgilas, Vardis, et al.
Veröffentlicht: (2025)
Multi-Stage Multi-Modal Pre-Training for Automatic Speech Recognition
von: Jain, Yash, et al.
Veröffentlicht: (2024)
von: Jain, Yash, et al.
Veröffentlicht: (2024)
LoASR-Bench: Evaluating Large Speech Language Models on Low-Resource Automatic Speech Recognition Across Language Families
von: Chen, Jianan, et al.
Veröffentlicht: (2026)
von: Chen, Jianan, et al.
Veröffentlicht: (2026)
Augmenting Automatic Speech Recognition Models with Disfluency Detection
von: Amann, Robin, et al.
Veröffentlicht: (2024)
von: Amann, Robin, et al.
Veröffentlicht: (2024)
Fairness of Automatic Speech Recognition: Looking Through a Philosophical Lens
von: Choi, Anna Seo Gyeong, et al.
Veröffentlicht: (2025)
von: Choi, Anna Seo Gyeong, et al.
Veröffentlicht: (2025)
Doing More with Less: Data Augmentation for Sudanese Dialect Automatic Speech Recognition
von: Mansour, Ayman
Veröffentlicht: (2026)
von: Mansour, Ayman
Veröffentlicht: (2026)
A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain
von: Obaidah, Qusai Abo, et al.
Veröffentlicht: (2024)
von: Obaidah, Qusai Abo, et al.
Veröffentlicht: (2024)
Benchmarking Automatic Speech Recognition for Indian Languages in Agricultural Contexts
von: S, Chandrashekar M, et al.
Veröffentlicht: (2026)
von: S, Chandrashekar M, et al.
Veröffentlicht: (2026)
Speech Retrieval-Augmented Generation without Automatic Speech Recognition
von: Min, Do June, et al.
Veröffentlicht: (2024)
von: Min, Do June, et al.
Veröffentlicht: (2024)
Handling Numeric Expressions in Automatic Speech Recognition
von: Huber, Christian, et al.
Veröffentlicht: (2024)
von: Huber, Christian, et al.
Veröffentlicht: (2024)
SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition
von: Dkhissi, Youness, et al.
Veröffentlicht: (2026)
von: Dkhissi, Youness, et al.
Veröffentlicht: (2026)
Benchmarking Rotary Position Embeddings for Automatic Speech Recognition
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
von: Zhang, Shucong, et al.
Veröffentlicht: (2025)
Multistage Fine-tuning Strategies for Automatic Speech Recognition in Low-resource Languages
von: Pillai, Leena G, et al.
Veröffentlicht: (2024)
von: Pillai, Leena G, et al.
Veröffentlicht: (2024)
Bias Vector: Mitigating Biases in Language Models with Task Arithmetic Approach
von: Shirafuji, Daiki, et al.
Veröffentlicht: (2024)
von: Shirafuji, Daiki, et al.
Veröffentlicht: (2024)
ASKD-Whisper: Adaptive Self-knowledge Distillation for Efficient and Low-Latency Automatic Speech Recognition
von: Lee, Junseok, et al.
Veröffentlicht: (2026)
von: Lee, Junseok, et al.
Veröffentlicht: (2026)
Semantically Corrected Amharic Automatic Speech Recognition
von: Adnew, Samuael, et al.
Veröffentlicht: (2024)
von: Adnew, Samuael, et al.
Veröffentlicht: (2024)
Improved Contextual Recognition In Automatic Speech Recognition Systems By Semantic Lattice Rescoring
von: Sudarshan, Ankitha, et al.
Veröffentlicht: (2023)
von: Sudarshan, Ankitha, et al.
Veröffentlicht: (2023)
Towards Unsupervised Speech Recognition at the Syllable-Level
von: Wang, Liming, et al.
Veröffentlicht: (2025)
von: Wang, Liming, et al.
Veröffentlicht: (2025)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
von: Lehečka, Jan, et al.
Veröffentlicht: (2024)
von: Lehečka, Jan, et al.
Veröffentlicht: (2024)
Automatic Summarization of Long Documents
von: Chhibbar, Naman, et al.
Veröffentlicht: (2024)
von: Chhibbar, Naman, et al.
Veröffentlicht: (2024)
WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data
von: Zhang, Ziheng, et al.
Veröffentlicht: (2026)
von: Zhang, Ziheng, et al.
Veröffentlicht: (2026)
Harnessing the Power of Artificial Intelligence to Vitalize Endangered Indigenous Languages: Technologies and Experiences
von: Pinhanez, Claudio, et al.
Veröffentlicht: (2024)
von: Pinhanez, Claudio, et al.
Veröffentlicht: (2024)
TICL+: A Case Study On Speech In-Context Learning for Children's Speech Recognition
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
von: Zheng, Haolong, et al.
Veröffentlicht: (2025)
CDGP: Automatic Cloze Distractor Generation based on Pre-trained Language Model
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
Learning from Child-Directed Speech in Two-Language Scenarios: A French-English Case Study
von: Binyamin, Liel, et al.
Veröffentlicht: (2026)
von: Binyamin, Liel, et al.
Veröffentlicht: (2026)
Reasoning Transfer for an Extremely Low-Resource and Endangered Language: Bridging Languages Through Sample-Efficient Language Understanding
von: Tran, Khanh-Tung, et al.
Veröffentlicht: (2025)
von: Tran, Khanh-Tung, et al.
Veröffentlicht: (2025)
It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition
von: Chen, Chen, et al.
Veröffentlicht: (2024)
von: Chen, Chen, et al.
Veröffentlicht: (2024)
VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain
von: Le-Duc, Khai
Veröffentlicht: (2024)
von: Le-Duc, Khai
Veröffentlicht: (2024)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
Integrating Linguistics and AI: Morphological Analysis and Corpus development of Endangered Toto Language of West Bengal
von: Guha, Ambalika, et al.
Veröffentlicht: (2025)
von: Guha, Ambalika, et al.
Veröffentlicht: (2025)
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers
von: Serai, Prashant, et al.
Veröffentlicht: (2024)
von: Serai, Prashant, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Language Complexity and Speech Recognition Accuracy: Orthographic Complexity Hurts, Phonological Complexity Doesn't
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024) -
Killkan: The Automatic Speech Recognition Dataset for Kichwa with Morphosyntactic Information
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2024) -
In-context Language Learning for Endangered Languages in Speech Recognition
von: Li, Zhaolin, et al.
Veröffentlicht: (2025) -
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025) -
NoLoR: An ASR-Based Framework for Expedited Endangered Language Documentation with Neo-Aramaic as a Case Study
von: Nazari, Matthew
Veröffentlicht: (2024)