Enabling ASR for Low-Resource Languages: A Comprehensive Dataset Creation Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yeroyan, Ara, Karpov, Nikolay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Methods to Increase the Amount of Data for Speech Recognition for Low Resource Languages
von: Ayrapetyan, Alexan, et al.
Veröffentlicht: (2025)
von: Ayrapetyan, Alexan, et al.
Veröffentlicht: (2025)
An ASR-Based Tutor for Learning to Read: How to Optimize Feedback to First Graders
von: Bai, Yu, et al.
Veröffentlicht: (2023)
von: Bai, Yu, et al.
Veröffentlicht: (2023)
Initial Decoding with Minimally Augmented Language Model for Improved Lattice Rescoring in Low Resource ASR
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
von: Murthy, Savitha, et al.
Veröffentlicht: (2024)
VoxKnesset: A Large-Scale Longitudinal Hebrew Speech Dataset for Aging Speaker Modeling
von: Marmor, Yanir, et al.
Veröffentlicht: (2026)
von: Marmor, Yanir, et al.
Veröffentlicht: (2026)
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection
von: Mandal, Atanu, et al.
Veröffentlicht: (2024)
von: Mandal, Atanu, et al.
Veröffentlicht: (2024)
Is Attention always needed? A Case Study on Language Identification from Speech
von: Mandal, Atanu, et al.
Veröffentlicht: (2021)
von: Mandal, Atanu, et al.
Veröffentlicht: (2021)
Bridging Auditory Perception and Language Comprehension through MEG-Driven Encoding Models
von: Ciferri, Matteo, et al.
Veröffentlicht: (2024)
von: Ciferri, Matteo, et al.
Veröffentlicht: (2024)
XLS-R Deep Learning Model for Multilingual ASR on Low- Resource Languages: Indonesian, Javanese, and Sundanese
von: Arisaputra, Panji, et al.
Veröffentlicht: (2024)
von: Arisaputra, Panji, et al.
Veröffentlicht: (2024)
Exploring Dynamic Parameters for Vietnamese Gender-Independent ASR
von: Leang, Sotheara, et al.
Veröffentlicht: (2025)
von: Leang, Sotheara, et al.
Veröffentlicht: (2025)
A Computational Approach to Analyzing Disrupted Language in Schizophrenia: Integrating Surprisal and Coherence Measures
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
von: Premananth, Gowtham, et al.
Veröffentlicht: (2025)
Alzheimer Disease Classification through ASR-based Transcriptions: Exploring the Impact of Punctuation and Pauses
von: Gómez-Zaragozá, Lucía, et al.
Veröffentlicht: (2023)
von: Gómez-Zaragozá, Lucía, et al.
Veröffentlicht: (2023)
Hybrid Deep Learning and Signal Processing for Arabic Dialect Recognition in Low-Resource Settings
von: Al-Shwayyat, Ghazal, et al.
Veröffentlicht: (2025)
von: Al-Shwayyat, Ghazal, et al.
Veröffentlicht: (2025)
Language Bias in Self-Supervised Learning For Automatic Speech Recognition
von: Storey, Edward, et al.
Veröffentlicht: (2025)
von: Storey, Edward, et al.
Veröffentlicht: (2025)
Beyond the Labels: Unveiling Text-Dependency in Paralinguistic Speech Recognition Datasets
von: Pešán, Jan, et al.
Veröffentlicht: (2024)
von: Pešán, Jan, et al.
Veröffentlicht: (2024)
Ultra Low Complexity Deep Learning Based Noise Suppression
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2023)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2023)
TaigiSpeech: A Low-Resource Real-World Speech Intent Dataset and Preliminary Results with Scalable Data Mining In-the-Wild
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2026)
von: Chang, Kai-Wei, et al.
Veröffentlicht: (2026)
TRI-DEP: A Trimodal Comparative Study for Depression Detection Using Speech, Text, and EEG
von: Nurfidausi, Annisaa Fitri, et al.
Veröffentlicht: (2025)
von: Nurfidausi, Annisaa Fitri, et al.
Veröffentlicht: (2025)
Neuro2Semantic: A Transfer Learning Framework for Semantic Reconstruction of Continuous Language from Human Intracranial EEG
von: Shams, Siavash, et al.
Veröffentlicht: (2025)
von: Shams, Siavash, et al.
Veröffentlicht: (2025)
Bottleneck Transformer-Based Approach for Improved Automatic STOI Score Prediction
von: Amartyaveer, et al.
Veröffentlicht: (2026)
von: Amartyaveer, et al.
Veröffentlicht: (2026)
Efficient ASR for Low-Resource Languages: Leveraging Cross-Lingual Unlabeled Data
von: Bandarupalli, Srihari, et al.
Veröffentlicht: (2025)
von: Bandarupalli, Srihari, et al.
Veröffentlicht: (2025)
Wavelet GPT: Wavelet Inspired Large Language Models
von: Verma, Prateek
Veröffentlicht: (2024)
von: Verma, Prateek
Veröffentlicht: (2024)
Resource-Efficient Separation Transformer
von: Della Libera, Luca, et al.
Veröffentlicht: (2022)
von: Della Libera, Luca, et al.
Veröffentlicht: (2022)
RIR-Mega-Speech: A Reverberant Speech Corpus with Comprehensive Acoustic Metadata and Reproducible Evaluation
von: Goswami, Mandip
Veröffentlicht: (2026)
von: Goswami, Mandip
Veröffentlicht: (2026)
Evaluating Standard and Dialectal Frisian ASR: Multilingual Fine-tuning and Language Identification for Improved Low-resource Performance
von: Amooie, Reihaneh, et al.
Veröffentlicht: (2025)
von: Amooie, Reihaneh, et al.
Veröffentlicht: (2025)
Two-component spatiotemporal template for activation-inhibition of speech in ECoG
von: Easthope, Eric
Veröffentlicht: (2024)
von: Easthope, Eric
Veröffentlicht: (2024)
Reading Miscue Detection in Primary School through Automatic Speech Recognition
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
von: Gao, Lingyun, et al.
Veröffentlicht: (2024)
Dilated CNNs for Periodic Signal Processing: A Low-Complexity Approach
von: Gildish, Eli, et al.
Veröffentlicht: (2026)
von: Gildish, Eli, et al.
Veröffentlicht: (2026)
Intelligent Fault Diagnosis of Type and Severity in Low-Frequency, Low Bit-Depth Signals
von: Spadini, Tito, et al.
Veröffentlicht: (2024)
von: Spadini, Tito, et al.
Veröffentlicht: (2024)
BUET Multi-disease Heart Sound Dataset: A Comprehensive Auscultation Dataset for Developing Computer-Aided Diagnostic Systems
von: Ali, Shams Nafisa, et al.
Veröffentlicht: (2024)
von: Ali, Shams Nafisa, et al.
Veröffentlicht: (2024)
Efficient Adapter Finetuning for Tail Languages in Streaming Multilingual ASR
von: Bai, Junwen, et al.
Veröffentlicht: (2024)
von: Bai, Junwen, et al.
Veröffentlicht: (2024)
Benchmarking Akan ASR Models Across Domain-Specific Datasets: A Comparative Evaluation of Performance, Scalability, and Adaptability
von: Mensah, Mark Atta, et al.
Veröffentlicht: (2025)
von: Mensah, Mark Atta, et al.
Veröffentlicht: (2025)
A Practitioner's Guide to Building ASR Models for Low-Resource Languages: A Case Study on Scottish Gaelic
von: Klejch, Ondřej, et al.
Veröffentlicht: (2025)
von: Klejch, Ondřej, et al.
Veröffentlicht: (2025)
A Large-Scale Evaluation of Speech Foundation Models
von: Yang, Shu-wen, et al.
Veröffentlicht: (2024)
von: Yang, Shu-wen, et al.
Veröffentlicht: (2024)
Learn and Don't Forget: Adding a New Language to ASR Foundation Models
von: Qian, Mengjie, et al.
Veröffentlicht: (2024)
von: Qian, Mengjie, et al.
Veröffentlicht: (2024)
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
von: Dhawan, Kunal, et al.
Veröffentlicht: (2024)
Canary-1B-v2 & Parakeet-TDT-0.6B-v3: Efficient and High-Performance Models for Multilingual ASR and AST
von: Sekoyan, Monica, et al.
Veröffentlicht: (2025)
von: Sekoyan, Monica, et al.
Veröffentlicht: (2025)
A Physics-Informed Neural Network-Based Approach for the Spatial Upsampling of Spherical Microphone Arrays
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
von: Miotello, Federico, et al.
Veröffentlicht: (2024)
DeepFilterGAN: A Full-band Real-time Speech Enhancement System with GAN-based Stochastic Regeneration
von: Serbest, Sanberk, et al.
Veröffentlicht: (2025)
von: Serbest, Sanberk, et al.
Veröffentlicht: (2025)
TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
von: Song, Xingchen, et al.
Veröffentlicht: (2024)
Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
von: Kim, Minsu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Methods to Increase the Amount of Data for Speech Recognition for Low Resource Languages
von: Ayrapetyan, Alexan, et al.
Veröffentlicht: (2025) -
An ASR-Based Tutor for Learning to Read: How to Optimize Feedback to First Graders
von: Bai, Yu, et al.
Veröffentlicht: (2023) -
Initial Decoding with Minimally Augmented Language Model for Improved Lattice Rescoring in Low Resource ASR
von: Murthy, Savitha, et al.
Veröffentlicht: (2024) -
VoxKnesset: A Large-Scale Longitudinal Hebrew Speech Dataset for Aging Speaker Modeling
von: Marmor, Yanir, et al.
Veröffentlicht: (2026) -
Attentive Fusion: A Transformer-based Approach to Multimodal Hate Speech Detection
von: Mandal, Atanu, et al.
Veröffentlicht: (2024)