LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Naouara, Hedi, Lorré, Jean-Pierre, Louradour, Jérôme |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024)
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
von: Mdhaffar, Salima, et al.
Veröffentlicht: (2024)
von: Mdhaffar, Salima, et al.
Veröffentlicht: (2024)
Towards Zero-Shot Text-To-Speech for Arabic Dialects
von: Doan, Khai Duy, et al.
Veröffentlicht: (2024)
von: Doan, Khai Duy, et al.
Veröffentlicht: (2024)
Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
von: Chen, Yushen, et al.
Veröffentlicht: (2026)
Leveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative Analysis
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)
Sagalee: an Open Source Automatic Speech Recognition Dataset for Oromo Language
von: Abu, Turi, et al.
Veröffentlicht: (2025)
von: Abu, Turi, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Hindi
von: Saha, Anish, et al.
Veröffentlicht: (2024)
von: Saha, Anish, et al.
Veröffentlicht: (2024)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
SAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech Recognition
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
von: Farooq, Muhammad Umar, et al.
Veröffentlicht: (2025)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
von: Adila, Aulia, et al.
Veröffentlicht: (2024)
Hybrid Deep Learning and Signal Processing for Arabic Dialect Recognition in Low-Resource Settings
von: Al-Shwayyat, Ghazal, et al.
Veröffentlicht: (2025)
von: Al-Shwayyat, Ghazal, et al.
Veröffentlicht: (2025)
RoDia: A New Dataset for Romanian Dialect Identification from Speech
von: Rotaru, Codrut, et al.
Veröffentlicht: (2023)
von: Rotaru, Codrut, et al.
Veröffentlicht: (2023)
How to Evaluate Automatic Speech Recognition: Comparing Different Performance and Bias Measures
von: Patel, Tanvina, et al.
Veröffentlicht: (2025)
von: Patel, Tanvina, et al.
Veröffentlicht: (2025)
Dynamic Data Pruning for Automatic Speech Recognition
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
von: Xiao, Qiao, et al.
Veröffentlicht: (2024)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
von: Sudo, Yui, et al.
Veröffentlicht: (2024)
YODAS: Youtube-Oriented Dataset for Audio and Speech
von: Li, Xinjian, et al.
Veröffentlicht: (2024)
von: Li, Xinjian, et al.
Veröffentlicht: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
von: Abdullah, Badr M., et al.
Veröffentlicht: (2025)
Joint Automatic Speech Recognition And Structure Learning For Better Speech Understanding
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
von: Hu, Jiliang, et al.
Veröffentlicht: (2025)
Benchmarking Automatic Speech Recognition Models for African Languages
von: Nahabwe, Alvin, et al.
Veröffentlicht: (2025)
von: Nahabwe, Alvin, et al.
Veröffentlicht: (2025)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Exploring Gender Disparities in Automatic Speech Recognition Technology
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
von: ElGhazaly, Hend, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Biomedical Data in Bengali Language
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
von: Kabir, Shariar, et al.
Veröffentlicht: (2024)
Overcoming Data Scarcity in Multi-Dialectal Arabic ASR via Whisper Fine-Tuning
von: Özyilmaz, Ömer Tarik, et al.
Veröffentlicht: (2025)
von: Özyilmaz, Ömer Tarik, et al.
Veröffentlicht: (2025)
Exploring Effective Distillation of Self-Supervised Speech Models for Automatic Speech Recognition
von: Wang, Yujin, et al.
Veröffentlicht: (2022)
von: Wang, Yujin, et al.
Veröffentlicht: (2022)
Convolutional Variational Autoencoders for Spectrogram Compression in Automatic Speech Recognition
von: Iakovenko, Olga, et al.
Veröffentlicht: (2024)
von: Iakovenko, Olga, et al.
Veröffentlicht: (2024)
Supporting SENCOTEN Language Documentation Efforts with Automatic Speech Recognition
von: Geng, Mengzhe, et al.
Veröffentlicht: (2025)
von: Geng, Mengzhe, et al.
Veröffentlicht: (2025)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
von: Sasu, David, et al.
Veröffentlicht: (2025)
von: Sasu, David, et al.
Veröffentlicht: (2025)
Transliterated Zero-Shot Domain Adaptation for Automatic Speech Recognition
von: Zhu, Han, et al.
Veröffentlicht: (2024)
von: Zhu, Han, et al.
Veröffentlicht: (2024)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
von: Hu, Ke, et al.
Veröffentlicht: (2025)
von: Hu, Ke, et al.
Veröffentlicht: (2025)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
von: Hsu, Ming-Hao, et al.
Veröffentlicht: (2024)
Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT
von: Yamauchi, Kazuki, et al.
Veröffentlicht: (2024)
von: Yamauchi, Kazuki, et al.
Veröffentlicht: (2024)
From Statistical Methods to Pre-Trained Models; A Survey on Automatic Speech Recognition for Resource Scarce Urdu Language
von: Sharif, Muhammad, et al.
Veröffentlicht: (2024)
von: Sharif, Muhammad, et al.
Veröffentlicht: (2024)
StoryTTS: A Highly Expressive Text-to-Speech Dataset with Rich Textual Expressiveness Annotations
von: Liu, Sen, et al.
Veröffentlicht: (2024)
von: Liu, Sen, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition System-Independent Word Error Rate Estimation
von: Park, Chanho, et al.
Veröffentlicht: (2024)
von: Park, Chanho, et al.
Veröffentlicht: (2024)
Hallucinations in Neural Automatic Speech Recognition: Identifying Errors and Hallucinatory Models
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
Contextualized End-to-end Automatic Speech Recognition with Intermediate Biasing Loss
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
von: Shakeel, Muhammad, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition for Non-Native English: Accuracy and Disfluency Handling
von: McGuire, Michael
Veröffentlicht: (2025)
von: McGuire, Michael
Veröffentlicht: (2025)
A Deep Learning Automatic Speech Recognition Model for Shona Language
von: Sirora, Leslie Wellington, et al.
Veröffentlicht: (2025)
von: Sirora, Leslie Wellington, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dialectal Coverage And Generalization in Arabic Speech Recognition
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2024) -
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
von: Mdhaffar, Salima, et al.
Veröffentlicht: (2024) -
Towards Zero-Shot Text-To-Speech for Arabic Dialects
von: Doan, Khai Duy, et al.
Veröffentlicht: (2024) -
Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis
von: Chen, Yushen, et al.
Veröffentlicht: (2026) -
Leveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative Analysis
von: Xu, Tianyi, et al.
Veröffentlicht: (2025)