Mići Princ -- A Little Boy Teaching Speech Technologies the Chakavian Dialect
Fuente:
arXiv
Salvato in:
| Autori principali: | Ljubešić, Nikola, Rupnik, Peter, Perinčić, Tea |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Identifying Primary Stress Across Related Languages and Dialects with Transformer-based Speech Encoder Models
di: Ljubešić, Nikola, et al.
Pubblicazione: (2025)
di: Ljubešić, Nikola, et al.
Pubblicazione: (2025)
The ParlaSpeech Collection of Automatically Generated Speech and Text Datasets from Parliamentary Proceedings
di: Ljubešić, Nikola, et al.
Pubblicazione: (2024)
di: Ljubešić, Nikola, et al.
Pubblicazione: (2024)
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
di: Blaschke, Verena, et al.
Pubblicazione: (2025)
di: Blaschke, Verena, et al.
Pubblicazione: (2025)
Dialectal Coverage And Generalization in Arabic Speech Recognition
di: Djanibekov, Amirbek, et al.
Pubblicazione: (2024)
di: Djanibekov, Amirbek, et al.
Pubblicazione: (2024)
Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT
di: Yamauchi, Kazuki, et al.
Pubblicazione: (2024)
di: Yamauchi, Kazuki, et al.
Pubblicazione: (2024)
Towards Zero-Shot Text-To-Speech for Arabic Dialects
di: Doan, Khai Duy, et al.
Pubblicazione: (2024)
di: Doan, Khai Duy, et al.
Pubblicazione: (2024)
Dolphin-CN-Dialect: Where Chinese Dialects Matter
di: Meng, Yangyang, et al.
Pubblicazione: (2026)
di: Meng, Yangyang, et al.
Pubblicazione: (2026)
RoDia: A New Dataset for Romanian Dialect Identification from Speech
di: Rotaru, Codrut, et al.
Pubblicazione: (2023)
di: Rotaru, Codrut, et al.
Pubblicazione: (2023)
Habibi: Laying the Open-Source Foundation of Unified-Dialectal Arabic Speech Synthesis
di: Chen, Yushen, et al.
Pubblicazione: (2026)
di: Chen, Yushen, et al.
Pubblicazione: (2026)
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
di: Mdhaffar, Salima, et al.
Pubblicazione: (2024)
di: Mdhaffar, Salima, et al.
Pubblicazione: (2024)
Voxlect: A Speech Foundation Model Benchmark for Modeling Dialects and Regional Languages Around the Globe
di: Feng, Tiantian, et al.
Pubblicazione: (2025)
di: Feng, Tiantian, et al.
Pubblicazione: (2025)
Bailing-TTS: Chinese Dialectal Speech Synthesis Towards Human-like Spontaneous Representation
di: Di, Xinhan, et al.
Pubblicazione: (2024)
di: Di, Xinhan, et al.
Pubblicazione: (2024)
Leveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative Analysis
di: Xu, Tianyi, et al.
Pubblicazione: (2025)
di: Xu, Tianyi, et al.
Pubblicazione: (2025)
LinTO Audio and Textual Datasets to Train and Evaluate Automatic Speech Recognition in Tunisian Arabic Dialect
di: Naouara, Hedi, et al.
Pubblicazione: (2025)
di: Naouara, Hedi, et al.
Pubblicazione: (2025)
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects
di: Yang, Sicheng, et al.
Pubblicazione: (2026)
di: Yang, Sicheng, et al.
Pubblicazione: (2026)
Exploring Energy-Based Models for Out-of-Distribution Detection in Dialect Identification
di: Hao, Yaqian, et al.
Pubblicazione: (2024)
di: Hao, Yaqian, et al.
Pubblicazione: (2024)
Learning Speech Representations with Variational Predictive Coding
di: Yeh, Sung-Lin, et al.
Pubblicazione: (2025)
di: Yeh, Sung-Lin, et al.
Pubblicazione: (2025)
Exploring Gender Disparities in Automatic Speech Recognition Technology
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
di: ElGhazaly, Hend, et al.
Pubblicazione: (2025)
CAFE A Novel Code switching Dataset for Algerian Dialect French and English
di: Lachemat, Houssam Eddine-Othman, et al.
Pubblicazione: (2024)
di: Lachemat, Houssam Eddine-Othman, et al.
Pubblicazione: (2024)
iMiGUE-Speech: A Spontaneous Speech Dataset for Affective Analysis
di: Kakouros, Sofoklis, et al.
Pubblicazione: (2026)
di: Kakouros, Sofoklis, et al.
Pubblicazione: (2026)
Streaming Speech-to-Confusion Network Speech Recognition
di: Filimonov, Denis, et al.
Pubblicazione: (2023)
di: Filimonov, Denis, et al.
Pubblicazione: (2023)
MAD Speech: Measures of Acoustic Diversity of Speech
di: Futeral, Matthieu, et al.
Pubblicazione: (2024)
di: Futeral, Matthieu, et al.
Pubblicazione: (2024)
VoxHakka: A Dialectally Diverse Multi-speaker Text-to-Speech System for Taiwanese Hakka
di: Chen, Li-Wei, et al.
Pubblicazione: (2024)
di: Chen, Li-Wei, et al.
Pubblicazione: (2024)
Textless Speech-to-Speech Translation With Limited Parallel Data
di: Diwan, Anuj, et al.
Pubblicazione: (2023)
di: Diwan, Anuj, et al.
Pubblicazione: (2023)
Efficient Dialect-Aware Modeling and Conditioning for Low-Resource Taiwanese Hakka Speech Processing
di: Peng, An-Ci, et al.
Pubblicazione: (2026)
di: Peng, An-Ci, et al.
Pubblicazione: (2026)
DisfluencySpeech -- Single-Speaker Conversational Speech Dataset with Paralanguage
di: Wang, Kyra, et al.
Pubblicazione: (2024)
di: Wang, Kyra, et al.
Pubblicazione: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
di: Abdullah, Badr M., et al.
Pubblicazione: (2025)
di: Abdullah, Badr M., et al.
Pubblicazione: (2025)
A Large Dataset of Spontaneous Speech with the Accent Spoken in São Paulo for Automatic Speech Recognition Evaluation
di: Lima, Rodrigo, et al.
Pubblicazione: (2024)
di: Lima, Rodrigo, et al.
Pubblicazione: (2024)
Teaching a Multilingual Large Language Model to Understand Multilingual Speech via Multi-Instructional Training
di: Denisov, Pavel, et al.
Pubblicazione: (2024)
di: Denisov, Pavel, et al.
Pubblicazione: (2024)
Speech Separation based on Contrastive Learning and Deep Modularization
di: Ochieng, Peter
Pubblicazione: (2023)
di: Ochieng, Peter
Pubblicazione: (2023)
SpeechColab Leaderboard: An Open-Source Platform for Automatic Speech Recognition Evaluation
di: Du, Jiayu, et al.
Pubblicazione: (2024)
di: Du, Jiayu, et al.
Pubblicazione: (2024)
Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation
di: He, Haorui, et al.
Pubblicazione: (2024)
di: He, Haorui, et al.
Pubblicazione: (2024)
Overcoming Data Scarcity in Multi-Dialectal Arabic ASR via Whisper Fine-Tuning
di: Özyilmaz, Ömer Tarik, et al.
Pubblicazione: (2025)
di: Özyilmaz, Ömer Tarik, et al.
Pubblicazione: (2025)
A Preliminary Analysis of Automatic Word and Syllable Prominence Detection in Non-Native Speech With Text-to-Speech Prosody Embeddings
di: Mondal, Anindita, et al.
Pubblicazione: (2024)
di: Mondal, Anindita, et al.
Pubblicazione: (2024)
DeSTA: Enhancing Speech Language Models through Descriptive Speech-Text Alignment
di: Lu, Ke-Han, et al.
Pubblicazione: (2024)
di: Lu, Ke-Han, et al.
Pubblicazione: (2024)
Towards Efficient Speech-Text Jointly Decoding within One Speech Language Model
di: Wu, Haibin, et al.
Pubblicazione: (2025)
di: Wu, Haibin, et al.
Pubblicazione: (2025)
Speech Recognition Model Improves Text-to-Speech Synthesis using Fine-Grained Reward
di: Wang, Guansu, et al.
Pubblicazione: (2025)
di: Wang, Guansu, et al.
Pubblicazione: (2025)
Emphasis Sensitivity in Speech Representations
di: Cassini, Shaun, et al.
Pubblicazione: (2025)
di: Cassini, Shaun, et al.
Pubblicazione: (2025)
Speech-IFEval: Evaluating Instruction-Following and Quantifying Catastrophic Forgetting in Speech-Aware Language Models
di: Lu, Ke-Han, et al.
Pubblicazione: (2025)
di: Lu, Ke-Han, et al.
Pubblicazione: (2025)
Evaluating Speech-to-Text x LLM x Text-to-Speech Combinations for AI Interview Systems
di: Allbert, Rumi, et al.
Pubblicazione: (2025)
di: Allbert, Rumi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Identifying Primary Stress Across Related Languages and Dialects with Transformer-based Speech Encoder Models
di: Ljubešić, Nikola, et al.
Pubblicazione: (2025) -
The ParlaSpeech Collection of Automatically Generated Speech and Text Datasets from Parliamentary Proceedings
di: Ljubešić, Nikola, et al.
Pubblicazione: (2024) -
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
di: Blaschke, Verena, et al.
Pubblicazione: (2025) -
Dialectal Coverage And Generalization in Arabic Speech Recognition
di: Djanibekov, Amirbek, et al.
Pubblicazione: (2024) -
Cross-Dialect Text-To-Speech in Pitch-Accent Language Incorporating Multi-Dialect Phoneme-Level BERT
di: Yamauchi, Kazuki, et al.
Pubblicazione: (2024)