MunTTS: A Text-to-Speech System for Mundari
Fuente:
arXiv
Guardado en:
| Autores principales: | , , , , , , |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
| _version_ | 1866913213175562240 |
|---|---|
| author | Gumma, Varun Hada, Rishav Yadavalli, Aditya Gogoi, Pamir Mondal, Ishani Seshadri, Vivek Bali, Kalika |
| author_facet | Gumma, Varun Hada, Rishav Yadavalli, Aditya Gogoi, Pamir Mondal, Ishani Seshadri, Vivek Bali, Kalika |
| contents | We present MunTTS, an end-to-end text-to-speech (TTS) system specifically for Mundari, a low-resource Indian language of the Austo-Asiatic family. Our work addresses the gap in linguistic technology for underrepresented languages by collecting and processing data to build a speech synthesis system. We begin our study by gathering a substantial dataset of Mundari text and speech and train end-to-end speech models. We also delve into the methods used for training our models, ensuring they are efficient and effective despite the data constraints. We evaluate our system with native speakers and objective metrics, demonstrating its potential as a tool for preserving and promoting the Mundari language in the digital age. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2401_15579 |
| institution | arXiv |
| publishDate | 2024 |
| record_format | arxiv |
| spellingShingle | MunTTS: A Text-to-Speech System for Mundari Gumma, Varun Hada, Rishav Yadavalli, Aditya Gogoi, Pamir Mondal, Ishani Seshadri, Vivek Bali, Kalika Computation and Language Sound Audio and Speech Processing We present MunTTS, an end-to-end text-to-speech (TTS) system specifically for Mundari, a low-resource Indian language of the Austo-Asiatic family. Our work addresses the gap in linguistic technology for underrepresented languages by collecting and processing data to build a speech synthesis system. We begin our study by gathering a substantial dataset of Mundari text and speech and train end-to-end speech models. We also delve into the methods used for training our models, ensuring they are efficient and effective despite the data constraints. We evaluate our system with native speakers and objective metrics, demonstrating its potential as a tool for preserving and promoting the Mundari language in the digital age. |
| title | MunTTS: A Text-to-Speech System for Mundari |
| topic | Computation and Language Sound Audio and Speech Processing |
| url | https://arxiv.org/abs/2401.15579 |