MunTTS: A Text-to-Speech System for Mundari

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Gumma, Varun, Hada, Rishav, Yadavalli, Aditya, Gogoi, Pamir, Mondal, Ishani, Seshadri, Vivek, Bali, Kalika
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866913213175562240
author Gumma, Varun
Hada, Rishav
Yadavalli, Aditya
Gogoi, Pamir
Mondal, Ishani
Seshadri, Vivek
Bali, Kalika
author_facet Gumma, Varun
Hada, Rishav
Yadavalli, Aditya
Gogoi, Pamir
Mondal, Ishani
Seshadri, Vivek
Bali, Kalika
contents We present MunTTS, an end-to-end text-to-speech (TTS) system specifically for Mundari, a low-resource Indian language of the Austo-Asiatic family. Our work addresses the gap in linguistic technology for underrepresented languages by collecting and processing data to build a speech synthesis system. We begin our study by gathering a substantial dataset of Mundari text and speech and train end-to-end speech models. We also delve into the methods used for training our models, ensuring they are efficient and effective despite the data constraints. We evaluate our system with native speakers and objective metrics, demonstrating its potential as a tool for preserving and promoting the Mundari language in the digital age.
format Preprint
id arxiv_https___arxiv_org_abs_2401_15579
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle MunTTS: A Text-to-Speech System for Mundari
Gumma, Varun
Hada, Rishav
Yadavalli, Aditya
Gogoi, Pamir
Mondal, Ishani
Seshadri, Vivek
Bali, Kalika
Computation and Language
Sound
Audio and Speech Processing
We present MunTTS, an end-to-end text-to-speech (TTS) system specifically for Mundari, a low-resource Indian language of the Austo-Asiatic family. Our work addresses the gap in linguistic technology for underrepresented languages by collecting and processing data to build a speech synthesis system. We begin our study by gathering a substantial dataset of Mundari text and speech and train end-to-end speech models. We also delve into the methods used for training our models, ensuring they are efficient and effective despite the data constraints. We evaluate our system with native speakers and objective metrics, demonstrating its potential as a tool for preserving and promoting the Mundari language in the digital age.
title MunTTS: A Text-to-Speech System for Mundari
topic Computation and Language
Sound
Audio and Speech Processing
url https://arxiv.org/abs/2401.15579