IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages
Fuente:
arXiv
Guardado en:
| Autores principales: | Javed, Tahir, Nawale, Janki Atul, George, Eldho Ittan, Joshi, Sakshi, Bhogale, Kaushal Santosh, Mehendale, Deovrat, Sethi, Ishvinder Virender, Ananthanarayanan, Aparna, Faquih, Hafsah, Palit, Pratiti, Ravishankar, Sneha, Sukumaran, Saranya, Panchagnula, Tripura, Murali, Sunjay, Gandhi, Kunal Sharad, R, Ambujavalli, M, Manickam K, Vaijayanthi, C Venkata, Karunganni, Krishnan Srinivasa Raghavan, Kumar, Pratyush, Khapra, Mitesh M |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
por: Javed, Tahir, et al.
Publicado: (2024)
por: Javed, Tahir, et al.
Publicado: (2024)
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
por: Sankar, Ashwin, et al.
Publicado: (2024)
por: Sankar, Ashwin, et al.
Publicado: (2024)
Recognizing Every Voice: Towards Inclusive ASR for Rural Bhojpuri Women
por: Joshi, Sakshi, et al.
Publicado: (2025)
por: Joshi, Sakshi, et al.
Publicado: (2025)
Empowering Low-Resource Language ASR via Large-Scale Pseudo Labeling
por: Bhogale, Kaushal Santosh, et al.
Publicado: (2024)
por: Bhogale, Kaushal Santosh, et al.
Publicado: (2024)
NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
por: Javed, Tahir, et al.
Publicado: (2025)
por: Javed, Tahir, et al.
Publicado: (2025)
FairI Tales: Evaluation of Fairness in Indian Contexts with a Focus on Bias and Stereotypes
por: Nawale, Janki Atul, et al.
Publicado: (2025)
por: Nawale, Janki Atul, et al.
Publicado: (2025)
IndicDLP: A Foundational Dataset for Multi-Lingual and Multi-Domain Document Layout Parsing
por: Nath, Oikantik, et al.
Publicado: (2025)
por: Nath, Oikantik, et al.
Publicado: (2025)
Towards Orthographically-Informed Evaluation of Speech Recognition Systems for Indian Languages
por: Bhogale, Kaushal Santosh, et al.
Publicado: (2026)
por: Bhogale, Kaushal Santosh, et al.
Publicado: (2026)
Digital Worlds, Mobile Lives: How ICTs Reshape Migration and Diaspora Communities
por: Tripura, Himadri
Publicado: (2018)
por: Tripura, Himadri
Publicado: (2018)
IndicLLMSuite: A Blueprint for Creating Pre-training and Fine-Tuning Datasets for Indian Languages
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2024)
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2024)
ELAICHI: Enhancing Low-resource TTS by Addressing Infrequent and Low-frequency Character Bigrams
por: Anand, Srija, et al.
Publicado: (2024)
por: Anand, Srija, et al.
Publicado: (2024)
Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2025)
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2025)
Rasa: Building Expressive Speech Synthesis Systems for Indian Languages in Low-resource Settings
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2024)
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2024)
Desarrollo de un pseudo-elemento de interfaz para el modelado de mampostería de ladrillo reforzado
por: S. Mehendale
Publicado: (2017)
por: S. Mehendale
Publicado: (2017)
DreamNet: A Multimodal Framework for Semantic and Emotional Analysis of Sleep Narratives
por: Panchagnula, Tapasvi
Publicado: (2025)
por: Panchagnula, Tapasvi
Publicado: (2025)
A Group Theoretic Construction of Batch Codes
por: Thomas, Eldho K.
Publicado: (2025)
por: Thomas, Eldho K.
Publicado: (2025)
Foraging with the Eyes: Dynamics in Human Visual Gaze and Deep Predictive Modeling
por: Panchagnula, Tejaswi V.
Publicado: (2025)
por: Panchagnula, Tejaswi V.
Publicado: (2025)
Enhancing Out-of-Vocabulary Performance of Indian TTS Systems for Practical Applications through Low-Effort Data Strategies
por: Anand, Srija, et al.
Publicado: (2024)
por: Anand, Srija, et al.
Publicado: (2024)
Can Vision-Language Models Evaluate Handwritten Math?
por: Nath, Oikantik, et al.
Publicado: (2025)
por: Nath, Oikantik, et al.
Publicado: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2026)
por: Khan, Mohammed Safi Ur Rahman, et al.
Publicado: (2026)
Finding Blind Spots in Evaluator LLMs with Interpretable Checklists
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
Social Media and Adolescents
por: Jaya Shrivastava, Dr. Urmila Pratiti
Publicado: (2026)
por: Jaya Shrivastava, Dr. Urmila Pratiti
Publicado: (2026)
Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages
por: Anand, Srija, et al.
Publicado: (2026)
por: Anand, Srija, et al.
Publicado: (2026)
A Glimpse into Tripura's Cultural Heritage through Language and Oral Tradition: An Analytical Study
por: Sarkar, Ratan, et al.
Publicado: (2025)
por: Sarkar, Ratan, et al.
Publicado: (2025)
The State Of TTS: A Case Study with Human Fooling Rates
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2025)
por: Varadhan, Praveen Srinivasa, et al.
Publicado: (2025)
Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India
por: Bhogale, Kaushal, et al.
Publicado: (2026)
por: Bhogale, Kaushal, et al.
Publicado: (2026)
Insight into material behavior via surface free energy calculations for common energetic materials
por: Janki Brahmbhatt, et al.
Publicado: (2024)
por: Janki Brahmbhatt, et al.
Publicado: (2024)
Hamiltonian Graphs and the Traveling Salesman Problem
por: Mehendale, Dhananjay P.
Publicado: (2007)
por: Mehendale, Dhananjay P.
Publicado: (2007)
On Gracefully Labeling Trees
por: Mehendale, Dhananjay P.
Publicado: (2005)
por: Mehendale, Dhananjay P.
Publicado: (2005)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
por: Singh, Anushka, et al.
Publicado: (2024)
por: Singh, Anushka, et al.
Publicado: (2024)
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models
por: Mundra, Nandini, et al.
Publicado: (2024)
por: Mundra, Nandini, et al.
Publicado: (2024)
RASMALAI: Resources for Adaptive Speech Modeling in Indian Languages with Accents and Intonations
por: Sankar, Ashwin, et al.
Publicado: (2025)
por: Sankar, Ashwin, et al.
Publicado: (2025)
From Shadows to Spotlight: Courageous Followers Shape Leaders Perceptions and Influence Followers’ Job Performance
por: Wajeeha Brar Ghias, et al.
Publicado: (2026)
por: Wajeeha Brar Ghias, et al.
Publicado: (2026)
From Literature to ReWA: Discussing Reproductive Well-being in HCI
por: Chowdhury, Hafsah Mahzabin, et al.
Publicado: (2025)
por: Chowdhury, Hafsah Mahzabin, et al.
Publicado: (2025)
Generative flow induced neural architecture search: Towards discovering optimal architecture in wavelet neural operator
por: Soin, Hartej, et al.
Publicado: (2024)
por: Soin, Hartej, et al.
Publicado: (2024)
Sr, Nd and Hf isotopic compositions of the clay size fraction and the silicate residues of Fe-Mn crusts from the Afanasi-Nikitin Seamount of Central Indian Basin
por: Neethu, Sukumaran
Publicado: (2025)
por: Neethu, Sukumaran
Publicado: (2025)
Genetic study on breeding stock of the Indian mackerel along the Indian coast
por: Sukumaran, Sandhya
Publicado: (2015)
por: Sukumaran, Sandhya
Publicado: (2015)
DexAssist: A Voice-Enabled Dual-LLM Framework for Accessible Web Navigation
por: Mehendale, Shridhar, et al.
Publicado: (2024)
por: Mehendale, Shridhar, et al.
Publicado: (2024)
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
por: Doddapaneni, Sumanth, et al.
Publicado: (2024)
Fitting sparse high-dimensional varying-coefficient models with Bayesian regression tree ensembles
por: Ghosh, Soham, et al.
Publicado: (2025)
por: Ghosh, Soham, et al.
Publicado: (2025)
Ejemplares similares
-
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
por: Javed, Tahir, et al.
Publicado: (2024) -
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
por: Sankar, Ashwin, et al.
Publicado: (2024) -
Recognizing Every Voice: Towards Inclusive ASR for Rural Bhojpuri Women
por: Joshi, Sakshi, et al.
Publicado: (2025) -
Empowering Low-Resource Language ASR via Large-Scale Pseudo Labeling
por: Bhogale, Kaushal Santosh, et al.
Publicado: (2024) -
NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
por: Javed, Tahir, et al.
Publicado: (2025)