RASMALAI: Resources for Adaptive Speech Modeling in Indian Languages with Accents and Intonations
Fuente:
arXiv
Saved in:
| Main Authors: | Sankar, Ashwin, Lacombe, Yoach, Thomas, Sherry, Varadhan, Praveen Srinivasa, Gandhi, Sanchit, Khapra, Mitesh M |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rasa: Building Expressive Speech Synthesis Systems for Indian Languages in Low-resource Settings
by: Varadhan, Praveen Srinivasa, et al.
Published: (2024)
by: Varadhan, Praveen Srinivasa, et al.
Published: (2024)
Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025)
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025)
Enhancing Out-of-Vocabulary Performance of Indian TTS Systems for Practical Applications through Low-Effort Data Strategies
by: Anand, Srija, et al.
Published: (2024)
by: Anand, Srija, et al.
Published: (2024)
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
by: Sankar, Ashwin, et al.
Published: (2024)
by: Sankar, Ashwin, et al.
Published: (2024)
The State Of TTS: A Case Study with Human Fooling Rates
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025)
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025)
ELAICHI: Enhancing Low-resource TTS by Addressing Infrequent and Low-frequency Character Bigrams
by: Anand, Srija, et al.
Published: (2024)
by: Anand, Srija, et al.
Published: (2024)
Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation
by: Varadhan, Praveen Srinivasa, et al.
Published: (2024)
by: Varadhan, Praveen Srinivasa, et al.
Published: (2024)
NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
by: Javed, Tahir, et al.
Published: (2025)
by: Javed, Tahir, et al.
Published: (2025)
Towards Building Large Scale Datasets and State-of-the-Art Automatic Speech Translation Systems for 14 Indian Languages
by: Sankar, Ashwin, et al.
Published: (2024)
by: Sankar, Ashwin, et al.
Published: (2024)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
by: Singh, Anushka, et al.
Published: (2024)
by: Singh, Anushka, et al.
Published: (2024)
Towards Orthographically-Informed Evaluation of Speech Recognition Systems for Indian Languages
by: Bhogale, Kaushal Santosh, et al.
Published: (2026)
by: Bhogale, Kaushal Santosh, et al.
Published: (2026)
Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages
by: Anand, Srija, et al.
Published: (2026)
by: Anand, Srija, et al.
Published: (2026)
IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages
by: Javed, Tahir, et al.
Published: (2024)
by: Javed, Tahir, et al.
Published: (2024)
Can Vision-Language Models Evaluate Handwritten Math?
by: Nath, Oikantik, et al.
Published: (2025)
by: Nath, Oikantik, et al.
Published: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
IndicLLMSuite: A Blueprint for Creating Pre-training and Fine-Tuning Datasets for Indian Languages
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2024)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2024)
An Empirical Comparison of Vocabulary Expansion and Initialization Approaches for Language Models
by: Mundra, Nandini, et al.
Published: (2024)
by: Mundra, Nandini, et al.
Published: (2024)
FairI Tales: Evaluation of Fairness in Indian Contexts with a Focus on Bias and Stereotypes
by: Nawale, Janki Atul, et al.
Published: (2025)
by: Nawale, Janki Atul, et al.
Published: (2025)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
by: Pulipaka, Sidharth, et al.
Published: (2025)
by: Pulipaka, Sidharth, et al.
Published: (2025)
Empowering Low-Resource Language ASR via Large-Scale Pseudo Labeling
by: Bhogale, Kaushal Santosh, et al.
Published: (2024)
by: Bhogale, Kaushal Santosh, et al.
Published: (2024)
Finding Blind Spots in Evaluator LLMs with Interpretable Checklists
by: Doddapaneni, Sumanth, et al.
Published: (2024)
by: Doddapaneni, Sumanth, et al.
Published: (2024)
Can Small Language Models Use What They Retrieve? An Empirical Study of Retrieval Utilization Across Model Scale
by: Pandey, Sanchit
Published: (2026)
by: Pandey, Sanchit
Published: (2026)
Mixture of LoRA Experts for Low-Resourced Multi-Accent Automatic Speech Recognition
by: Bagat, Raphaël, et al.
Published: (2025)
by: Bagat, Raphaël, et al.
Published: (2025)
LID Models are Actually Accent Classifiers: Implications and Solutions for LID on Accented Speech
by: Bafna, Niyati, et al.
Published: (2025)
by: Bafna, Niyati, et al.
Published: (2025)
ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
by: He, Xiangheng, et al.
Published: (2024)
by: He, Xiangheng, et al.
Published: (2024)
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
by: Javed, Tahir, et al.
Published: (2024)
by: Javed, Tahir, et al.
Published: (2024)
Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models
by: Ahuja, Sanchit, et al.
Published: (2026)
by: Ahuja, Sanchit, et al.
Published: (2026)
A Linguistically Motivated Analysis of Intonational Phrasing in Text-to-Speech Systems: Revealing Gaps in Syntactic Sensitivity
by: Pouw, Charlotte, et al.
Published: (2025)
by: Pouw, Charlotte, et al.
Published: (2025)
Performant ASR Models for Medical Entities in Accented Speech
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Cross-Lingual Auto Evaluation for Assessing Multilingual LLMs
by: Doddapaneni, Sumanth, et al.
Published: (2024)
by: Doddapaneni, Sumanth, et al.
Published: (2024)
Accent Vector: Controllable Accent Manipulation for Multilingual TTS Without Accented Data
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
TRIDENT: A Redundant Architecture for Caribbean-Accented Emergency Speech Triage
by: Galbraith, Elroy, et al.
Published: (2025)
by: Galbraith, Elroy, et al.
Published: (2025)
Quantifying Speaker Embedding Phonological Rule Interactions in Accented Speech Synthesis
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
by: Lertpetchpun, Thanathai, et al.
Published: (2026)
Pairwise Evaluation of Accent Similarity in Speech Synthesis
by: Zhong, Jinzuomu, et al.
Published: (2025)
by: Zhong, Jinzuomu, et al.
Published: (2025)
PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions
by: Jin, Sicheng, et al.
Published: (2026)
by: Jin, Sicheng, et al.
Published: (2026)
Cued Speech Generation Leveraging a Pre-trained Audiovisual Text-to-Speech Model
by: Sankar, Sanjana, et al.
Published: (2025)
by: Sankar, Sanjana, et al.
Published: (2025)
Improving Accented Speech Recognition using Data Augmentation based on Unsupervised Text-to-Speech Synthesis
by: Do, Cong-Thanh, et al.
Published: (2024)
by: Do, Cong-Thanh, et al.
Published: (2024)
Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
by: Lee, Wonjun, et al.
Published: (2026)
by: Lee, Wonjun, et al.
Published: (2026)
PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech
by: Menta, Venkata Pushpak Teja
Published: (2026)
by: Menta, Venkata Pushpak Teja
Published: (2026)
DOSA: A Dataset of Social Artifacts from Different Indian Geographical Subcultures
by: Seth, Agrima, et al.
Published: (2024)
by: Seth, Agrima, et al.
Published: (2024)
Similar Items
-
Rasa: Building Expressive Speech Synthesis Systems for Indian Languages in Low-resource Settings
by: Varadhan, Praveen Srinivasa, et al.
Published: (2024) -
Phir Hera Fairy: An English Fairytaler is a Strong Faker of Fluent Speech in Low-Resource Indian Languages
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025) -
Enhancing Out-of-Vocabulary Performance of Indian TTS Systems for Practical Applications through Low-Effort Data Strategies
by: Anand, Srija, et al.
Published: (2024) -
IndicVoices-R: Unlocking a Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS
by: Sankar, Ashwin, et al.
Published: (2024) -
The State Of TTS: A Case Study with Human Fooling Rates
by: Varadhan, Praveen Srinivasa, et al.
Published: (2025)