NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Javed, Tahir, Bhogale, Kaushal, Khapra, Mitesh M.
Format: Preprint
Publié: 2025
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866915367761215488
author Javed, Tahir
Bhogale, Kaushal
Khapra, Mitesh M.
author_facet Javed, Tahir
Bhogale, Kaushal
Khapra, Mitesh M.
contents We introduce Nirantar, a comprehensive framework for evaluating continual learning (CL) in multilingual and multi-domain ASR. Designed to reflect real-world CL challenges, Nirantar leverages data collected incrementally across 22 languages and 208 districts in India through natural episodes. This enables evaluation across Language-Incremental (LIL), Domain-Incremental (DIL), and the novel Language-Incremental Domain-Incremental Learning (LIDIL) scenarios. Unlike prior work that relies on simulated episodes, Nirantar presents dynamic, non-uniform language and domain shifts, making it an ideal testbed for CL research. With 3250 hours of human-transcribed speech, including 1720 hours newly introduced in this work, our framework enables systematic benchmarking of CL methods. We evaluate existing approaches and demonstrate that no single method performs consistently well, underscoring the need for more robust CL strategies.
format Preprint
id arxiv_https___arxiv_org_abs_2507_00534
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
Javed, Tahir
Bhogale, Kaushal
Khapra, Mitesh M.
Computation and Language
We introduce Nirantar, a comprehensive framework for evaluating continual learning (CL) in multilingual and multi-domain ASR. Designed to reflect real-world CL challenges, Nirantar leverages data collected incrementally across 22 languages and 208 districts in India through natural episodes. This enables evaluation across Language-Incremental (LIL), Domain-Incremental (DIL), and the novel Language-Incremental Domain-Incremental Learning (LIDIL) scenarios. Unlike prior work that relies on simulated episodes, Nirantar presents dynamic, non-uniform language and domain shifts, making it an ideal testbed for CL research. With 3250 hours of human-transcribed speech, including 1720 hours newly introduced in this work, our framework enables systematic benchmarking of CL methods. We evaluate existing approaches and demonstrate that no single method performs consistently well, underscoring the need for more robust CL strategies.
title NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
topic Computation and Language
url https://arxiv.org/abs/2507.00534