首页|NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data

NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data

来源：

英文摘要

We introduce Nirantar, a comprehensive framework for evaluating continual learning (CL) in multilingual and multi-domain ASR. Designed to reflect real-world CL challenges, Nirantar leverages data collected incrementally across 22 languages and 208 districts in India through natural episodes. This enables evaluation across Language-Incremental (LIL), Domain-Incremental (DIL), and the novel Language-Incremental Domain-Incremental Learning (LIDIL) scenarios. Unlike prior work that relies on simulated episodes, Nirantar presents dynamic, non-uniform language and domain shifts, making it an ideal testbed for CL research. With 3250 hours of human-transcribed speech, including 1720 hours newly introduced in this work, our framework enables systematic benchmarking of CL methods. We evaluate existing approaches and demonstrate that no single method performs consistently well, underscoring the need for more robust CL strategies.

作者：Tahir Javed、Kaushal Bhogale、Mitesh M. Khapra

作者单位：

学科分类：语言学南亚语系（澳斯特罗-亚细亚语系）

推荐引用：Tahir Javed,Kaushal Bhogale,Mitesh M. Khapra.NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data[EB/OL].(2025-07-01)[2025-07-16].https://arxiv.org/abs/2507.00534.点此复制

NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data

NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data

评论