2025 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU)(2025)
Indian Institute of Science (IISc)
被引用0|浏览3
摘要
We present MADASR 2.0, a challenge at ASRU 2025 aimed at advancing multilingual and multidialectal automatic speech recognition (ASR) in low-resource Indian languages. Building on the 2023 edition, it introduces a subset of the RESPIN corpus, over 1200 hours of read speech across 8 languages and 33 dialects, with test sets including both read and spontaneous speech. The challenge comprises four tracks varying by training data size and external resource usage, and supports auxiliary tasks like language and dialect identification. We detail the dataset, tasks, baselines, and submissions and analyse trends across tracks and speech styles. Results highlight the continued difficulty of spontaneous ASR, the benefits of multitask and transfer learning, and effective strategies for building dialect-aware ASR systems. MADASR 2.0 offers a standardised benchmark to support future research on inclusive and scalable ASR for linguistically diverse populations.
更多
查看译文
关键词
Multilingual ASR,Dialectal ASR,Indian Languages,Low-Resource Speech Recognition,Benchmark Challenge,RESPIN Corpus