Search papers, labs, and topics across Lattice.
This study investigates the impact of linguistic relatedness on cross-lingual transfer in large multilingual automatic speech recognition (ASR) systems, particularly for low-resource African languages. By employing a controlled experimental design across various factors and ASR models, the research reveals that pre-adaptation using related auxiliary languages does not yield significant transfer improvements when target-language data is scarce. These findings challenge the assumption that linguistic relatedness is a reliable predictor of cross-lingual transfer success in large-scale ASR applications.
Pre-adaptation on linguistically related languages fails to enhance performance in large ASR systems, questioning a common assumption in cross-lingual transfer strategies.
Extending automatic speech recognition (ASR) to low-resource African languages is constrained by the prohibitive demands of data collection at scale. A promising direction is to leverage linguistic relatedness to enhance cross-lingual transfer from a related auxiliary language to the low-resource target by sequentially adapting on both. Although this strategy has shown meaningful improvements in small ASR models, its effectiveness in large ASR remains unclear. We extend this framework to large multilingual ASR through a systematic controlled experimental design spanning six factors, two Africa-centric corpora, and four large ASR models, isolating whether linguistic relatedness reliably predicts cross-lingual transfer gains in this setting. Across all conditions, pre-adaptation on related auxiliary languages yields no practically meaningful transfer improvements given minimal target-language data, suggesting that linguistic relatedness alone may not reliably predict cross-lingual transfer gains in large multilingual ASR, or constitute an effective strategy for extending such models to low-resource languages.