Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Unified Gradient Projection: Language-Balanced Continual Learning for Multilingual Low-Resource ASR

Domain:

natural language processing

Record type:

paper
Creator:
RenLinAi,Tan
Publisher:
arXiv
Host:avatar
Large-scale pretrained ASR models such as Whisper exhibit strong multilingual capabilities. However, fine-tuning on low-resource languages often causes catastrophic forgetting. Although continual learning mitigates this issue, existing methods struggle to regulate cross-task interference in multilingual settings, where dominant languages bias optimization. We propose Unified Gradient Projection (UGP), which constrains parameter updates using reference gradients from language-balanced replay in a unified projection space. By equalizing per-language contributions in the projection, UGP reduces dominant-language bias and improves cross-lingual stability. We further show that combining gradient-level projection with data-level replay yields complementary gains in stability and plasticity. Across diverse low-resource language groups and model scales, UGP enables effective adaptation while substantially mitigating forgetting. On Whisper-large-v3, it achieves near-zero average forgetting. Accepted by Interspeech 2026

Visit

doi.org

Tasks

automatic speech recognitionspeech processingtransfer learning

Tags

Computation and Language (cs.CL)Sound (cs.SD)Audio and Speech Processing (eess.AS)FOS: Computer and information sciencesFOS: Electrical engineering, electronic engineering, information engineering

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

Continual-learning for Modelling Low-Resource Languages from Large Language ModelsMultilingual Projection for Parsing Truly Low-Resource LanguagesContinual Mixed-Language Pre-Training for Extremely Low-Resource Neural Machine TranslationCross-lingual NER Performance via Annotation Projection vs. Multilingual Language Models in Low-Resource LanguagesInstructAlign: High-and-Low Resource Language Alignment via Continual CrosslingualInstruction Tuningbuumba641/Low-Resource-Multilingual-ASR-for-Zambian-Languages-Nyanja-Tonga-and-Bemba

Continual-learning for Modelling Low-Resource Languages from Large Language Models

Modelling a language model for a multi-lingual scenario includes several potential challenges, among

Multilingual Projection for Parsing Truly Low-Resource Languages

We propose a novel approach to cross-lingual part-of-speech tagging and dependency parsing for truly

Continual Mixed-Language Pre-Training for Extremely Low-Resource Neural Machine Translation

The data scarcity in low-resource languages has become a bottleneck to building robust neural machin

Cross-lingual NER Performance via Annotation Projection vs. Multilingual Language Models in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

InstructAlign: High-and-Low Resource Language Alignment via Continual CrosslingualInstruction Tuning

Large language models (LLMs) that are tuned with instructions have demonstrated remarkable capabilit

buumba641/Low-Resource-Multilingual-ASR-for-Zambian-Languages-Nyanja-Tonga-and-Bemba

# Low-Resource ASR for Zambian Languages (Bemba, Nyanja, Tonga) **Final Year Research Project (UNZA