Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Adapting Multilingual Speech Representation Model for a New, Underresourced Language through Multilingual Fine-tuning and Continued Pretraining

Domain:

natural language processing

Record type:

papermodel
Creator:
NowPtaMurNie
Host:avatar
In recent years, neural models learned through self-supervised pretraining on large scale multilingual text or speech data have exhibited promising results for underresourced languages, especially when a relatively large amount of data from related language(s) is available. While the technology has a potential for facilitating tasks carried out in language documentation projects, such as speech transcription, pretraining a multilingual model from scratch for every new language would be highly impractical. We investigate the possibility for adapting an existing multilingual wav2vec 2.0 model for a new language, focusing on actual fieldwork data from a critically endangered tongue: Ainu. Specifically, we (i) examine the feasibility of leveraging data from similar languages also in fine-tuning; (ii) verify whether the model's performance can be improved by further pretraining on target language data. Our results show that continued pretraining is the most effective method to adapt a wav2vec 2.0 model for a new language and leads to considerable reduction in error rates. Furthermore, we find that if a model pretrained on a related speech variety or an unrelated language with similar phonological characteristics is available, multilingual fine-tuning using additional data from that language can have positive impact on speech recognition performance when there is very little labeled data in the target language. 14 pages

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Computation and LanguageMachine LearningAudio and Speech Processing

Similar

Multilingual language model Adaptive Fine-Tuning: A Study on African LanguagesAdapting Pre-trained Language Models to African Languages via Multilingual Adaptive Fine-TuningMULTILINGUAL ADAPTIVE FINE-TUNING (MAFT)Comparison of Intermediate-Task Fine-Tuning and Multilingual Fine-Tuning for Zero-Shot Low-Resource Language AccuracyMultilingual Language Model Pretraining using Machine-translated DataRevisiting Multilingual Data Mixtures in Language Model Pretraining

Multilingual language model Adaptive Fine-Tuning: A Study on African Languages

Multilingual pre-trained language models (PLMs) have demonstrated impressive performance on several downstream tasks on both high-resourced and low-resourced languages. However, there is still a large performance drop for languages unseen during pre-training, espec

Adapting Pre-trained Language Models to African Languages via Multilingual Adaptive Fine-Tuning

Multilingual pre-trained language models (PLMs) have demonstrated impressive performance on several downstream tasks for both high-resourced and low-resourced languages. However, there is still a large performance drop for languages unseen during pre-training, espe

MULTILINGUAL ADAPTIVE FINE-TUNING (MAFT)

We introduce MAFT as an approach to adapt a multi-lingual PLM to a new set of languages. Adapting PLMs has been shown to be effective when adapting to a new domain (Gururangan et al., 2020) or language (Pfeiffer et al., 2020; Alabi et al., 2020; Adelani et al., 202

Comparison of Intermediate-Task Fine-Tuning and Multilingual Fine-Tuning for Zero-Shot Low-Resource Language Accuracy

Accuracy of English-language Question Answering (QA) systems has improved significantly in recent ye

Multilingual Language Model Pretraining using Machine-translated Data

High-resource languages such as English, enables the pretraining of high-quality large language mode

Revisiting Multilingual Data Mixtures in Language Model Pretraining

The impact of different multilingual data mixtures in pretraining large language models (LLMs) has b