Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Adapting Monolingual Models: Data can be Scarce when Language Similarity is High

Domaine:

natural language processing

Type de record:

paper
Créateur:
de BarNisWie
Hôte:avatar
For many (minority) languages, the resources needed to train large models are not available. We investigate the performance of zero-shot transfer learning with as little data as possible, and the influence of language similarity in this process. We retrain the lexical layers of four BERT-based models using data from two low-resource target language varieties, while the Transformer layers are independently fine-tuned on a POS-tagging task in the model's source language. By combining the new lexical layers and fine-tuned Transformer layers, we achieve high task performance for both target languages. With high language similarity, 10MB of data appears sufficient to achieve substantial monolingual transfer performance. Monolingual BERT-based models generally achieve higher downstream task performance after retraining the lexical layer than multilingual BERT, even when the target language is included in the multilingual model. Findings of ACL 2021 Camera Ready

Visit

arxiv.org

Tasks

part of speech taggingtransfer learning

Tags

Computation and Language

Similaires

Adapting Chat Language Models Using Only Target Unlabeled Language DataWhen Language Transfer is NegativeWhat Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language ModelsGoldfish: Monolingual Language Models for 350 LanguagesRegional Bias in Monolingual English Language ModelsLugha-Llama: Adapting Large Language Models for African Languages

Adapting Chat Language Models Using Only Target Unlabeled Language Data

Vocabulary expansion (VE) is the de-facto approach to language adaptation of large language models (

When Language Transfer is Negative

This paper analyses morpho-syntactic interference errors committed by learners of French as a foreig

What Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language Models

Compression techniques have been crucial in advancing machine learning by enabling efficient trainin

Goldfish: Monolingual Language Models for 350 Languages

For many low-resource languages, the only available language models are large multilingual models tr

Regional Bias in Monolingual English Language Models

Abstract In Natural Language Processing (NLP), pre-trained language models (LLMs) are wide

Lugha-Llama: Adapting Large Language Models for African Languages

Large language models (LLMs) have achieved impressive results in a wide range of natural language ap