Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Why Low-Resource NLP Needs More Than Cross-Lingual Transfer: Lessons Learned from Luxembourgish

Domain:

natural language processing

Record type:

paper
Creator:
PhiGuoKleBis
Host:avatar
Cross-lingual transfer has become a central paradigm for extending natural language processing (NLP) technologies to low-resource languages. By leveraging supervision from high-resource languages, multilingual language models can achieve strong task performance with little or no labeled target-language data. However, it remains unclear to what extent cross-lingual transfer can substitute for language-specific efforts. In this paper, we synthesize prior research findings and data collection results on Luxembourgish, which, despite its typological proximity to high-resource languages and its presence in a multilingual context, remains insufficiently represented in modern NLP technologies. Across findings, we observe a fundamental interdependence between cross-lingual transfer and language-specific efforts. Cross-lingual transfer can substantially improve target-language performance, but its success depends critically on the availability of sufficiently high-quality, task-aligned target-language data. At the same time, such resources, particularly in low-resource settings, are typically too limited in scale to drive strong performance on their own. Instead, such resources reach their full potential only when leveraged within a cross-lingual framework. We therefore argue that cross-lingual transfer and language-specific efforts should not be viewed as competing alternatives. Instead, they function as complementary components of a sustainable low-resource NLP pipeline. Based on these insights, we provide practical guidelines for integrating and balancing cross-lingual transfer with language-specific development in sustainable low-resource NLP pipelines. Accepted at BigPicture Workshop 2026 (co-located with ACL 2026)

Visit

arxiv.org

Tasks

transfer learning

Tags

Computation and LanguageArtificial Intelligence

Similar

Multilingual Intermediate-Task Training for Low-Resource Cross-Lingual TransferCross-lingual transfer of multilingual models on low resource African LanguagesIntermediate-Task Training for Low-Resource Zero-Shot Cross-Lingual TransferCross-lingual Transfer Accuracy and Task Similarity in Low-Resource LanguagesCross-lingual Transfer Effects on Euphemism Detection in Low-Resource LanguagesCross-lingual Transfer Accuracy in Low-Resource Languages via Task Similarity

Multilingual Intermediate-Task Training for Low-Resource Cross-Lingual Transfer

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Cross-lingual transfer of multilingual models on low resource African Languages

Large multilingual models have significantly advanced natural language processing (NLP) research. Ho

Intermediate-Task Training for Low-Resource Zero-Shot Cross-Lingual Transfer

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Cross-lingual Transfer Accuracy and Task Similarity in Low-Resource Languages

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Cross-lingual Transfer Effects on Euphemism Detection in Low-Resource Languages

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Cross-lingual Transfer Accuracy in Low-Resource Languages via Task Similarity

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni