Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

How to Improve LLMs' Performance on Specific Languages: A Perspective on LLM-Derived Language Similarity

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
AssShiXuaZen
Éditeur:
Und
Hôte:avatar
Large language models (LLMs) exhibit uneven performance across languages. In language-specific applications, practitioners often rely on target-language corpora or cross-lingual transfer to achieve better performance. However, traditional linguistic typology, commonly used as a transfer language selection strategy in previous studies, may not align with LLM's perception of language similarity. This work proposes LLM-based language similarity as a novel perspective for selecting effective fine-tuning languages. We construct a framework to quantify the similarity within each language pair through both the lenses of language-specific performance patterns and cross-lingual transferability, ultimately deriving three similarity score matrices. Moreover, we observe a counter-intuitive phenomenon: super-additive transfer effect, where fine-tuning on a certain language yields higher performance than fine-tuning directly on the target language. Additionally, due to the absence of an existing dataset meeting our experimental requirements, we construct and release M4CQ-Pro dataset, which features domain-diverse distribution of 135 tasks and content consistency across 31 languages (including over 20 medium- and low-resource languages), with 61518 manually reviewed high-quality questions per language. We evaluate our approach on representative multilingual LLMs and results show that all three LLM-based similarity measures effectively guide fine-tuning language selection, outperforming traditional linguistic similarity, with the integrated measure achieving the best results. Our approach provides not only a novel perspective on language similarity, but also practical baselines for selecting fine-tuning languages.

Visit

doi.org

Tasks

language modelingtransfer learning

Similaires

How Does Alignment Enhance LLMs' Multilingual Capabilities? A Language Neurons PerspectiveImpact of Language Similarity Metrics on Cross-Lingual NER Performance in Low-Resource LanguagesWhere Are We? Evaluating LLM Performance on African LanguagesCulturally-Grounded Chain-of-Thought (CG-CoT):Enhancing LLM Performance on Culturally-Specific Tasks in Low-Resource LanguagesLLM Probe: Evaluating LLMs for Low-Resource LanguagesPredicting Machine Translation Performance on Low-Resource Languages: The Role of Domain Similarity

How Does Alignment Enhance LLMs' Multilingual Capabilities? A Language Neurons Perspective

Multilingual Alignment is an effective and representative paradigm to enhance LLMs' multilingual cap

Impact of Language Similarity Metrics on Cross-Lingual NER Performance in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Where Are We? Evaluating LLM Performance on African Languages

Africa’s rich linguistic heritage remains underrepresented in NLP, largely due to historical policies that favor foreign languages and create significant data inequities. In this paper, we integrate theoretical insights on Africa’s language landscape with an empiri

Culturally-Grounded Chain-of-Thought (CG-CoT):Enhancing LLM Performance on Culturally-Specific Tasks in Low-Resource Languages

Large Language Models (LLMs) struggle with culturally-specific reasoning tasks, particularly in low-

LLM Probe: Evaluating LLMs for Low-Resource Languages

Despite rapid advances in large language models (LLMs), their linguistic abilities in low-resource a

Predicting Machine Translation Performance on Low-Resource Languages: The Role of Domain Similarity

Fine-tuning and testing a multilingual large language model is expensive and challenging for low-res