Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Fairness in Multilingual Large Language Models: Addressing the Language Disparity Gap in AI Systems

Domain:

natural language processing

Record type:

paper
Creator:
UpaDurLal
Publisher:
Edt
Host:
Current Large Language Models (LLMs) exhibit significant performance disparities across languages, with English and high-resource languages receiving disproportionate model capacity and training data while speakers of African, Southeast Asian, and Indigenous languages face substantially degraded service quality. This research addresses the critical challenge of fairness in multilingual LLMs by surveying recent developments (2023–2026), analyzing underserved language groups, and proposing methodological approaches to close the language fairness gap. We identify three primary dimensions of unfairness: data scarcity in low-resource languages, suboptimal model architectures for multilingual transfer, and inadequate fairness evaluation metrics. Through analysis of existing benchmarks (XGLUE, Masakhane, FLORES-200, NLLB), we demonstrate that performance parity across language families requires integrated approaches combining data augmentation, architectural innovations, and culturally-informed fairness metrics. Our work introduces the Cross-Lingual Fairness Index (CLFI), a novel metric extending the PEER (Probability of Equal Expected Rank) framework to LLM generation tasks, enabling quantitative assessment of language equity. Case studies from initiatives including Masakhane, IndicNLP, and Google's No Language Left Behind (NLLB) demonstrate feasibility of targeted interventions. We conclude that achieving fairness in multilingual LLMs requires sustained investment in low-resource languages, participatory involvement of native speakers, and adoption of language-aware evaluation protocols throughout the model development lifecycle. Index Terms—Algorithmic Bias, AI Localization, Cross-Lingual Transfer, Fairness Metrics, Language Equity, Language Fairness, Low-Resource Languages, Multilingual LLMs

Visit

doi.org

Similar

Faux Polyglot: A Study on Information Disparity in Multilingual Large Language ModelsQuantifying Language Disparities in Multilingual Large Language ModelsBridging language gaps in multilingual large language modelsIsolating Culture Neurons in Multilingual Large Language ModelsMultilingual Emotion Neurons in Large Audio-Language ModelsAssessing Dialect Fairness and Robustness of Large Language Models in Reasoning Tasks

Faux Polyglot: A Study on Information Disparity in Multilingual Large Language Models

Although the multilingual capability of LLMs offers new opportunities to overcome the language barri

Quantifying Language Disparities in Multilingual Large Language Models

Results reported in large-scale multilingual evaluations are often fragmented and confounded by fact

Bridging language gaps in multilingual large language models

Large language models (LLMs) have revolutionized natural language processing, yet significant perfor

Isolating Culture Neurons in Multilingual Large Language Models

Language and culture are deeply intertwined, yet it has been unclear how and where multilingual larg

Multilingual Emotion Neurons in Large Audio-Language Models

Emotion is central to human communication, and its expression varies across languages. Large audio-l

Assessing Dialect Fairness and Robustness of Large Language Models in Reasoning Tasks

Language is not monolithic. While benchmarks, including those designed for multiple languages, are o