Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Evaluation of geographical distortions in language models

Domain:

natural language processing

Record type:

paper
Creator:
DecIntRocTei
Editor:
TerObsDépCen
Publisher:
CCSDSpringer-Verlag
Host:avatar
International audience Geographic bias in language models (LMs) is an underexplored dimension of model fairness, despite growing attention being given to other social biases. We investigate whether LMs provide equally accurate representations across all global regions and propose a benchmark of four indicators to detect undertrained and underperforming areas: (i) indirect assessment of geographic training data coverage via tokenizer analysis, (ii) evaluation of basic geographic knowledge, (iii) detection of geographic distortions, and (iv) visualization of performance disparities through maps. Applying this framework to ten widely used encoder-and decoder-based models, we find systematic overrepresentation of Western countries and consistent underrepresentation of several African, Eastern European, and Middle Eastern regions, leading to measurable performance gaps. We further analyse the impact of these biases on downstream tasks, particularly in crisis response, and show that regions most vulnerable to natural disasters are often those with poorer LM coverage. Our findings underscore the need for geographically balanced LMs to ensure equitable and effective global applications.

Visit

hal.science

Tasks

language modeling

Tags

Spatial informationLLMNLPBias[INFO.INFO-CL]Computer Science [cs]/Computation and Language [cs.CL]

Licenses

https://about.hal.science/hal-authorisation-v1/info:eu-repo/semantics/OpenAccess

Similar

Controlled Evaluation of Syntactic Knowledge in Multilingual Language ModelsControlled Evaluation of Syntactic Knowledge in Multilingual Language ModelsHealMed: Multilingual Evaluation of Large Language Models in MedicineEvaluation Mirage: A Layered Evaluation of Large Language Models and Language Identification for African NLPSinhala Encoder-only Language Models and EvaluationEvaluation of Arabic Large Language Models on Moroccan Dialect

Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models

Language models (LMs) are capable of acquiring elements of human-like syntactic knowledge. Targeted

Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models

This benchmark is made up of "targeted syntactic evaluation tests for three low-resource languages (

HealMed: Multilingual Evaluation of Large Language Models in Medicine

We present HealMed, an expert-reviewed benchmark for multilingual evaluation of large language model

Evaluation Mirage: A Layered Evaluation of Large Language Models and Language Identification for African NLP

David Ifeoluwa Adelani (Supervisor) As Large Language Models (LLMs) are increasingly deployed in glo

Sinhala Encoder-only Language Models and Evaluation

Recently, language models (LMs) have produced excellent results in many natural language processing

Evaluation of Arabic Large Language Models on Moroccan Dialect

Large Language Models (LLMs) have shown outstanding performance in many Natural Language Processing