Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

LSR: Linguistic Safety Robustness Benchmark for Low-Resource West African Languages

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
Far
Hôte:avatar
Safety alignment in large language models relies predominantly on English-language training data. When harmful intent is expressed in low-resource languages, refusal mechanisms that hold in English frequently fail to activate. We introduce LSR (Linguistic Safety Robustness), the first systematic benchmark for measuring cross-lingual refusal degradation in West African languages: Yoruba, Hausa, Igbo, and Igala. LSR uses a dual-probe evaluation protocol - submitting matched English and target-language probes to the same model - and introduces Refusal Centroid Drift (RCD), a metric that quantifies how much of a model's English refusal behavior is lost when harmful intent is encoded in a target language. We evaluate Gemini 2.5 Flash across 14 culturally grounded attack probes in four harm categories. English refusal rates hold at approximately 90 percent. Across West African languages, refusal rates fall to 35-55 percent, with Igala showing the most severe degradation (RCD = 0.55). LSR is implemented in the Inspect AI evaluation framework and is available as a PR-ready contribution to the UK AISI's inspect_evals repository. A live reference implementation and the benchmark dataset are publicly available. 6 pages. Reference implementation: huggingface.co. Dataset: huggingface.co

Visit

arxiv.org

Languages

HausaIgalaYoruba

Tags

Computation and LanguageArtificial Intelligence

Similaires

English Intermediate-Task Training for Robustness in Low-Resource XTREME Benchmark LanguagesLow Resource Neural Machine Translation: A Benchmark for Five African LanguagesShadman19/llm-benchmark-low-resource-languagesMultilingual Intermediate-Task Training for Robustness in Low-Resource LanguagesLINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct LanguagesCross-lingual NER Model Robustness in Low-Resource Languages

English Intermediate-Task Training for Robustness in Low-Resource XTREME Benchmark Languages

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Low Resource Neural Machine Translation: A Benchmark for Five African Languages

Recent advents in Neural Machine Translation (NMT) have shown improvements in low-resource language (LRL) translation tasks. In this work, we benchmark NMT between English and five African LRL pairs (Swahili, Amharic, Tigrigna, Oromo, Somali [SATOS]). We collected

Shadman19/llm-benchmark-low-resource-languages

Evaluating open-source LLMs on Bengali, Swahili and Tamil vs English baseline # 🌍 LLM Benchmark for

Multilingual Intermediate-Task Training for Robustness in Low-Resource Languages

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages

In this paper, we present the LingOly benchmark, a novel benchmark for advanced reasoning abilities

Cross-lingual NER Model Robustness in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages