Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review

Domaine:

natural language processing

Type de record:

paper
Créateur:
LemUzoAnyKap
Éditeur:
arXiv
Hôte:avatar
Large Language Models (LLMs) have achieved substantial progress in safety alignment, yet their safety guarantees remain significantly weaker in low-resource and multilingual settings than in high-resource languages. In this paper, we conduct a Systematic Literature Review (SLR) of LLM safety alignment in low-resource languages by adopting the PRISMA 2020 methodology. Out of roughly 1,500 papers identified from Semantic Scholar, arXiv, and OpenAlex, 50 relevant studies have been selected and analyzed. Our review is organized around four themes: safety alignment methods, multilingual safety risks, evaluation benchmarks, and cross-lingual transferability. We further propose a taxonomy of safety alignment approaches based on three adaptation mechanisms: data adaptation, objective optimization, and mechanistic alignment. Across literature, translated English benchmarks fail to sufficiently represent culturally rooted harms, and multilingual models are more vulnerable to cross-lingual jailbreaks, code-switching attacks, and safety degradation in underrepresented languages. These failures are driven by several key factors, including uneven multilingual pre-training coverage, insufficient native-language preference data, poor transfer of safety representations, and a lack of culturally aware evaluation frameworks. The review also notes that many low-resource languages, especially African languages, have fewer safety benchmarks available than other multilingual regions. Overall, the results reveal a persistent multilingual safety gap, and suggest that future progress will require culturally grounded benchmarks, participatory data collection, balanced multilingual pre-training, and scalable multilingual alignment methods. The paper was accepted at LM4UC workshop organize by IJCAI. I added a screenshot of the decision (Open Review)

Visit

doi.org

Tasks

transfer learning

Tags

Computation and Language (cs.CL)Artificial Intelligence (cs.AI)Machine Learning (cs.LG)FOS: Computer and information sciences

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature ReviewToxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource LanguagesUnlocking LLM Safeguards for Low-Resource Languages via Reasoning and Alignment with Minimal Training DataA Systematic Literature Review on Bias Evaluation and Mitigation in Automatic Speech Recognition Models for Low-Resource African LanguagesShadman19/llm-benchmark-low-resource-languagesaadilganigaie/LLM-for-Low-Resource-Languages

Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature Review

ASR has achieved remarkable global progress, yet African low-resource languages remain rigorously un

Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages

The advancement of Large Language Models (LLMs) has transformed natural language processing; however

Unlocking LLM Safeguards for Low-Resource Languages via Reasoning and Alignment with Minimal Training Data

Recent advances in LLMs have enhanced AI capabilities, but also increased the risk posed by maliciou

A Systematic Literature Review on Bias Evaluation and Mitigation in Automatic Speech Recognition Models for Low-Resource African Languages

With recent advancements in speech recognition, it is crucial to ensure that automatic speech recogn

Shadman19/llm-benchmark-low-resource-languages

Evaluating open-source LLMs on Bengali, Swahili and Tamil vs English baseline # 🌍 LLM Benchmark for

aadilganigaie/LLM-for-Low-Resource-Languages

Extending the Vocabulary of Large Language Models for Low-Resource Languages In the realm of natural