Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Unlocking LLM Safeguards for Low-Resource Languages via Reasoning and Alignment with Minimal Training Data

Domain:

natural language processing

Record type:

papermodeldataset
Creator:
CheZhaLinHou
Host:avatar
Recent advances in LLMs have enhanced AI capabilities, but also increased the risk posed by malicious requests, highlighting the need for effective LLM safeguards to detect such queries. Existing approaches largely rely on classifier-based methods that lack interpretability and perform poorly on low-resource languages. To address these limitations, we propose ConsistentGuard, a novel reasoning-based multilingual safeguard, which enhances explainability via reasoning and boosts knowledge transfer between languages through alignment. With only 1,000 training samples, our method demonstrates superior performance on three datasets across six languages, outperforming larger models trained with significantly more data, and exhibits strong interpretability and generalization ability. We also contribute a multilingual benchmark extension and release our codes to support future research. Accepted to MRL Workshop at EMNLP 2025

Visit

arxiv.org

Tags

Computation and Language

Similar

Mitigating Cross-Lingual Performance Degradation in Low-Resource Languages via Intermediate-Task Training on English ReasoningLLM Safety Alignment in Low-Resource Languages: A Systematic Literature ReviewCross-lingual NER Generalization via Embedding Alignment in Low-Resource Languagesaadilganigaie/LLM-for-Low-Resource-LanguagesLoanword Identification in Low-Resource Languages with Minimal SupervisionEnhancing Cross-Lingual Transfer for Low-Resource Languages via Intermediate-Task Training

Mitigating Cross-Lingual Performance Degradation in Low-Resource Languages via Intermediate-Task Training on English Reasoning

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

LLM Safety Alignment in Low-Resource Languages: A Systematic Literature Review

Large Language Models (LLMs) have achieved substantial progress in safety alignment, yet their safet

Cross-lingual NER Generalization via Embedding Alignment in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

aadilganigaie/LLM-for-Low-Resource-Languages

Extending the Vocabulary of Large Language Models for Low-Resource Languages In the realm of natural

Loanword Identification in Low-Resource Languages with Minimal Supervision

Bilingual resources play a very important role in many natural language processing tasks, especially

Enhancing Cross-Lingual Transfer for Low-Resource Languages via Intermediate-Task Training

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni