Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Improving Low-Resource Translation with Dictionary-Guided Fine-Tuning and RL: A Spanish-to-Wayuunaiki Study

Domaine:

natural language processing

Type de record:

paper
Créateur:
MosRobRodMan
Hôte:avatar
Low-resource machine translation remains a significant challenge for large language models (LLMs), which often lack exposure to these languages during pretraining and have limited parallel data for fine-tuning. We propose a novel approach that enhances translation for low-resource languages by integrating an external dictionary tool and training models end-to-end using reinforcement learning, in addition to supervised fine-tuning. Focusing on the Spanish-Wayuunaiki language pair, we frame translation as a tool-augmented decision-making problem in which the model can selectively consult a bilingual dictionary during generation. Our method combines supervised instruction tuning with Guided Reward Policy Optimization (GRPO), enabling the model to learn both when and how to use the tool effectively. BLEU similarity scores are used as rewards to guide this learning process. Preliminary results show that our tool-augmented models achieve up to +3.37 BLEU improvement over previous work, and a 18% relative gain compared to a supervised baseline without dictionary access, on the Spanish-Wayuunaiki test set from the AmericasNLP 2025 Shared Task. We also conduct ablation studies to assess the effects of model architecture and training strategy, comparing Qwen2.5-0.5B-Instruct with other models such as LLaMA and a prior NLLB-based system. These findings highlight the promise of combining LLMs with external tools and the role of reinforcement learning in improving translation quality in low-resource language settings.

Visit

arxiv.org

Tasks

machine translation

Tags

Computation and LanguageArtificial Intelligence

Similaires

Fine-Tuning mBART-50 for Akkadian-to-English Translation: A Transfer Learning Approach for Low-ResourceFine-Tuning LLMs for Low-Resource Dialect Translation: The Case of LebaneseFine Tuning Methods for Low-resource LanguagesSAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast AsiaAdapting Large Language Models to Low-Resource Tibetan: A Two-Stage Continual and Supervised Fine-Tuning StudyFine-Tuning Whisper for Kinyarwanda: A Practical Approach to Low-Resource ASR Development

Fine-Tuning mBART-50 for Akkadian-to-English Translation: A Transfer Learning Approach for Low-Resource

Overview The research paper "Fine-Tuning mBART-50 for Akkadian-to-English Translation" by Frank Mor

Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese

This paper examines the effectiveness of Large Language Models (LLMs) in translating the low-resourc

Fine Tuning Methods for Low-resource Languages

The rise of Large Language Models has not been inclusive of all cultures. The models are mostly trai

SAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast Asia

The vision of an inclusive World Wide Web is impeded by a severe linguistic divide, particularly for

Adapting Large Language Models to Low-Resource Tibetan: A Two-Stage Continual and Supervised Fine-Tuning Study

Adapting large language models (LLMs) to low-resource languages remains a major challenge due to dat

Fine-Tuning Whisper for Kinyarwanda: A Practical Approach to Low-Resource ASR Development