Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Adapting Language-Specific LLMs to a Reasoning Model in One Day via Model Merging -- An Open Recipe

Domaine:

natural language processing

Type de record:

papermodel
Créateur:
PipTavManTha
Hôte:avatar
This paper investigates data selection and model merging methodologies aimed at incorporating advanced reasoning capabilities such as those of DeepSeek R1 into language-specific large language models (LLMs), with a particular focus on the Thai LLM. Our goal is to enhance the reasoning capabilities of language-specific LLMs while maintaining their target language abilities. DeepSeek R1 excels in reasoning but primarily benefits high-resource languages such as English and Chinese. However, low-resource languages remain underserved due to the dominance of English-centric training data and model optimizations, which limit performance in these languages. This limitation results in unreliable code-switching and diminished effectiveness on tasks in low-resource languages. Meanwhile, local and regional LLM initiatives have attempted to bridge this gap by developing language-specific LLMs that focus on improving local linguistic fidelity. We demonstrate that, with only publicly available datasets and a computational budget of $120, it is possible to enhance the reasoning capabilities of language-specific LLMs to match the level of DeepSeek R1, without compromising their performance on target language tasks. 9 pages

Visit

arxiv.org

Tasks

language modelingtransfer learning

Tags

Computation and LanguageArtificial Intelligence

Similaires

Family Matters: Language Transfer and Merging for Adapting Small LLMs to FaroeseAdapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via AdaptersLanguage-specific model example.pps121/LLM-Cultural-Model-MergingMed-CoReasoner: Reducing Language Disparities in Medical Reasoning via Language-Informed Co-ReasoningLatxa: An Open Language Model and Evaluation Suite for Basque

Family Matters: Language Transfer and Merging for Adapting Small LLMs to Faroese

We investigate strategies for adapting small, efficient language models to Faroese, a low-resource N

Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters

This paper explores the integration of graph knowledge from linguistic ontologies into multilingual

Language-specific model example.

Model details for language-specific model generated from YouTube data for daily campaigns in Ghan

pps121/LLM-Cultural-Model-Merging

A research project on Model Merging. Two fine-tuned culturally aligned models (e.g., African and Lat

Med-CoReasoner: Reducing Language Disparities in Medical Reasoning via Language-Informed Co-Reasoning

While reasoning-enhanced large language models perform strongly on English medical tasks, a persiste

Latxa: An Open Language Model and Evaluation Suite for Basque

We introduce Latxa, a family of large language models for Basque ranging from 7 to 70 billion parame