Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context

Domain:

natural language processing

Record type:

paper
Creator:
KooKim
Host:avatar
Large Language Models (LLMs) have achieved impressive progress across a wide range of tasks, yet their heavy reliance on English-centric training data leads to significant performance degradation in non-English languages. While existing multilingual prompting methods emphasize reformulating queries into English or enhancing reasoning capabilities, they often fail to incorporate the language- and culture-specific grounding that is essential for some queries. To address this limitation, we propose EMCEE (Extracting synthetic Multilingual Context and merging), a simple yet effective framework that enhances the multilingual capabilities of LLMs by explicitly extracting and utilizing query-relevant knowledge from the LLM itself. In particular, EMCEE first extracts synthetic context to uncover latent, language-specific knowledge encoded within the LLM, and then dynamically merges this contextual insight with reasoning-oriented outputs through a judgment-based selection mechanism. Extensive experiments on four multilingual benchmarks covering diverse languages and tasks demonstrate that EMCEE consistently outperforms prior approaches, achieving an average relative improvement of 16.4% overall and 31.7% in low-resource languages. ACL 2026 Main

Visit

arxiv.org

Tags

Computation and LanguageArtificial Intelligence

Similar

Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation EngineeringAdapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via AdaptersLLMs Beyond English: Scaling the Multilingual Capability of LLMs with Cross-Lingual FeedbackDo Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues TestImproving Multilingual Math Reasoning for African LanguagesBerkGuny/Sequential-fine-tuning-for-improving-everyday-cultural-knowledge-in-multilingual-LLMs

Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering

Large Language Models (LLMs) and Large Vision-Language Models (LVLMs) demonstrate strong reasoning c

Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters

This paper explores the integration of graph knowledge from linguistic ontologies into multilingual

LLMs Beyond English: Scaling the Multilingual Capability of LLMs with Cross-Lingual Feedback

To democratize large language models (LLMs) to most natural languages, it is imperative to make thes

Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test

This paper explores the moral judgment and moral reasoning abilities exhibited by Large Language Mod

Improving Multilingual Math Reasoning for African Languages

Researchers working on low-resource languages face persistent challenges due to limited data availab

BerkGuny/Sequential-fine-tuning-for-improving-everyday-cultural-knowledge-in-multilingual-LLMs

Developed a two-stage multilingual LLM fine-tuning pipeline that improves culturally grounded questi