Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Leveraging LLM For Synchronizing Information Across Multilingual Tables

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
KhiKatAnaRot
Hôte:avatar
The vast amount of online information today poses challenges for non-English speakers, as much of it is concentrated in high-resource languages such as English and French. Wikipedia reflects this imbalance, with content in low-resource languages frequently outdated or incomplete. Recent research has sought to improve cross-language synchronization of Wikipedia tables using rule-based methods. These approaches can be effective, but they struggle with complexity and generalization. This paper explores large language models (LLMs) for multilingual information synchronization, using zero-shot prompting as a scalable solution. We introduce the Information Updation dataset, simulating the real-world process of updating outdated Wikipedia tables, and evaluate LLM performance. Our findings reveal that single-prompt approaches often produce suboptimal results, prompting us to introduce a task decomposition strategy that enhances coherence and accuracy. Our proposed method outperforms existing baselines, particularly in Information Updation (1.79%) and Information Addition (20.58%), highlighting the model strength in dynamically updating and enriching data across architectures. 17 Pages, 11 Tables, 2 Figures

Visit

arxiv.org

Tags

Computation and Language

Similaires

Leveraging LLMs for Synthesizing Training Data Across Many Languages in Multilingual Dense RetrievalMUSSEL: Enhanced Bayesian polygenic risk prediction leveraging information across multiple ancestry groupsA framework for information extraction from tables in biomedical literatureImproving Methodologies for LLM Evaluations Across Global LanguagesMaking a MIRACL: Multilingual Information Retrieval Across a Continuum of Languagesdefimanu3l/african-audio-multilingual-llm

Leveraging LLMs for Synthesizing Training Data Across Many Languages in Multilingual Dense Retrieval

There has been limited success for dense retrieval models in multilingual retrieval, due to uneven a

MUSSEL: Enhanced Bayesian polygenic risk prediction leveraging information across multiple ancestry groups

Polygenic risk scores (PRSs) are now showing promising predictive performance on a wide variety of c

A framework for information extraction from tables in biomedical literature

The scientific literature is growing exponentially, and professionals are no more able to cope with

Improving Methodologies for LLM Evaluations Across Global Languages

As frontier AI models are deployed globally, it is essential that their behaviour remains safe and r

Making a MIRACL: Multilingual Information Retrieval Across a Continuum of Languages

MIRACL (Multilingual Information Retrieval Across a Continuum of Languages) is a multilingual datase

defimanu3l/african-audio-multilingual-llm

A final year Computer Science project implementing an African multilingual Audio-based Multi-lingual