Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

In Data or Invisible: Toward a Better Digital Representation of Low-Resource Languages with Knowledge Graphs

Domain:

natural language processing

Record type:

paper
Creator:
Mbe
Host:avatar
Emerging digital technologies are exacerbating the existing divide in Open Access Data (OAD) between high-and low-resource languages, excluding many communities from participating in the global digital transformation. In this PhD proposal, we aim to address this gap, focusing on the language coverage of Linked Open Data knowledge graphs (LOD KGs). First, we identify key variables that characterize language distribution in LOD, including the number of Wikipedia articles per language edition and the number of language-tagged entities in LOD KGs. These variables are analyzed across three major multilingual LOD KGs, DBpedia, BabelNet, and Wikidata, providing insights into the representation and distribution of languages within LOD. Building on this analysis, we intend to study the impact of cross-lingual transfer candidate selection on the task of multilingual KG completion. In particular, we plan to investigate strategies based on linguistic proximity and the availability of curated annotated alignments between languages. Language proximity also motivates us to explore the benefits of analogical reasoning that relies on (dis)similarities and has not yet been investigated to identify correspondences across languages to improve KG completion performance and enhance language coverage in LOD.

Visit

arxiv.org

Tags

Artificial Intelligence

Similar

Multilingual Knowledge Graphs and Low-Resource Languages: A ReviewAdapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via AdaptersSuvarthiSarkar/Creation-of-Knowledge-Graphs-for-Low-resource-LanguageToward robust representation for low-resource automatic speech recognitionToward Robust Multilingual Adaptation of LLMs for Low-Resource LanguagesAfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages

Multilingual Knowledge Graphs and Low-Resource Languages: A Review

There is a lack of multilingual data to support applications in a large number of languages, especia

Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters

This paper explores the integration of graph knowledge from linguistic ontologies into multilingual

SuvarthiSarkar/Creation-of-Knowledge-Graphs-for-Low-resource-Language

In this paper, we have worked on two different research problems: performing cross-lingual Knowledge

Toward robust representation for low-resource automatic speech recognition

Vers une représentation robuste pour la reconnaissance automatique de la parole des langues peu doté

Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages

Large language models (LLMs) continue to struggle with low-resource languages, primarily due to limi

AfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages

Language model compression through knowledge distillation has emerged as a promising approach for de