Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Development of a Multilingual Lexicon Based on Sentiment Analysis for Low-Resource Languages

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Mik
Éditeur:
Spr
Hôte:
Abstract The multilingual landscape of South Africa and the Democratic Republic of Congo (DRC) presents considerable challenges for multilingual translation due to the scarcity of accurately labeled datasets. Existing approaches, based on monolingual datasets and machine translation methods, often fail to address mixed-language contexts and nuances of sentiment polarity. This study aims to address these gaps by developing a multilingual lexicon initially designed for French, now enriched with translations and sentiment scores for English, Afrikaans, Sepedi, and Zulu. A corpus of 3,000 words and 1,000 sentences was created, and machine learning techniques such as random forests, support vector machines (SVM), decision trees, and the Naive Bayes classifier were applied to the lexicon. Furthermore, the study leverages a transformer-based model achieving remarkable performance with 99% precision and 98% accuracy in contextual sentiment prediction. Explainable artificial intelligence (XAI) was integrated to clarify model predictions, thus improving confidence in multilingual translation. The results demonstrate the usefulness of the lexicon in improving low-resource language translation and sentiment analysis, laying the foundation for scalable AI solutions in linguistically diverse contexts.

Visit

doi.org

Tasks

sentiment analysistext classification

Languages

AfrikaansSotho, Northern

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment LexiconA Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AITriLex: A Framework for Multilingual Sentiment Analysis in Low-Resource South African LanguagesHYBRID DEEP LEARNING MODELS FOR MULTILINGUAL SENTIMENT ANALYSIS IN LOW-RESOURCE LANGUAGESgift-xipu/Multilingual-Sentiment-Analysis-Lexicon-for-African-Languages-using-LLMsD-LexeCan: A Dynamic Lexicon-Based Framework for Sentiment Analysis in Tarifit, a Low-Resource Multiscript Language

Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon

Improving multilingual language models capabilities in low-resource languages is generally difficult

A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI

South Africa and the Democratic Republic of Congo (DRC) present a complex linguistic landscape with

TriLex: A Framework for Multilingual Sentiment Analysis in Low-Resource South African Languages

Low-resource African languages remain underrepresented in sentiment analysis, limiting both lexical

HYBRID DEEP LEARNING MODELS FOR MULTILINGUAL SENTIMENT ANALYSIS IN LOW-RESOURCE LANGUAGES

Multilingual sentiment analysis poses significant challenges, especially in the context of languages

gift-xipu/Multilingual-Sentiment-Analysis-Lexicon-for-African-Languages-using-LLMs

# Multilingual Sentiment Analysis Lexicon for African Languages using LLMS ## Introduction This is

D-LexeCan: A Dynamic Lexicon-Based Framework for Sentiment Analysis in Tarifit, a Low-Resource Multiscript Language

Sentiment analysis for low-resource languages remains challenging due to limited annotated data, ort