Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Advancing sentiment analysis for low-resourced african languages using pre-trained language models

Domain:

natural language processing

Record type:

paper
Creator:
KoeMphTur
Publisher:
Pub
Host:
While sentiment analysis systems excel in high-resource languages, most African languages facing limited resources, remain under-represented. This gap leaves a significant portion of the world’s population without access to technologies in their native languages. However, multilingual pre-trained language models (PLM) offer a promising approach for sentiment analysis in low-resource languages. Although the absence of large data in African languages poses a challenge for developing PLMs, fine-tuning and task adaptation of existing multilingual PLMs is an alternative solution. This paper explores the use of multilingual PLMs for sentiment analysis in five Southern African languages: Sepedi , Sesotho , Setswana , isiXhosa , and isiZulu . We leverage existing PLMs and fine-tune them for this specific task, avoiding training the models from scratch. Our work expands on the SAfriSenti corpus, a Twitter sentiment dataset for these languages. We employ various annotation techniques to create a labelled dataset and perform benchmark experiments utilising various multilingual PLMs. Our findings demonstrate the effectiveness of multilingual PLM, particularly for closely-related languages (Sotho-Tswana), where the ensemble PLMs method achieved an average weighted F1 score above 63%. In particular, Nguni closely-related languages achieved an even higher average weighted F1 score, exceeding 77%, highlighting the potential of PLMs for sentiment analysis in South African languages.

Visit

doi.org

Tasks

sentiment analysistext classification

Languages

BirwaNgwoSetswanaSotho, NorthernSotho, SouthernXhosaZulu

Licenses

http://creativecommons.org/licenses/by/4.0/

Similar

Fig 5 - Advancing sentiment analysis for low-resourced african languages using pre-trained language modelsFig 4 - Advancing sentiment analysis for low-resourced african languages using pre-trained language modelsExplainable Pre-Trained Language Models for Sentiment Analysis in Low-Resourced LanguagesPre-Trained Transformer-Based Models for Text Classification Using Low-Resourced Ewe Language

Fig 5 - Advancing sentiment analysis for low-resourced african languages using pre-trained language models

Mean F1 scores of PLMs fine-tuned on closely related African languages, with 95% confidence inter

Fig 4 - Advancing sentiment analysis for low-resourced african languages using pre-trained language models

Mean F1 scores for all models with 95% confidence intervals across five African languages. The er

Explainable Pre-Trained Language Models for Sentiment Analysis in Low-Resourced Languages

Sentiment analysis is a pivotal tool for gauging the public’s perception and understanding

Pre-Trained Transformer-Based Models for Text Classification Using Low-Resourced Ewe Language

Despite a few attempts to automatically crawl Ewe text from online news portals and magazines, the A