Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

AfriSenti: A Twitter Sentiment Analysis Benchmark for African Languages

Domaine:

natural language processing

Type de record:

paper

Africa is home to over 2000 languages from over six language families and has the highest linguistic diversity among all continents. This includes 75 languages with at least one million speakers each. Yet, there is little NLP research conducted on African languages. Crucial in enabling such research is the availability of high-quality annotated datasets. In this paper, we introduce AfriSenti, which consists of 14 sentiment datasets of 110,000+ tweets in 14 African languages (Amharic, Algerian Arabic, Hausa, Igbo, Kinyarwanda, Moroccan Arabic, Mozambican Portuguese, Nigerian Pidgin, Oromo, Swahili, Tigrinya, Twi, Xitsonga, and Yorùbá) from four language families annotated by native speakers. The data is used in SemEval 2023 Task 12, the first Afro-centric SemEval shared task. We describe the data collection methodology, annotation process, and related challenges when curating each of the datasets. We conduct experiments with different sentiment classification baselines and discuss their usefulness. We hope AfriSenti enables new work on under-represented languages.

Visit

arxiv.org

Connected records

dataset

Tasks

sentiment analysistext classification

Languages

AkanAmharicArabic, Algerian SpokenArabic, Moroccan SpokenBwamu, CwiDinka, SoutheasternHausaIgboKinyarwandaOromo+5

Tags

afrisenti

Licenses

https://github.com/hausanlp/NaijaSenti#license

Similaires

AFRISENTI-SEMEVAL SHARED TASK 12: SENTIMENT ANALYSIS FOR 15 LOW-RESOURCE AFRICAN LANGUAGES USING TWITTER DATASETDuluthNLP at SemEval-2023 Task 12: AfriSenti-SemEval: Sentiment Analysis for Low-resource African Languages using Twitter DatasetTeam ISCL_WINTER at SemEval-2023 Task 12:AfriSenti-SemEval: Sentiment Analysis for Low-resource African Languages using Twitter DatasetSemEval-2023 Task 12: Sentiment Analysis for African Languages (AfriSenti-SemEval)Sentiment Analysis Across Multiple African Languages: A Current BenchmarkCodeHermez/African-Langs-For-Sentiment-Analysis-Using-AfriSenti-Datasets-Sentiment-Analysis-In-African-Langs-

AFRISENTI-SEMEVAL SHARED TASK 12: SENTIMENT ANALYSIS FOR 15 LOW-RESOURCE AFRICAN LANGUAGES USING TWITTER DATASET

AFRISENTI-SEMEVAL SHARED TASK 12: SENTIMENT ANALYSIS FOR 15 LOW-RESOURCE AFRICAN LANGUAGES USING TWITTER DATASET

Poster presented at the Deep Learning Indaba 2022 by Shamsuddeen Muhammad

DuluthNLP at SemEval-2023 Task 12: AfriSenti-SemEval: Sentiment Analysis for Low-resource African Languages using Twitter Dataset

Team ISCL_WINTER at SemEval-2023 Task 12:AfriSenti-SemEval: Sentiment Analysis for Low-resource African Languages using Twitter Dataset

SemEval-2023 Task 12: Sentiment Analysis for African Languages (AfriSenti-SemEval)

We present the first Africentric SemEval Shared task, Sentiment Analysis for African Languages (Afri

Sentiment Analysis Across Multiple African Languages: A Current Benchmark

Sentiment analysis is a fundamental and valuable task in NLP. However, due to limitations in data an

CodeHermez/African-Langs-For-Sentiment-Analysis-Using-AfriSenti-Datasets-Sentiment-Analysis-In-African-Langs-

A Comparative Study of Monolingual and Multilingual Transfer Learning Strategies with Code-Mixing An