Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

mteb/NaijaSenti

Domain:

natural language processing

Record type:

dataset
Creator:
mteb
Host:
NaijaSenti An MTEB dataset Massive Text Embedding Benchmark NaijaSenti is the first large-scale human-annotated Twitter sentiment dataset for the four most widely spoken languages in Nigeria — Hausa, Igbo, Nigerian-Pidgin, and Yorùbá — consisting of around 30,000 annotated tweets per language, including a significant fraction of code-mixed tweets. Task category t2c Domains Social, Written Reference NaijaSenti: A Nigerian Twit…

Visit

huggingface.co

Tasks

sentiment analysistext classification

Languages

HausaIgboPidgin, NigerianYoruba

Tags

mtebtext

Licenses

cc-by-4.0

Similar

NaijaSentishmuhammadd/NaijaSentimteb/MIRACLRerankingmteb/NollySentiBitextMiningmteb/MasakhaNEWSClusteringP2Pmteb/AfriSentiClassification

NaijaSenti

NaijaSenti is the first large-scale human-annotated Twitter sentiment dataset for the four most wide

shmuhammadd/NaijaSenti

This is a Lacuna Funded Project to develop sentiment and emotion corpus for three Nigerian languages

mteb/MIRACLReranking

MIRACLReranking An MTEB dataset Massive Text Embedding Benchmark MIRACL (Multilingual Information R

mteb/NollySentiBitextMining

NollySentiBitextMining An MTEB dataset Massive Text Embedding Benchmark NollySenti is Nollywood mov

mteb/MasakhaNEWSClusteringP2P

MasakhaNEWSClusteringP2P An MTEB dataset Massive Text Embedding Benchmark Clustering of news articl

mteb/AfriSentiClassification

AfriSentiClassification An MTEB dataset Massive Text Embedding Benchmark AfriSenti is the largest s