Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Camer-Hate-FR: An Annotated Dataset for Hate Speech Detection in Cameroonian French.

Domain:

natural language processing

Record type:

dataset
Creator:
MesNzeFOTTch
Editor:
UNI
Publisher:
Men
Host:avatar
This dataset, titled Camer-Hate-FR, provides a valuable resource for detecting hate speech within the unique linguistic context of Cameroonian French. The data consists of 46,825 messages collected between January and June 2025 from public Cameroonian social media sources, including Facebook pages, YouTube channels, and WhatsApp groups. Existing hate speech detection models, primarily trained on standard European French, perform poorly on Cameroonian data due to the prevalent use of local slang, code-switching with English and indigenous languages (Camfranglais), and nuanced cultural contexts. This dataset was created to address this gap. Each message has been manually annotated by three native speakers as either 'hateful' (haineux) or 'non-hateful' (non_haineux), with the final label determined by a majority vote. The dataset is provided as a single CSV file and includes the original text, the annotation counts, the final vote, and the justifications provided by annotators. All data has been fully anonymized to protect user privacy. This resource is designed to train, validate, and benchmark machine learning models for content moderation, facilitate sociolinguistic analysis, and spur the development of more inclusive and effective NLP technologies for Francophone Africa.

Visit

doi.orgdata.mendeley.com

Tasks

code switchinghate speech detectiontext classification

Tags

Computer ScienceComputational LinguisticsSocial MediaNatural Language ProcessingMachine LearningCameroonAfrica CultureSentiment Analysis

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

HausaHate: An Expert Annotated Corpus for Hausa Hate Speech DetectionAmharic dataset for hate speech detectionAmharic Facebook Dataset for Hate Speech detectionHateTune: Tunisian Dialect Hate Speech Detection DatasetLuckilyeee/Hate-Speech-DetectionLusanji/Hate-Speech-Detection

HausaHate: An Expert Annotated Corpus for Hausa Hate Speech Detection

We introduce the first expert annotated corpus of Facebook comments for Hausa hate speech detection.

Amharic dataset for hate speech detection

the dataset is collected from social media such as facebook and telegram. the dataset is further pro

Amharic Facebook Dataset for Hate Speech detection

This dataset is collected from Facebook pages of activists who write their posts using Geez script a

HateTune: Tunisian Dialect Hate Speech Detection Dataset

Luckilyeee/Hate-Speech-Detection

Achieving Hate Speech Detection in a Low Resource Setting # Achieving Hate Speech Detection in a Lo

Lusanji/Hate-Speech-Detection

Hate speech detection in audio for English and Kiswahili languages # Automatic hate speech detectio