Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Amharic text dataset extracted from memes for hate speech detection or classification

Domain:

natural language processing

Record type:

dataset
Creator:
HayAbemeqmeq
Editor:
meq
Publisher:
Men
Host:avatar
the dataset is collected from social media such as facebook and telegram. the dataset is further processed. the collection are orginal_cleaned: this dataset is neither stemed nor stopword are remove: stopword_removed: in this dataset stopwords are removed but not stemmed and in stemed datset is stemmed and stopwords are removed. stemming is done using hornmorpho developed by Michael Gesser( available at HornMorpho) all datasets are normalized and free from noise such as punctuation marks and emojs.

Visit

doi.orgdata.mendeley.com

Tasks

hate speech detectiontext classification

Languages

Amharic

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode