Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Tigrinya Abusive Language Detection (TiALD) Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
fga
Host:
TiALD is a large-scale, multi-task benchmark dataset for abusive language detection in the Tigrinya language. It consists of 13,717 YouTube comments annotated for abusiveness, sentiment, and topic tasks. The dataset includes comments written in both the Ge’ez script and prevalent non-standard Latin transliterations to mirror real-world usage.

Visit

huggingface.co

Tasks

hate speech detectiontext classification

Languages

GeezTigrigna

Tags

tigrinyaabusive-language-detectionhate-speech-detectiontopic-classificationsentiment-analysislow-resource

Licenses

cc-by-4.0