Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

yor-sarc

Domain:

natural language processing

Record type:

dataset
Creator:
toh
Host:
Yor-Sarc is a gold-standard dataset for sarcasm detection in Yorùbá, a tonal and morphologically rich low-resource African language spoken by over 50 million people. The dataset was created to address the scarcity of high-quality annotated resources for figurative language understanding in African NLP. It contains 436 manually annotated Yorùbá text instances labeled for binary sarcasm classification. Dataset Details

Visit

huggingface.co

Tasks

text classification

Languages

Yoruba

Tags

sarcasm detectiontext classificationsentiment analysisfigurative languagelow-resource NLPafrican languages

Licenses

cc-by-4.0

Similar

Yor-Sarc: A gold-standard dataset for sarcasm detection in a low-resource African languageptrdvn/kakugo-yorapertium/apertium-yornaijavoices/mms-tts-yornaijavoices/naijavoices-tts-yornaijavoices/mms-tts-yor-finetuned

Yor-Sarc: A gold-standard dataset for sarcasm detection in a low-resource African language

Sarcasm detection poses a fundamental challenge in computational semantics, requiring models to reso

ptrdvn/kakugo-yor

[Paper] [Code] [Model] A synthetically generated conversation dataset for training in Yoruba.

apertium/apertium-yor

Apertium linguistic data for Yoruba Yoruba apertium-yor ==========================================

naijavoices/mms-tts-yor

naijavoices/naijavoices-tts-yor

naijavoices/mms-tts-yor-finetuned