Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

toloka/JEEM

Domain:

natural language processing

Record type:

dataset
Creator:
tol
Host:
📊 Curated by: Toloka, MBZUAI 🌐 Language(s): Modern Standard Arabic, 🇯🇴 Jordanian dialect, 🇪🇬 Egyptian dialect, 🇦🇪 Emirati dialect, 🇲🇦 Moroccan dialect 🔍 Dataset Description

Visit

huggingface.co

Languages

Arabic, Egyptian SpokenArabic, Moroccan Spoken

Tags

Dialectal Arabic

Licenses

mit

Similar

JEEM: Vision-Language Understanding in Four Arabic DialectsChallenges of Amharic Hate Speech Data Annotation Using Yandex Toloka Crowdsourcing PlatformThe 5Js in Ethiopia: Amharic Hate Speech Data Annotation Using Toloka Crowdsourcing Platform

JEEM: Vision-Language Understanding in Four Arabic Dialects

We introduce JEEM, a benchmark designed to evaluate Vision-Language Models (VLMs) on visual understa

Challenges of Amharic Hate Speech Data Annotation Using Yandex Toloka Crowdsourcing Platform

This paper presents an Amharic hate speech annotation using the Yandex Toloka crowdsourcing platform

The 5Js in Ethiopia: Amharic Hate Speech Data Annotation Using Toloka Crowdsourcing Platform