Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

MBZUAI-Paris/DarijaMMLU

Domaine:

natural language processing

Type de record:

dataset
Créateur:
MBZ
Hôte:
DarijaMMLU is an evaluation benchmark designed to assess large language models' (LLM) performance in Moroccan Darija, a variety of Arabic. It consists of 22,027 multiple-choice questions, translated from selected subsets of the Massive Multitask Language Understanding (MMLU) and ArabicMMLU benchmarks to measure model performance on 44 subjects in Darija. Supported Tasks

Visit

huggingface.co

Tasks

language modelingquestion answering

Languages

Arabic, Algerian SpokenArabic, Moroccan Spoken

Licenses

mit

Similaires

MBZUAI-Paris/DarijaBenchMBZUAI-Paris/DarijaAlpacaEvalMBZUAI-Paris/DarijaHellaSwagMBZUAI-Paris/DarijaStoryMBZUAI-Paris/EgyptianHellaSwagMBZUAI-Paris/MoroccanWikipedia-QA

MBZUAI-Paris/DarijaBench

Note the ODC-BY license, indicating that different licenses apply to subsets of the data. This means

MBZUAI-Paris/DarijaAlpacaEval

Dataset Summary

MBZUAI-Paris/DarijaHellaSwag

DarijaHellaSwag is a challenging multiple-choice benchmark designed to evaluate machine reading comp

MBZUAI-Paris/DarijaStory

DarijaStory is a story completion dataset. It consists of 4,392 long stories scraped from 9esa, a we

MBZUAI-Paris/EgyptianHellaSwag

EgyptianHellaSwag is a challenging multiple-choice benchmark designed to evaluate machine reading co

MBZUAI-Paris/MoroccanWikipedia-QA

Dataset Summary