DarijaHellaSwag is a challenging multiple-choice benchmark designed to evaluate machine reading comp
DarijaStory is a story completion dataset. It consists of 4,392 long stories scraped from 9esa, a we
Note the ODC-BY license, indicating that different licenses apply to subsets of the data. This means
EgyptianHellaSwag is a challenging multiple-choice benchmark designed to evaluate machine reading co
Dataset Summary
DarijaMMLU is an evaluation benchmark designed to assess large language models' (LLM) performance in