This dataset card aims to be a base template for new datasets. It has been generated using this raw
A large-scale dataset of 130,851 Darija (Moroccan Arabic) comments collected from YouTube and synthe
VISH-DARIJA-TTS is a synthetic Moroccan Darija text-audio dataset designed for research on vishing,
The Moroccan Darija Wiki Dataset consists of 10,044 parallel text samples of Moroccan Darija sourced
Darija Open Dataset (DODa) is an open-source project for the Moroccan dialect. With more than 10,000
Darija (Moroccan Arabic) Stories Dataset is a large-scale collection of stories written in Moroccan