A large-scale dataset of 130,851 Darija (Moroccan Arabic) comments collected from YouTube and synthe
VISH-DARIJA-TTS is a synthetic Moroccan Darija text-audio dataset designed for research on vishing,
Disparate biases associated with datasets and trained classifiers in hateful and abusive content ide
The Moroccan Darija offensive language detection dataset is a human-labeled dataset consisting of a