This dataset card aims to be a base template for new datasets. It has been generated using this raw
Hear Morocco: Audio Dataset with Dual Script Transcripts
The Moroccan Darija Wiki Dataset consists of 10,044 parallel text samples of Moroccan Darija sourced
Darija Open Dataset (DODa) is an open-source project for the Moroccan dialect. With more than 10,000
Darija (Moroccan Arabic) Stories Dataset is a large-scale collection of stories written in Moroccan
The Moroccan Darija Wiki Audio Dataset consists of 551 parallel text and speech samples of Moroccan