This dataset is a comprehensive, carefully curated collection of text data specifically for Moroccan
This dataset is designed for text classification in Moroccan Darija, a dialect spoken in Morocco. It
This dataset has been created with Argilla. As shown in the sections below, this dataset can be load
AtlasOCRBench is a comprehensive evaluation benchmark tailored specifically for Moroccan Darija (Moroccan Arabic dialect) OCR tasks. This dataset was created to measure the real-world performance of OCR models on Darija text, addressing the unique challenges posed