Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

atlasia/Darija_LID_Anootation_10k

Domain:

natural language processing

Record type:

dataset
Creator:
atl
Host:
This dataset has been created with Argilla. As shown in the sections below, this dataset can be loaded into your Argilla server as explained in Load with Argilla, or used directly with the datasets library in Load with datasets. To load with Argilla, you'll just need to install Argilla as pip install argilla --upgrade and then use the following code: import argilla as rg

Visit

huggingface.co

Tasks

language identification

Languages

Arabic, Algerian Spoken

Tags

rlfhargillahuman-feedback

Similar

atlasia/Atlasetatlasia/moroccan_darija_domain_classifier_datasetatlasia/darija_bible_alignedatlasia/bible_darija_text_onlyatlasia/moroccan_darija_corpusatlasia/AtlasOCRBench

atlasia/Atlaset

This dataset is a comprehensive, carefully curated collection of text data specifically for Moroccan

atlasia/moroccan_darija_domain_classifier_dataset

This dataset is designed for text classification in Moroccan Darija, a dialect spoken in Morocco. It

atlasia/darija_bible_aligned

This dataset contains aligned audio segments from the Moroccan Arabic (Darija) Bible translation, sp

atlasia/bible_darija_text_only

atlasia/moroccan_darija_corpus

atlasia/AtlasOCRBench

AtlasOCRBench is a comprehensive evaluation benchmark tailored specifically for Moroccan Darija (Moroccan Arabic dialect) OCR tasks. This dataset was created to measure the real-world performance of OCR models on Darija text, addressing the unique challenges posed