Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
zagBisAldBes
Éditeur:
arXiv
Hôte:avatar
We present ArabicDialectSafety, a human-curated Arabic safety dataset of 25,071 prompts covering six Arabic varieties: Modern Standard Arabic, Syrian, Egyptian, Algerian, Palestinian, and Moroccan. The dataset is annotated with dialect labels and seven fine-grained harm categories. We introduce a dual-task evaluation framework for binary safe/unsafe detection and granular harm classification across dialects. Benchmarking seven supervised and generative models, we find that fine-tuned MARBERTv2 achieves the strongest performance, with Macro-F1 scores of 0.95 for binary classification and 0.90 for granular classification, substantially outperforming prompted frontier LLMs, including Arabic-specialized models. Our analyses show that dialect conditioning is most effective when integrated at the representation level, while significant performance gaps remain for low-resource Maghrebi dialects. We further evaluate seven frontier LLMs as response generators on harmful dialectal Arabic prompts and observe unsafe generation rates below 5 percent across models. We release the dataset and code upon acceptance to support future research on dialect-aware Arabic safety evaluation. Warning: This paper contains examples of harmful and potentially offensive content included solely for research purposes. 13 pages, 2 figures, 9 tables

Visit

doi.org

Tasks

hate speech detectiontext classification

Languages

Arabic, Algerian Spoken

Tags

Computation and Language (cs.CL)FOS: Computer and information sciences

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Sudanese Arabic Dialect Identification BenchmarkVoxArabica: A Robust Dialect-Aware Arabic Speech Recognition SystemSentiment Analysis on Arabic Dialects: A Multi-Dialect BenchmarkDeep Language Detection for Indian Code: A Context-Aware Deep Learning for Multilingual Content ClassificationReliable Generative Data Augmentation for Arabic Text Classification: A Length-Aware Data-Centric ApproachImproving Arabic Dialect Processing in IoT Systems: A Comparative Study of Baseline and Dialect-Aware AI Models

Sudanese Arabic Dialect Identification Benchmark

A curated benchmark dataset for Arabic dialect identification, with special focus on Sudanese Arabic

VoxArabica: A Robust Dialect-Aware Arabic Speech Recognition System

Arabic is a complex language with many varieties and dialects spoken by over 450 millions all around

Sentiment Analysis on Arabic Dialects: A Multi-Dialect Benchmark

Deep Language Detection for Indian Code: A Context-Aware Deep Learning for Multilingual Content Classification

The multilingual character of India has been seen in the online discussions whereby Hindi, gujarati,

Reliable Generative Data Augmentation for Arabic Text Classification: A Length-Aware Data-Centric Approach

Data augmentation has become a critical strategy for improving text classification performance, part

Improving Arabic Dialect Processing in IoT Systems: A Comparative Study of Baseline and Dialect-Aware AI Models

Context: Bringing voice-controlled interfaces into Internet of Things (IoT) systems has created fres