Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Dialect Bias in Arabic Toxicity Detection Under Controlled Distributional Shifts

Domaine:

natural language processing

Type de record:

paper
Créateur:
Mah
Éditeur:
MDP
Hôte:
Toxicity detection models can behave inconsistently across Arabic dialects, yet the level of training-data imbalance at which dialect-conditioned bias becomes statistically detectable remains unquantified. This paper presents a controlled framework for measuring such bias under class-conditional distributional shift. Four transformer encoders (AraBERT, MARBERT, CAMeLBERT-DA, and XLM-RoBERTa) are fine-tuned on a balanced toxicity corpus comprising texts in the Saudi, Egyptian, and Tunisian dialects across five random seeds, after which the representation of Saudi-dialect texts in the toxic class is experimentally increased at three injection levels, λ∈{0.10,0.20,0.30}, under two designs: one that jointly alters toxic-class composition and class balance, and one that isolates toxic-class dialectal composition. Bias emergence is quantified through a statistical bias-sensitivity threshold BSTB, the lowest injection level at which the Saudi-centered false-positive-rate gap is positive and statistically significant under a permutation test. False-negative-rate disparities are additionally evaluated, while ERASER-based comprehensiveness and sufficiency measures and SHAP scenario analysis serve as exploratory attribution diagnostics. Under the first design, BSTB is reached at λ=0.20 or λ=0.30 in most model–seed runs. Under the second design, no model–seed run reaches BSTB within the tested range. The matched-control analysis detects no significant marker-specific comprehensiveness effect.

Visit

doi.org

Tasks

hate speech detectiontext classification

Languages

Arabic, Tunisian Spoken

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

AraTox: A Multi-Dialect, Multi-Label Arabic Dataset for Toxicity DetectionNizar-Charrada/Tunisian-Dialect-Toxicity-Detectiontesnimmmeee/Tunisian-Dialect-Toxicity-Detection-using-Fine_Tuned_BertAutomatic Dialect Detection in Arabic Broadcast Speech0khacha/darija-toxicity-detectionSentiment Analysis in Poems in Misurata Sub-dialect -- A Sentiment Detection in an Arabic Sub-dialect

AraTox: A Multi-Dialect, Multi-Label Arabic Dataset for Toxicity Detection

AraTox is a multi-dialect, multi-label Arabic dataset for toxicity detection. It contains annotated

Nizar-Charrada/Tunisian-Dialect-Toxicity-Detection

Tunisian Dialect-Specific Toxicity Detection # Tunisian Dialect Toxicity Detection **Note: This AI

tesnimmmeee/Tunisian-Dialect-Toxicity-Detection-using-Fine_Tuned_Bert

new repo # Tunisian-Dialect-Toxicity-Detection-using-Fine_Tuned_Bert ### README for Tunisian-Dialec

Automatic Dialect Detection in Arabic Broadcast Speech

We investigate different approaches for dialect identification in Arabic broadcast speech, using pho

0khacha/darija-toxicity-detection

NLP pipeline for toxicity detection in Moroccan Darija and Arabizi. Powered by a fine-tuned ArabERT

Sentiment Analysis in Poems in Misurata Sub-dialect -- A Sentiment Detection in an Arabic Sub-dialect

Over the recent decades, there has been a significant increase and development of resources for Arab