Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Investigating Offensive Language Detection in a Low-Resource Setting with a Robustness Perspective

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
IsrAnaMohAsm
Éditeur:
MDP
Hôte:
Moroccan Darija, a dialect of Arabic, presents unique challenges for natural language processing due to its lack of standardized orthographies, frequent code switching, and status as a low-resource language. In this work, we focus on detecting offensive language in Darija, addressing these complexities. We present three key contributions that advance the field. First, we introduce a human-labeled dataset of Darija text collected from social media platforms. Second, we explore and fine-tune various language models on the created dataset. This investigation identifies a Darija RoBERTa-based model as the most effective approach, with an accuracy of 90% and F1 score of 85%. Third, we evaluate the best model beyond accuracy by assessing properties such as correctness, robustness and fairness using metamorphic testing and adversarial attacks. The results highlight potential vulnerabilities in the model’s robustness, with the model being susceptible to attacks such as inserting dots (29.4% success rate), inserting spaces (24.5%), and modifying characters in words (18.3%). Fairness assessments show that while the model is generally fair, it still exhibits bias in specific cases, with a 7% success rate for attacks targeting entities typically subject to discrimination. The key finding is that relying solely on offline metrics such as the F1 score and accuracy in evaluating machine learning systems is insufficient. For low-resource languages, the recommendation is to focus on identifying and addressing domain-specific biases and enhancing pre-trained monolingual language models with diverse and noisier data to improve their robustness and generalization capabilities in diverse linguistic scenarios.

Visit

doi.org

Tasks

hate speech detectiontext classification

Languages

Arabic, Algerian SpokenArabic, Moroccan Spoken

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

A Forensic Linguistic Dataset for Offensive Content Detection in Low-Resource Language: SetswanaDetection of Offensive Language and ITS Severity for Low Resource LanguageCross-lingual NER robustness in low-resource languages with source language diversityCorrelation between Source Language Diversity and Synthetic Data Robustness in Low-Resource Grammatical Error Detectiona-ibrahimi/Moroccan-Darija-Offensive-Language-Detection-DatasetPACMAN: a framework for pulse oximeter digit detection and reading in a low-resource setting

A Forensic Linguistic Dataset for Offensive Content Detection in Low-Resource Language: Setswana

Developing Monolingual Setswana Datasets for Offensive Content Detection Reproducibility Package, Me

Detection of Offensive Language and ITS Severity for Low Resource Language

Continuous proliferation of hate speech in different languages on social media has drawn significant

Cross-lingual NER robustness in low-resource languages with source language diversity

Multilingual Language Models (MLLMs) exhibit robust cross-lingual transfer capabilities, or the abil

Correlation between Source Language Diversity and Synthetic Data Robustness in Low-Resource Grammatical Error Detection

Grammatical Error Detection (GED) methods rely heavily on human annotated error corpora. However, th

a-ibrahimi/Moroccan-Darija-Offensive-Language-Detection-Dataset

The Moroccan Darija Offensive Language Detection Dataset is a human-labeled dataset consisting of Mo

PACMAN: a framework for pulse oximeter digit detection and reading in a low-resource setting

In light of the COVID-19 pandemic, patients were required to manually input their daily oxygen satur