Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Racial Bias in Hate Speech and Abusive Language Detection Datasets

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
DavBhaWeb
Hôte:avatar
Technologies for abusive language detection are being developed and applied with little consideration of their potential biases. We examine racial bias in five different sets of Twitter data annotated for hate speech and abusive language. We train classifiers on these datasets and compare the predictions of these classifiers on tweets written in African-American English with those written in Standard American English. The results show evidence of systematic racial bias in all datasets, as classifiers trained on them tend to predict that tweets written in African-American English are abusive at substantially higher rates. If these abusive language detection systems are used in the field they will therefore have a disproportionate negative impact on African-American social media users. Consequently, these systems may discriminate against the groups who are often the targets of the abuse we are trying to detect. To appear in the proceedings of the Third Abusive Language Workshop (sites.google.com) at the Annual Meeting for the Association for Computational Linguistics 2019. Please cite the published version

Visit

arxiv.org

Tasks

hate speech detectiontext classification

Tags

Computation and LanguageMachine Learning

Similaires

Intersectional Bias in Hate Speech and Abusive Language DatasetsDemoting Racial Bias in Hate Speech DetectionThe Risk of Racial Bias in Hate Speech DetectionHate Speech Detection and Racial Bias Mitigation in Social Media based on BERT modelAfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languagesismailateam/-SudaHate-Sudanese-Hate-Speech-and-Abusive-Dataset

Intersectional Bias in Hate Speech and Abusive Language Datasets

Algorithms are widely applied to detect hate speech and abusive language in social media. We investi

Demoting Racial Bias in Hate Speech Detection

In current hate speech datasets, there exists a high correlation between annotators' perceptions of

The Risk of Racial Bias in Hate Speech Detection

We investigate how annotators{'} insensitivity to differences in dialect can lead to racial bias in automatic hate speech detection models, potentially amplifying harm against minority populations. We first uncover unexpected correlations between surface markers of

Hate Speech Detection and Racial Bias Mitigation in Social Media based on BERT model

Disparate biases associated with datasets and trained classifiers in hateful and abusive content ide

AfriHate: A Multilingual Collection of Hate Speech and Abusive Language Datasets for African Languages

Hate speech and abusive language are global phenomena that need socio-cultural background knowledge

ismailateam/-SudaHate-Sudanese-Hate-Speech-and-Abusive-Dataset

A dataset of Sudanese Arabic text labeled for hate, abusive, and normal content, part of a multi-dia