Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Findings: Suum Cuique: Studying Bias in Taboo Detection with a Community Perspective

Domain:

natural language processing

Record type:

paper
Creator:
AssKhaRusSri
Publisher:
Und
Host:avatar
Prior research has discussed and illustrated the need to consider linguistic norms at the community level when studying taboo (hateful/offensive/toxic etc.) language. However, a methodology for doing so, that is firmly founded on community language norms is still largely absent. This can lead both to biases in taboo text classification and limitations in our understanding of the causes of bias. We propose a method to study bias in taboo classification and annotation where a community perspective is front and center. This is accomplished by using special classifiers tuned for each community's language. In essence, these classifiers represent community level language norms. We use these to study bias and find, for example, biases are largest against African Americans (7/10 datasets and all 3 classifiers examined). In contrast to previous papers we also study other communities and find, for example, strong biases against South Asians. In a small scale user study we illustrate our key idea which is that common utterances, i.e., those with high alignment scores with a community (community classifier confidence scores) are unlikely to be regarded taboo. Annotators who are community members contradict taboo classification decisions and annotations in a majority of instances. This paper is a significant step toward reducing false positive taboo decisions that over time harm minority communities.

Visit

doi.orgunderline.io

Tasks

hate speech detectiontext classification

Similar

Exploring the Racial Bias in Pain Detection with a Computer Vision ModelAlgorithmic Bias in Recidivism Prediction: A Causal PerspectiveInvestigating Offensive Language Detection in a Low-Resource Setting with a Robustness PerspectiveDemoting Racial Bias in Hate Speech DetectionBias detection and mitigation in Recommendation systemsAnnotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection

Exploring the Racial Bias in Pain Detection with a Computer Vision Model

People detect painful expressions more easily in members of their racial ingroup than outgroup. Here

Algorithmic Bias in Recidivism Prediction: A Causal Perspective

ProPublica's analysis of recidivism predictions produced by Correctional Offender Management Profili

Investigating Offensive Language Detection in a Low-Resource Setting with a Robustness Perspective

Moroccan Darija, a dialect of Arabic, presents unique challenges for natural language processing due

Demoting Racial Bias in Hate Speech Detection

In current hate speech datasets, there exists a high correlation between annotators' perceptions of

Bias detection and mitigation in Recommendation systems

Bias detection and mitigation in Recommendation systems

Poster presented at the Deep Learning Indaba 2023 by Nadiera Mustapha

Annotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection

The perceived toxicity of language can vary based on someone's identity and beliefs, but this variat