Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Improving Zero-Shot Cross-Lingual Hate Speech Detection with Pseudo-Label Fine-Tuning of Transformer Language Models

Domain:

natural language processing

Record type:

paper
Creator:
HarIgnArkGar
Publisher:
Ass
Host:
Hate speech has proliferated on social media platforms in recent years. While this has been the focus of many studies, most works have exclusively focused on a single language, generally English. Low-resourced languages have been neglected due to the dearth of labeled resources. These languages, however, represent an important portion of the data due to the multilingual nature of social media. This work presents a novel zero-shot, cross-lingual transfer learning pipeline based on pseudo-label fine-tuning of Transformer Language Models for automatic hate speech detection. We employ our pipeline on benchmark datasets covering English (source) and 6 different non-English (target) languages written in 3 different scripts. Our pipeline achieves an average improvement of 7.6% (in terms of macro-F1) over previous zero-shot, cross-lingual models. This demonstrates the feasibility of high accuracy automatic hate speech detection for low-resource languages. We release our code and models at github.com.

Visit

doi.org

Tasks

hate speech detectiontext classificationtransfer learning

Similar

Label modification and bootstrapping for zero-shot cross-lingual hate speech detectionMultimodal Pretraining and Sequential Fine-Tuning for Zero-Shot Cross-Lingual Euphemism DetectionContrastive Learning Effects on Zero-Shot Cross-Lingual Euphemism Detection in XLM-R Fine-TuningSequential Fine-Tuning Language Variation in Zero-Shot Euphemism DetectionCross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on RomanianDomain Similarity Impact on Multilingual Hate Speech Detection Generalization in Zero-Shot Cross-Lingual Transfer

Label modification and bootstrapping for zero-shot cross-lingual hate speech detection

The goal of hate speech detection is to filter negative online content aiming at certain groups of p

Multimodal Pretraining and Sequential Fine-Tuning for Zero-Shot Cross-Lingual Euphemism Detection

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Contrastive Learning Effects on Zero-Shot Cross-Lingual Euphemism Detection in XLM-R Fine-Tuning

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Sequential Fine-Tuning Language Variation in Zero-Shot Euphemism Detection

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian

Relation extraction (RE) for low-resource languages is typically constrained by the lack of annotate

Domain Similarity Impact on Multilingual Hate Speech Detection Generalization in Zero-Shot Cross-Lingual Transfer

Automatic detection of abusive online content such as hate speech, offensive language, threats, etc.