Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Inculcating Context for Emoji Powered Bengali Hate Speech Detection using Extended Fuzzy SVM and Text Embedding Models

Domaine:

natural language processing

Type de record:

paper
Créateur:
SayAmiDevVar
Éditeur:
Ass
Hôte:
The massive growth of social webs offer opportunities to communicate with diverse languages, unstructured text, informal posts, misspelled contents and emojis. Social media users feel comfortable to express their emotions specially emotions with high intensity (hate speech) in their mother tongue. Hate speech in any form targets groups and individuals that may trigger antisocial activities, hate crimes, and terrorist acts. Bengali social media users use Bengali for posting implicit or indirect hate text. Existing Bengali hate speech detection research considers explicit hate speech detection but in actual hate is expressed more in implicit way. In order to detect both implicit and explicit hate speech from low resource content, social webs need highly efficient automated tools. Researchers applied discriminative learning approaches (i.e. SVM, MLP, CNN) to distinguish hate text with only clear-cut outcomes in detecting direct hate speech. The proposed novel Bengali hate speech detection model considers two parallel approaches: (i) It applies extended fuzzy SVM classifier for class imbalanced dataset (FSVMCIL) and multilingual BERT (mBERT) text embedding model to detect first hate label; (ii) Morphological analysis method to detect implicit and explicit hate content with the hate similarity (HS) scheme for second hate label. Linking both labeling methods, this research extracts contextual Bengali hate speech from informal text. This novel HS method considers Word2Vec word embedding model and Bengali hate lexicon. It also considers emoji to text conversion for efficient contextual analysis. This study also conducts extensive experiments for various categories with the Bengali hate speech dataset. It also evaluates the proposed model performance considering weighted F1 score, precision, recall and accuracy parameters. Results reveal significant improvement in Bengali hate speech detection with 2.35% increase in F1- score and 9.11 % increase in accuracy.

Visit

doi.org

Tasks

hate speech detectiontext classification

Similaires

Combining FastText and Glove Word Embedding for Offensive and Hate speech Text DetectionHate Speech and Offensive Language Detection in BengaliDetecting Online Hate Speech Using Context Aware ModelsHATE SPEECH DETECTION ON SOCIAL MEDIA FOR AMHARIC TEXT USING DEEP LEARNING APPROACHA Context-Aware and Target-Adaptive Multilingual Framework for Hate Speech Detection in Code-Switched Social Media TextAmharic text dataset extracted from memes for hate speech detection or classification: Amharic Language Hate Speech Detection System from Facebook Image Post Using Deep Learning System

Combining FastText and Glove Word Embedding for Offensive and Hate speech Text Detection

Combining FastText and Glove Word Embedding for Offensive and Hate speech Text Detection

Poster presented at the Deep Learning Indaba 2022 by Nabil BADRI

Hate Speech and Offensive Language Detection in Bengali

Social media often serves as a breeding ground for various hateful and offensive content. Identifyin

Detecting Online Hate Speech Using Context Aware Models

In the wake of a polarizing election, the cyber world is laden with hate speech. Context accompanyin

HATE SPEECH DETECTION ON SOCIAL MEDIA FOR AMHARIC TEXT USING DEEP LEARNING APPROACH

HATE SPEECH DETECTION ON SOCIAL MEDIA FOR AMHARIC TEXT USING DEEP LEARNING APPROACH

A Context-Aware and Target-Adaptive Multilingual Framework for Hate Speech Detection in Code-Switched Social Media Text

The rapid expansion of social media has accelerated the spread of hate speech, particularly within m

Amharic text dataset extracted from memes for hate speech detection or classification: Amharic Language Hate Speech Detection System from Facebook Image Post Using Deep Learning System

the dataset is collected from social media such as facebook and telegram. the dataset is further pro