Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

sajib-kumar/Mitigating-Bangla-Extrinsic-Gender-Bias

Domain:

natural language processing

Record type:

dataset
Creator:
Saj
Host:
# Mitigating-Extrinsic-Gender-Bias-in-Bangla-Text-Classification In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languages. To assess this bias, we construct four manually annotated, task-specific benchmark datasets for sentiment analysis, toxicity detection, hate speech detection, and sarcasm detection. Each dataset is augmented using nuanced gender perturbations, where we systematically swap gendered names and terms while preserving semantic content, enabling minimal-pair evaluation of gender-driven prediction shifts. We then propose RandSymKL, a randomized debiasing strategy integrated with symmetric KL divergence and cross-entropy loss to mitigate the bias across task-specific pretrained models. To our knowledge, this approach is the first to integrate these elements in a unified strategy for extrinsic gender bias mitigation focused on the classification task. Our approach was evaluated against existing bias mitigation methods, with results showing that our technique not only effectively reduces bias but also maintains competitive accuracy compared to other baseline approaches. # Repository Structure The repository has three folders. **Approach** contains the code and results of all 8 approaches we applied to detect and mitigate extrinsic gender bias. In each approach subdirectory, there are two files for each task (sentiment analysis, sarcasm detection, hatespeech detection, and toxicity detection). One contains the implementation, and one contains the obtained results. Additionally it contains two files named average_kldm_ALL_tasks_lambda_1_3_5.ipynb , randomized_kldm_ALL_tasks_lambda_1_3_5.ipynb which we used for experimenting with different values of lambda for our RandSymKL and its closest baseline (AvgSymKL_MF). **Data** directory contains the datasets for all the tasks, along with a csv file that contains the gendered terms. **Gender Name Alteration** contains the …

Visit

github.com

Tasks

hate speech detectionsentiment analysistext classification

Similar

Assessing and Mitigating Bias in Medical Artificial IntelligenceEvaluating and Mitigating Inherent Linguistic Bias of African American English through InferenceMitigating Publication Bias - Views and experiences of principal investigators based in South Africa on interventions to reduce publication biassanjeev-kumar-patel/Forest-FireDeep Generative Views to Mitigate Gender Classification Bias Across Gender-Race GroupsBias in gender representation, by country and subject.

Assessing and Mitigating Bias in Medical Artificial Intelligence

Background: Deep learning algorithms derived in homogeneous populations may be poorly

Evaluating and Mitigating Inherent Linguistic Bias of African American English through Inference

Recent studies show that NLP models trained on standard English texts tend to produce biased outcome

Mitigating Publication Bias - Views and experiences of principal investigators based in South Africa on interventions to reduce publication bias

Randomised trials are essential for evidence-based healthcare, yet an estimated 25–50% remain unpubl

sanjeev-kumar-patel/Forest-Fire

A machine learning web application that predicts the Fire Weather Index (FWI) for Algerian forest re

Deep Generative Views to Mitigate Gender Classification Bias Across Gender-Race Groups

Published studies have suggested the bias of automated face-based gender classification algorithms a

Bias in gender representation, by country and subject.

Note: These figures show the predicted mean share of gendered words that are female. This measure