Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Multi-Pipeline Approach for Sentiment Analysis of West-African Pidgin

Domain:

natural language processing

Record type:

paper
Creator:
OguAinAyeLaw
Publisher:
Zenodo
Host:avatar
This paper presents a multi-pipeline approach to sentiment analysis, with the aim of improving both the accuracy and relevance of the results. Sentiment analysis of West African Pidgin has historically been fragmented, often involving the training of new models with pidgin data, and typically focusing on general sentiment polarity. This study seeks to address these gaps by adopting a holistic, multi-pipeline system for the comprehensive analysis of sentiment polarity in Pidgin text. A subject classifier was developed using the Logistic Regression algorithm to predict the relevance of a body of text to the specific subject matter. Data was collected from Twitter and processed into tokens, which were then used for training and evaluation. This enabled the model to handle a wide range of informal and context-specific words commonly found in Pidgin. For the sentiment analysis itself, a cross-lingual model, RoBERTa (XLM-R), was fine-tuned and expanded through transfer learning using the AfriBERTa model, developed by Ogueji et. al (2021). This fine-tuned model achieved an average F1-score of 74.5 over five runs, demonstrating its effectiveness in sentiment classification. The subject classifier also performed efficiently, achieving an accuracy of 0.81 in identifying relevant text. This multi-pipeline system demonstrates significant promise in enhancing sentiment analysis for Pidgin text, being the first to combine subject classification and cross-lingual sentiment analysis techniques. The results show that the proposed approach can be a valuable tool in natural language processing for underrepresented languages such as West African Pidgin.

Visit

doi.orgzenodo.org

Tasks

sentiment analysistext classificationtransfer learning

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

679190375/Sentiment-Analysis-for-Cameroonian-Pidgin-EnglishSynthetic Data Generation Pipeline for Low-Resource Swahili Sentiment Analysis: Multi-LLM Judging with Human ValidationIammteo/Sentiment-analysis-model-for-Nigerian-pidgin-Low-resource-language-tabularis-ai/Synthetic-Data-Generation-Pipeline-for-Low-Resource-Swahili-Sentiment-AnalysisCodeHermez/African-Langs-For-Sentiment-Analysis-Using-AfriSenti-Datasets-Sentiment-Analysis-In-African-Langs-SanaNGU/Sentiment-Analysis-for-African-Languages

679190375/Sentiment-Analysis-for-Cameroonian-Pidgin-English

NLP-based sentiment analysis for Cameroonian Pidgin English text classification. ## 🗣️ Sentiment An

Synthetic Data Generation Pipeline for Low-Resource Swahili Sentiment Analysis: Multi-LLM Judging with Human Validation

Iammteo/Sentiment-analysis-model-for-Nigerian-pidgin-Low-resource-language-

This is a machine learning project of a Nigerian Pidgin sentiment analysis model specifically design

tabularis-ai/Synthetic-Data-Generation-Pipeline-for-Low-Resource-Swahili-Sentiment-Analysis

# Synthetic Data for Low-Resource Swahili Language Sentiment Analysis This repository contains the

CodeHermez/African-Langs-For-Sentiment-Analysis-Using-AfriSenti-Datasets-Sentiment-Analysis-In-African-Langs-

A Comparative Study of Monolingual and Multilingual Transfer Learning Strategies with Code-Mixing An

SanaNGU/Sentiment-Analysis-for-African-Languages

This Repo describe the code for SemEval 23 Task 12 : AfriSenti-SemEval Shared Task 12 ### To run th