Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

EMOTION DETECTION AND CLASSIFICATION ON TIGRIGNA SOCIAL MEDIA TEXTS USING TRANSFORMER MODELS

Domain:

natural language processing

Record type:

datasetpaper
Creator:
Geb
Editor:
MekMek
Publisher:
Mek
Host:avatar
The rapid growth of social media has reshaped emotional expression, producing large-scale digital data for social, cultural, and political analysis, thereby highlighting the importance of reliable automated emotion detection tools. Despite advances in Natural Language Processing (NLP), Tigrigna remains underrepresented, with existing multilingual models often underperforming due to limited annotated data, lack of tailored resources, and linguistic complexity. To address this gap, this study introduces transformer-based models tailored for emotion detection and classification in Tigrigna social media texts, focusing on four emotion categories: happiness, sadness, neutral, and disgust. A total of 4,000 Tigrigna sentences were collected from Facebook and YouTube and manually annotated with a high Inter-Annotator Agreement. To expand and balance the corpus, 6,000 additional sentences were generated using data augmentation techniques, including backtranslation and synonym replacement, resulting in a final dataset of 10,000 sentences. Following preprocessing, including normalization, tokenization, and cleaning, the data was split into training (8,000), validation (1,000), and testing (1,000) subsets. Three transformer-based models namely XLM-RoBERTa, tiBERT, and the Tigrigna-specific tiRoBERTa were fine-tuned and evaluated using Macro-F1, precision, and recall metrics to address class imbalance. The results demonstrated progressive improvements across models: XLM-R achieved an F1-score of 81%, tiBERT 84.4%, and tiRoBERTa 88%, with tiRoBERTa outperforming the others across all emotion categories, particularly in distinguishing subtle distinctions between sadness and happiness. Misclassifications between neutral and disgust persisted, reflecting data-related issues, model-specific challenges, and the low-resource nature of Tigrigna. Data augmentation improved F1-scores by 2–10% across models, underscoring its crucial role in enhancing performance in low-resource NLP tasks. The study concludes that transformer models, when culturally and linguistically adapted, are highly effective for Tigrigna emotion detection. Future research should expand Tigrigna-specific pretraining corpora, explore advanced augmentation, investigate hybrid architectures, and integrate multimodal data (e.g., combining text with images or videos). Applying these findings via APIs and dashboards can support researchers, policymakers, and organizations in leveraging Tigrigna social media for informed decision-making.

Visit

doi.orgrepository.mu.edu.et

Tasks

emotion identificationtext classification

Languages

Tigrigna

Tags

Data augmentationEmotion detectionLow-resource NLPTigrignaSocial mediaTransformer modelstiRoBERTa

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

Multi-label Emotion Classification on Social Media Comments using Deep learningHATE SPEECH DETECTION FOR TIGRIGNA LANGUAGE ON SOCIAL MEDIA USING DEEP LEARNINGDetection of Somali-written Fake News and Toxic Messages on the Social Media Using Transformer-based Language ModelsCUET_Novice@DravidianLangTech 2025: Abusive Comment Detection in Malayalam Text Targeting Women on Social Media Using Transformer-Based ModelsEmotion Classification for Amharic Social Media Text Comments Using Deep LearningExploring Amharic Sentiment Analysis from Social Media Texts: Building Annotation Tools and Classification Models.

Multi-label Emotion Classification on Social Media Comments using Deep learning

Abstract Social media is an online platform that people use to develop social networks or

HATE SPEECH DETECTION FOR TIGRIGNA LANGUAGE ON SOCIAL MEDIA USING DEEP LEARNING

Hate speech on social media poses a significant challenge to online safety and social harmony, with

Detection of Somali-written Fake News and Toxic Messages on the Social Media Using Transformer-based Language Models

The fact that everyone with a social media account can create and share content, and the increasing

CUET_Novice@DravidianLangTech 2025: Abusive Comment Detection in Malayalam Text Targeting Women on Social Media Using Transformer-Based Models

Social media has become a widely used platform for communication and entertainment, but it has also

Emotion Classification for Amharic Social Media Text Comments Using Deep Learning

Exploring Amharic Sentiment Analysis from Social Media Texts: Building Annotation Tools and Classification Models.

This paper presents the study of sentiment analysis for Amharic social media texts. As the number of