Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Sentiment Classification for Under-Resourced Language Using Word2Vec Neural Network: Amharic Language Social Media Text

Domain:

natural language processing

Record type:

paper
Creator:
ZewJen
Publisher:
Zenodo
Host:avatar
Sentiment classification becomes popular task in social network texts which express opinions on different issue to analyze and produce useful knowledge. However, many linguistic computational resources are available only for English language. In the recent years, due to the emergence of social media platforms, opinion-rich resources are booming abundant for under-resourced languages with the need to perform Sentiment Analysis. On the other hand,most of the existing researches focus on how to extract the effective features, such as lexical and syntactic features,while limited work has been done on semantic features, which can make more contributions to both under-resourced and resourceful languages. In this paper, we proposed sentiment classification based on Word2Vec for Amharic Language text on political domain. The Word2Vec establishes the neural network models to learn the vector representations of words to extract the deep semantic relationships. Firstly, we cluster the similar features together and apply language modeling Ngram to check sentiment-bearing Co-occurring Terms (COT). Word2Vec and TF-IDF were used to learn the word representations as a candidate feature vector. Secondly, The Gradient-Boosting Tree(GBT) and Random forest machine learning classifiers were used to train and test in the Apache Spark platform. In our experiments, we use the Amharic language in Ethiopia and adopt a standard natural language pre-processing techniques on the crawled Facebook datasets to categorize into positive and negative opinions. Experimental results of feature extraction using Word2Vec technique performs better in the GBT classifier achieving an average accuracy of 82.29%. Therefore, our proposed approach can successfully discriminate among posts and comments expressing positive and negative opinions.

Visit

doi.orgzenodo.org

Tasks

embeddingssentiment analysistext classification

Languages

Amharic

Tags

Amharic text SentimentWord2Vec SemanticSocial MediaUnder-resourced Language

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcodeOpen Accessinfo:eu-repo/semantics/openAccess

Similar

SENTIMENT CLASSIFICATION FOR SOCIAL MEDIA AMHARIC TEXT USING DEEP NEURAL NETWORK APPROACH: POLITICS DOMAINSentiment Classification from Social Media Amharic Text Using a Deep Neural Network Approach: Politics DomainDocument Classification for the Under-resourced Amharic LanguageDeep Learning-Based Sentiment Classification of Social Network Texts in Amharic LanguageEmotion Classification for Amharic Social Media Text Comments Using Deep LearningShort Text Language Identification for Under Resourced Languages

SENTIMENT CLASSIFICATION FOR SOCIAL MEDIA AMHARIC TEXT USING DEEP NEURAL NETWORK APPROACH: POLITICS DOMAIN

Social media is a tool that political parties utilize to genuinely communicate with the public about

Sentiment Classification from Social Media Amharic Text Using a Deep Neural Network Approach: Politics Domain

Document Classification for the Under-resourced Amharic Language

NLP is severely hampered by a scarcity of digital resources. This is especially true for Amharic, a

Deep Learning-Based Sentiment Classification of Social Network Texts in Amharic Language

Emotion Classification for Amharic Social Media Text Comments Using Deep Learning

Short Text Language Identification for Under Resourced Languages

The paper presents a hierarchical naive Bayesian and lexicon based classifier for short text languag