Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Towards a New Lexicon-Based Features Vector for Sentiment Analysis: Application to Moroccan Arabic Tweets

Domain:

natural language processing

Record type:

paperdataset
Creator:
GarKha
Editor:
LabSysUniLab
Publisher:
CCSDSpringer International Publishing
Host:avatar
International audience The emergence of the Web 2.0 technology generated a huge amount of raw data by enabling Internet users to post their opinions and reviews on the web. This data plays an important role in decision making for many peoples and organizations. An example of valuable insights that can be extracted from user’s posts is their opinions and sentiments regarding topics, events, services, products, etc. The English language has been the subject of extensive research on sentiment analysis. The proposed solutions are largely dominated by the use of two main analysis approaches based on machine learning techniques and the lexical approach. This work focuses on the second one to analyze the sentiments expressed in Moroccan tweets written in Arabic language : Standard Arabic (SA) and Moroccan Dialect (MD), and proposes a new method for extracting characteristics and representing data. The main idea of this method is to represent the text as a weight vector of feelings. Due to the lack of resources (databases and lexicon dictionaries) for the Arabic language, especially for the Moroccan one, this work starts with the construction of a corpus of 18.000 valid tweets based on 36 114 collected tweets that are manually tagged and classified as MD or SA. Then describes the steps of the construction of the Moroccan Senti-lexicon, a dictionary of 30.000 words labeled as positive, negative or neutral. The results of this study prove to be superior to those obtained by other comparable state of the art approaches.

Visit

hal.science

Tasks

sentiment analysistext classification

Languages

Arabic, Moroccan Spoken

Tags

[INFO]Computer Science [cs]

Similar

A Proposed Lexicon-Based Sentiment Analysis Approach for the Vernacular Algerian ArabicMSA-Moroccan Dialect: A Multimodal Sentiment Analysis Dataset for Moroccan Arabic (Darija)Tweets for sentiment analysisA Sentiment analysis approach for Arabic dialects texts analysis based on automatic translation: Application to the Algerian dialect.AUTOMATED SPEECH SENTIMENT ANALYSIS FOR MOROCCAN DIALECT SPEAKERS USING DEEP LEARNING AND MFCC-BASED FEATURESHausa Lexicon-based Sentiment Analysis Dataset for NLP

A Proposed Lexicon-Based Sentiment Analysis Approach for the Vernacular Algerian Arabic

MSA-Moroccan Dialect: A Multimodal Sentiment Analysis Dataset for Moroccan Arabic (Darija)

This dataset provides the first publicly available multimodal resource for sentiment analysis in Mor

Tweets for sentiment analysis

This is the dataset used in the research manuscript “Sentiment Analysis of Tweets: Political Climate

A Sentiment analysis approach for Arabic dialects texts analysis based on automatic translation: Application to the Algerian dialect.

AUTOMATED SPEECH SENTIMENT ANALYSIS FOR MOROCCAN DIALECT SPEAKERS USING DEEP LEARNING AND MFCC-BASED FEATURES

Abstract Sentiment analysis is used in several fields, such as teleconsultations in the medical fie

Hausa Lexicon-based Sentiment Analysis Dataset for NLP

NLP Data for the Hausa Language