Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

MSA-Moroccan Dialect: A Multimodal Sentiment Analysis Dataset for Moroccan Arabic (Darija)

Domaine:

natural language processing

Type de record:

dataset
Créateur:
AyoNfa
Éditeur:
Uni
Éditeur:
Men
Hôte:avatar
This dataset provides the first publicly available multimodal resource for sentiment analysis in Moroccan Arabic (Darija). It consists of 3,040 samples extracted from authentic Moroccan podcasts, each containing aligned text (transcript), audio (speech waveform), and visual (features) modalities. Samples are manually annotated with sentiment labels (positive, negative, neutral) by native speakers. The dataset is designed to support research in multimodal machine learning, cross-modal fusion, and low-resource dialectal NLP.

Visit

doi.orgdata.mendeley.com

Tasks

sentiment analysisspeech processingtext classification

Languages

Arabic, Algerian SpokenArabic, Moroccan Spoken

Tags

Computer ScienceArtificial IntelligenceComputer VisionNatural Language ProcessingSpeech AnalysisArabic LanguageMultimodality StudiesSentiment AnalysisLarge Language Model

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode