Logo Lanfrica

HASSANIYA Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
El
Éditeur:
UniUni
Éditeur:
Men
Hôte:avatar
The attached file is the first Mauritanian dialect dataset called “HASSANIYA” containing two thousand records classified into three categories: positive, negative and neutral. This dataset was collected using web scraping tools from comments posted on the Facebook platform, and Label Studio was used to annotate each record.