darija <-> english dataset # Darija Open Dataset Welcome to the Darija Open Dataset (DODa), a
Algerian Forest Fires Data from 2012
This repository accompanies the project "A Machine Learning Model for Predicting Finger Millet Grain
A Text Recognition (TR) and Optical Character Recognition (OCR) dataset for the Tigrinya language. T
This dataset contains 125 hours of transcribed speech recordings about agriculture in Wolof, Pulaar and Sereer, the three most widely spoken languages in Senegal.
The attached file is the first Mauritanian dialect dataset called “HASSANIYA” containing two thousan