The Swahili Verb Conjugation Dataset is an extensive resource containing over 319,156 meticulously c
This research developed a Kencorpus Swahili Question Answering Dataset KenSwQuAD from raw data of Swahili language, which is a low resource language predominantly spoken in Eastern African and also has speakers in other parts of the world. Question Answering datase
Labelled Swahili text samples for civic, health, financial, and educational NLP tasks in East Africa
Today African languages are spoken by more than a billion people, yet in the world of machine transl
This dataset, the Sheng-English