The "all-categories" version of za-isizulu-siswati-news for isizulu news.
The paper uses 5-fold cross validation to train models, so here all data is put in the train split.
The original dataset was not loadable at the time by calling the following:
from datasets import load_dataset
dataset = load_dataset("dsfsi/za-isizulu-siswati-news")