A Setswana (ISO 639-3: tsn) Twitter sentiment dataset of 3,555 tweets annotated by three native-speaker annotators. The dataset is released alongside full per-annotation timestamps, language-identification metadata, and per-annotator labels to enable both downstream modelling and annotation-quality research.
Three classification splits plus a full config are provided.
Split
Examples
Use
train
2,762