This dataset contains sentiment-labeled text data in Pedi for binary sentiment classification (Positive/Negative). Sentiments are extracted and processed from the English meanings of the sentences using DistilBERT for sentiment classification. The dataset is part of a larger collection of African language sentiment analysis resources.
Total samples: 422,975
Positive sentiment: 255703 (60.5%)