
This corpus contains high-quality BBC Igbo and BBC Pidgin news snippets annotated by Bytte AI for intent, sentiment, content quality, sentence segmentation, and Named Entity Recognition (NER). Comprising 63 Igbo and 91 Pidgin samples per task, it provides a rich, multilingual benchmark for NLP research. Designed for low-resource African languages, it enables model training, evaluation, and benchmarking on real-world journalistic text.