The Swahili Stop-Words Dataset is a curated collection of function words that carry minimal semantic weight and are commonly omitted during text preprocessing in Natural Language Processing (NLP) workflows.
While these words are essential for the syntactic structure of Swahili, they can be excluded from most computational tasks without compromising the overall semantic integrity of the text.