This dataset is designed to aid in the development of tools for detecting and correcting non-vocabulary errors in Yoruba text. It comprises a collection of Yoruba sentences and paragraphs, each annotated to highlight common errors such as grammatical mistakes, tonal errors, spelling inconsistencies, and improper word usage.