This dataset contains child speech samples from typically developing Afrikaans and Sesotho sa Leboa-speaking children in South Africa. The recordings, totaling at least 700 minutes per language, were collected in naturalistic interactions between children and trained speech-language therapists using standardized toys and books. Data were gathered in home and clinical settings to support linguistic analysis, speech transcription, and the development of automated tools for early identification and intervention in multilingual contexts.
For more details, review the readme.txt for Afrikaans and Sesotho sa Leboa