Logo Lanfrica

Child speech database for the South African context: Speech samples of typically developing Afrikaans and Sesotho sa Leboa-speaking children

Domain:

natural language processinghealthcare

Record type:

dataset
Creator:
BorDe Van
Editor:
Win
Publisher:
University of PretoriaSAD
Host:avatar
This dataset contains child speech samples from typically developing Afrikaans and Sesotho sa Leboa-speaking children in South Africa. The recordings, totaling at least 700 minutes per language, were collected in naturalistic interactions between children and trained speech-language therapists using standardized toys and books. Data were gathered in home and clinical settings to support linguistic analysis, speech transcription, and the development of automated tools for early identification and intervention in multilingual contexts. For more details, review the readme.txt for Afrikaans and Sesotho sa Leboa