Logo Lanfrica

DhoNam: Dholuo Speech dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Mas
Hôte:
DhoNam: Dholuo Speech dataset is a speech corpus designed to supercharge Automatic Speech Recognition (ASR) and other speech technologies for Dholuo, one of Kenya’s major indigenous languages. This dataset contains native-speaker audio recordings collected through a platform where users read aloud a displayed sentence. The dataset includes the audio recordings and the corresponding prompt/sentence that was read.