The “DhoNam: Dholuo Speech dataset” is a speech corpus designed to supercharge Automatic Speech Recognition (ASR) and other speech technologies for Dholuo, one of Kenya’s major indigenous languages. This dataset contains native-speaker audio recordings collected through a plat…
Notes / challenges: Fair Forward portfolio. Dataset