Logo Lanfrica

A large scale collection of voice data in Kinyarwanda to allow the building of inclusive language technology

Domaine:

natural language processing

Type de record:

dataset
This effort created the largest open-source voice dataset of diverse Kinyarwanda speakers for speech recognition (speech-to-text). It collected more than 2380 hours of AI voice data in Kinyarwanda (by 03/2026). You can use this resource to build AI systems understanding spoke… Notes / challenges: Fair Forward portfolio. Dataset

Languages

Similaires