This effort created the largest open-source voice dataset of diverse Kinyarwanda speakers for speech recognition (speech-to-text). It collected more than 2380 hours of AI voice data in Kinyarwanda (by 03/2026). You can use this resource to build AI systems understanding spoke…
Notes / challenges: Fair Forward portfolio. Dataset