Welcome to the NaijaVoices dataset. The NaijaVoices dataset consists of 1,800 hours of authentic speech (from over 5,000 diverse speakers!) and expert curated text in Igbo, Hausa, and Yoruba. ~600 hours for each of the three languages. It also boasts of adequate female representation and balanced age-range distribution (young to old speakers). For more about the dataset info visit our website: naijavoices.com. By using this dataset, you acknowledge reading and accepting to use this dataset according to the data usage terms and conditions: naijavoices.com
Arvix: The NaijaVoices Dataset: Cu…