TWB Voice 1.0 is a multilingual speech corpus containing read speech data in three languages from Nigeria: Hausa, Shuwa Arabic, and Kanuri. This dataset was created as part of the TWB Voice project by CLEAR Global (formerly Translators without Borders) to support automatic speech recognition (ASR) development for underrepresented languages.
Languages