Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

za-african-next-voices

Domain:

natural language processing

Record type:

dataset
Creator:
dsf
Host:
Note: This dataset is a compressed version of za-african-next-voices. It was compressed to .opus format using a 32k bitrate. Swivuriso: ZA-African Next Voices-Compressed

Visit

huggingface.co

Languages

NdebeleSetswanaSotho, SouthernTsongaVendaXhosaZulu

Licenses

cc-by-4.0

Similar

za-african-next-voicesSwivuriso: ZA-African Next Voiceskesbeast23/za-african-next-voices-tonalkesbeast23/za-african-next-voices-difficulty-scoresMali African Next VoicesRwanda African Next Voices

za-african-next-voices

Swivuriso is a large-scale multilingual speech dataset targeting over 3000 hours of audio across 7 S

Swivuriso: ZA-African Next Voices

Swivuriso is a 3000-hour multilingual speech dataset developed as part of the African Next Voices project, to support the development and benchmarking of automatic speech recognition (ASR) technologies in seven South African languages. Covering

kesbeast23/za-african-next-voices-tonal

This dataset contains tonal (F0/pitch) metadata extracted from dsfsi-anv/za-african-next-voices. zu

kesbeast23/za-african-next-voices-difficulty-scores

Mali African Next Voices

The AfVoices dataset is the largest open corpus of spontaneous Bambara speech at its release in late 2025. It contains 423 hours of segmented audio and 612 hours of original raw recordings collected across southern Mali. Speech was recorded in natural, conversation

Rwanda African Next Voices

The dataset was created by Digital Umuganda and made possible through funding from the Gates Foundation. The data spans five high-impact domains — Health, Government, Financial Services, Education, and Agriculture — to support robust ASR model development in both c