Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

African Next Voices: Ethiopia

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Digital Umuganda
AfriVoice Ethiopia is an open-source speech corpus for ASR development covering five Ethiopian languages: Amharic, Afaan Oromo, Sidama, Wolaytta, and Tigrinya.

Visit

huggingface.cozenodo.orgzenodo.orgzenodo.orgzenodo.orgzenodo.org

Tasks

automatic speech recognitionspeech processing

Languages

AmharicOromoSidamoTigrignaWolaytta

Tags

African Next VoicesANVASREthiopiaDigital Umuganda

Licenses

CC BY 4.0

Similaires

za-african-next-voicesza-african-next-voicesMali African Next VoicesRwanda African Next VoicesSwivuriso: ZA-African Next Voiceskesbeast23/za-african-next-voices-tonal

za-african-next-voices

Swivuriso is a large-scale multilingual speech dataset targeting over 3000 hours of audio across 7 S

za-african-next-voices

Note: This dataset is a compressed version of za-african-next-voices. It was compressed to .opus for

Mali African Next Voices

The AfVoices dataset is the largest open corpus of spontaneous Bambara speech at its release in late 2025. It contains 423 hours of segmented audio and 612 hours of original raw recordings collected across southern Mali. Speech was recorded in natural, conversation

Rwanda African Next Voices

The dataset was created by Digital Umuganda and made possible through funding from the Gates Foundation. The data spans five high-impact domains — Health, Government, Financial Services, Education, and Agriculture — to support robust ASR model development in both c

Swivuriso: ZA-African Next Voices

Swivuriso is a 3000-hour multilingual speech dataset developed as part of the African Next Voices project, to support the development and benchmarking of automatic speech recognition (ASR) technologies in seven South African languages. Covering

kesbeast23/za-african-next-voices-tonal

This dataset contains tonal (F0/pitch) metadata extracted from dsfsi-anv/za-african-next-voices. zu