Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Automatic Speech Recognition For African Languages Model Hub - a Hugging Face Space by asr-africa

Domaine:

natural language processing

Type de record:

project

Previous studies have led to the collection of a considerable number of hours of open-source ASR data, for example, the work done in India where over 1000’s of hours of data were collected for low-resource Indian languages. In this research, we would like to transfer the learnings from these successes and replicate the same model for low-resource African languages. For example, the aspects around the use of speech data from noisy and non-noisy environments. However, to ensure that we proceed in a cost-efficient and sustainable approach, we deem it necessary to understand the amount of data that we need to collect for African languages. Hence, we propose to leverage the Mozilla Common Voice (MCV) platform and other appropriate and openly available / open-source repositories of African language datasets to build automatic speech recognition models and test their performance to learn if the data collected was sufficient. Platform overview A preview of what the platform contains and how to navigate. Use the links and tabs in the top navigation to jump to demos, datasets, results, or evaluation details. 1. Benchmark Datasets: A multilingual collection covering over 17 African languages, built from open corpora (e.g., Common Voice, Fleurs, NCHLT, ALFFA, Naija Voices). Each dataset is cleaned, validated, and partitioned into training, development, and test splits to ensure fair benchmarking.

  1. Model Collections: Fine-tuned ASR models derived from Wav2Vec2 XLS-R, Whisper, MMS, and W2V-BERT, adapted for African phonetic, tonal, and orthographic features. These are hosted as public collections on Hugging Face.

  2. Evaluation Scenarios: Designed to test data efficiency, domain adaptation, and speech-type robustness — e.g., how models generalize from read speech to spontaneous dialogue, or from education to agricultural domains.

  3. ASR Demo Interface: A Gradio-powered live testing tool, allowing users to upload or record audio, view transcriptions, and submit structured feedback via the integrated backend API.

  4. Quantitative Results: Comprehensive analysis of model performance across training hours and data scales (1–400 hours), visualized through Word Error Rate (WER) and Character Error Rate (CER) trends. Findings show clear data scaling laws, with XLS-R and W2V-BERT models performing best under low-resource conditions.

  5. Human Evaluation Framework: A structured qualitative evaluation conducted with 20 native-language evaluators across 12 languages. Evaluators assessed accuracy, meaning preservation, orthography, and error types (e.g., named entities, punctuation, diacritics). This data is publicly available in the curated ASR_Evaluation_dataset.

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Tags

huggingfaceasrbenchmarkmodel hub

Similaires

Speech Resource Finder - a Hugging Face Space by CLEAR-GlobalMuphulusiDzivhani/Automatic-Speech-Recognition-ASR-and-Topic-Modeling-for-African-LanguagesAutomatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature ReviewBenchmarking Automatic Speech Recognition Models for African LanguagesMultilingual Automatic Speech Recognition for Kinyarwanda, Swahili, and Luganda: Advancing ASR in Select East African LanguagesTowards speech technology for South African languages: automatic speech recognition in Xhosa

Speech Resource Finder - a Hugging Face Space by CLEAR-Global

Search for speech resources, such as ASR and TTS, for different languages. Enter a language name or code to get details on supported services and models.

Almost 4 billion people speak languages with little or no speech technology support. This tool makes

MuphulusiDzivhani/Automatic-Speech-Recognition-ASR-and-Topic-Modeling-for-African-Languages

COS802 Project – Automatic Speech Recognition (ASR) and Topic Modeling for African Languages # 📊 CO

Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature Review

ASR has achieved remarkable global progress, yet African low-resource languages remain rigorously un

Benchmarking Automatic Speech Recognition Models for African Languages

Automatic speech recognition (ASR) for African languages remains constrained by limited labeled data

Multilingual Automatic Speech Recognition for Kinyarwanda, Swahili, and Luganda: Advancing ASR in Select East African Languages

Multilingual Automatic Speech Recognition for Kinyarwanda, Swahili, and Luganda: Advancing ASR in Select East African Languages

Poster presented at the Deep Learning Indaba 2023 by Samuel Rutunda

Towards speech technology for South African languages: automatic speech recognition in Xhosa