Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Nigerian Common Voice Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Ala
Host:
The Nigerian Common Voice Dataset is a comprehensive dataset consisting of 158 hours of audio recordings and corresponding transcription (sentence). This dataset includes metadata like accent, locale that can help improve the accuracy of speech recognition engines. This dataset is specifically curated to address the gap in speech and language

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

HausaIgboYoruba

Licenses

apache-2.0