Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Common Voice Corpus 15.0

Domain:

natural language processing

Record type:

dataset
Creator:
fsi
Host:
This dataset is an unofficial version of the Mozilla Common Voice Corpus 15. It was downloaded and converted from the project's website commonvoice.mozilla.org. Languages

Visit

huggingface.co

Languages

AfrikaansAmazighAmharicBasaaGandaHausaIgboJulaKinyarwandaSwahili+4

Tags

mozillafoundation

Licenses

cc

Similar

Common Voice Corpus 17.0Common Voice Corpus 16.0Common Voice Corpus 16Common Voice: A Massively-Multilingual Speech CorpusCommon VoiceCommon Voice Basaa

Common Voice Corpus 17.0

This dataset is an unofficial version of the Mozilla Common Voice Corpus 17. It was downloaded and c

Common Voice Corpus 16.0

This dataset is an unofficial version of the Mozilla Common Voice Corpus 16. It was downloaded and c

Common Voice Corpus 16

The Common Voice dataset consists of a unique MP3 and corresponding text file. Many of the 30328 rec

Common Voice: A Massively-Multilingual Speech Corpus

The Common Voice corpus is a massively-multilingual collection of transcribed speech intended for speech technology research and development. Common Voice is designed for Automatic Speech Recognition purposes but can be useful in other domains (e.g. language identi

Common Voice

Common Voice is Mozilla's initiative to help teach machines how real people speak. The dataset currently consists of 7,335 validated hours of speech in 60 languages, but we’re always adding more voices and languages.

Common Voice Basaa

Voice data collection and distribution Interface for Basaa language using Mozilla's Common Voice infrastructure.