Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Common Voice Corpus 15.0

Domaine:

natural language processing

Type de record:

dataset
Créateur:
fsi
Hôte:
This dataset is an unofficial version of the Mozilla Common Voice Corpus 15. It was downloaded and converted from the project's website commonvoice.mozilla.org. Languages

Visit

huggingface.co

Languages

AfrikaansAmazighAmharicBasaaGandaHausaIgboJulaKinyarwandaSwahili+4

Tags

mozillafoundation

Licenses

cc

Similaires

Common Voice Corpus 16.0Common Voice Corpus 16Common Voice Corpus 17.0Common Voice: A Massively-Multilingual Speech CorpusCommon VoiceCommon Voice Basaa

Common Voice Corpus 16.0

This dataset is an unofficial version of the Mozilla Common Voice Corpus 16. It was downloaded and c

Common Voice Corpus 16

The Common Voice dataset consists of a unique MP3 and corresponding text file. Many of the 30328 rec

Common Voice Corpus 17.0

This dataset is an unofficial version of the Mozilla Common Voice Corpus 17. It was downloaded and c

Common Voice: A Massively-Multilingual Speech Corpus

The Common Voice corpus is a massively-multilingual collection of transcribed speech intended for speech technology research and development. Common Voice is designed for Automatic Speech Recognition purposes but can be useful in other domains (e.g. language identi

Common Voice

Common Voice is Mozilla's initiative to help teach machines how real people speak. The dataset currently consists of 7,335 validated hours of speech in 60 languages, but we’re always adding more voices and languages.

Common Voice Basaa

Voice data collection and distribution Interface for Basaa language using Mozilla's Common Voice infrastructure.