Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Common Voice Scripted Speech 25.0

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Pea
Hôte:
This repository is being prepared as a row-normalized multilingual ASR dataset built from Mozilla Data Collective Common Voice Scripted Speech 25.0. The normalized rows are the main deliverable: audio, sentence, locale, language, split, source_dataset_id, source_archive, and upstream Common Voice metadata where present. During staging, the original MDC .tar.gz archives are preserved

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

AfrikaansAjaAmharicBaatonumBafiaBafutBakokoBamenyamBamunBankon+38

Tags

common-voicemozilla-data-collectivescripted-speechmultilingual-asrasr

Licenses

cc0-1.0

Similaires

Common Voice Scripted Speech 26.0 - AmharicCommon Voice Scripted Speech 26.0 - BuluCommon Voice Scripted Speech 26.0 - MokpweCommon Voice Scripted Speech 26.0 - TunenCommon Voice Scripted Speech 26.0 - KotokoliCommon Voice Scripted Speech 26.0 - Kihemba

Common Voice Scripted Speech 26.0 - Amharic

A collection of read speech recordings in Amharic (አማርኛ).

Common Voice Scripted Speech 26.0 - Bulu

A collection of read speech recordings in Bulu (bum).

Common Voice Scripted Speech 26.0 - Mokpwe

A collection of read speech recordings in Mokpwe (bri).

Common Voice Scripted Speech 26.0 - Tunen

A collection of read speech recordings in Tunen (tvu).

Common Voice Scripted Speech 26.0 - Kotokoli

A collection of read speech recordings in Kotokoli (kdh).

Common Voice Scripted Speech 26.0 - Kihemba

A collection of read speech recordings in Kihemba (hem).