Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Galsenaicommunity/Wolof-Common-Voice

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Gal
Hôte:
Wolof Text Data Collection for recording on the Mozilla Common Voice platform. # Wolof-Common-Voice Wolof Text Data Collection for recording on the Mozilla Common Voice platform. The data comes from the Masakhane's corpus collected as part of the MasakhaNER project. # Structure The project is structured as follows: ``` . ├── data/ │   └── raw Wolof is now referenced on Common Voice and you can enter your email on the platform to follow the progress of the project 🥳 # Replicate this project for your language If you wish to have your language referenced on Common Voice, just go to the platform and click on `LANGUAGES` then `Request a Language`. Fill in the information and wait for the Mozilla team to contact you. You will then have to open a Github Issue for language localization and fill a template. When you get to this stage, you can use our template as inspiration to fill yours. You will then have to collect textual data in your language and upload them to the sentence collector taking into account the prerequisites advised by Mozilla (cf Upload section). For optimal recording conditions, we advise you to also follow the indications provided in this document. > If you need any feedback, you can send us an email at galsenaimeetups[at]gmail[dot]com # License This work is licensed under a Creative Commons Attribution-ShareAlike 4.0 International License .

Visit

github.com

Languages

Wolof

Licenses

CC-BY-SA-4.0

Similaires

Galsenaicommunity/Wolof-NMTGalsenaicommunity/Wolof-ASRCommon VoiceCommon Voice BasaaNigerian Common Voice DatasetCommon Voice Baoulé (bci)

Galsenaicommunity/Wolof-NMT

# French-Wolof Translator A modular, production-ready French-Wolof translation system built on Face

Galsenaicommunity/Wolof-ASR

State of the art ASR models for the Wolof Language # Wolof-ASR State of the art ASR models for the

Common Voice

Common Voice is Mozilla's initiative to help teach machines how real people speak. The dataset currently consists of 7,335 validated hours of speech in 60 languages, but we’re always adding more voices and languages.

Common Voice Basaa

Voice data collection and distribution Interface for Basaa language using Mozilla's Common Voice infrastructure.

Nigerian Common Voice Dataset

The Nigerian Common Voice Dataset is a comprehensive dataset consisting of 158 hours of audio record

Common Voice Baoulé (bci)

Dataset audio en langue baoulé (code ISO 639-3 : bci), extrait de Mozilla Common Voice. Le baoulé es