Wolof Text Data Collection for recording on the Mozilla Common Voice platform.
# Wolof-Common-Voice
Wolof Text Data Collection for recording on the Mozilla Common Voice platform.
The data comes from the Masakhane's corpus collected as part of the MasakhaNER project.
# Structure
The project is structured as follows:
```
.
├── data/
│ └── raw Wolof is now referenced on Common Voice and you can enter your email on the platform to follow the progress of the project 🥳
# Replicate this project for your language
If you wish to have your language referenced on Common Voice, just go to the platform and click on `LANGUAGES` then `Request a Language`. Fill in the information and wait for the Mozilla team to contact you. You will then have to open a Github Issue for language localization and fill a template. When you get to this stage, you can use our template as inspiration to fill yours.
You will then have to collect textual data in your language and upload them to the sentence collector taking into account the prerequisites advised by Mozilla (cf Upload section).
For optimal recording conditions, we advise you to also follow the indications provided in this document.
> If you need any feedback, you can send us an email at galsenaimeetups[at]gmail[dot]com
# License
This work is licensed under a Creative Commons Attribution-ShareAlike 4.0 International License .