Logo Lanfrica
fr
Accueil
Atlas
Analyses
Documentation
Sign in
Retour
Abdullah804/yoruba-subset
Domaine:
natural language processing
Type de record:
dataset
Créateur:
Abd
Hôte:
Visit
Actions
Share
Report an issue
This is a 100-hour subset of the Naija Voices Yoruba speech dataset, curated by @Abdullah804. 101,391 samples Converted to 16kHz WAV format Includes metadata: path, text, speaker_id, gender, age_range, duration Prepared for ASR/TTS finetuning.
Visit
huggingface.co
Tasks
automatic speech recognition
speech processing
text to speech
Languages
Yoruba