Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Bomitaba-TTS-Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Ins
Hôte:
The dataset comprises three components: audio clips, an audio mapping file, and raw audio of Bomitaba, a Bantu language spoken in the Congo. Each audio clip is paired with its corresponding transcription. There are 2,613 transcribed audio clips, totalling 182 minutes and 4 seconds. There are two raw audio files totalling 121 minutes and 14.24 seconds. The audio mapping file contains 2,610 lines. Each line begins with the name of an audio file, followed by a tab, then the corresponding text exce

Visit

mozilladatacollective.com

Tasks

text to speechspeech processing

Languages

Bomitaba

Tags

mdcmozilla data collectiveTTSWAVTSV

Licenses

Nwulite Obodo Open Data Licence 1.0 (NOODL-1.0)