Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Duala-TTS-Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Ins
Host:
This dataset comprises 1,521 high-quality audio recordings of read speech produced by a single Duala speaker over several sessions. Duala (ISO 639-3: dua), also known as Douala, is a Bantu language of the Niger-Congo family spoken primarily in the Littoral Region of Cameroon, notably in the city of Douala and its surrounding areas. It is a low-resource language with limited existing digital speech resources, making this dataset a significant contribution to natural language processing efforts for the language. Audio files are provided in MP3 format (approx. 147 MB), totalling 4 hours, 30 minutes and 41.64 seconds of speech. The dataset includes 16 audio/sentence mapping files in TSV format, containing 1,521 aligned audio/sentence pairs in total. Transcriptions follow the General Alphabet of Cameroonian Languages, a standardised orthographic system based on the Latin alphabet augmented with phonetic characters and diacritical marks used to represent tonal and phonological features of Cameroonian languages. The recordings draw on narrative texts relating to colonial encounters and experiences. These narratives originally existed as oral and audio recordings and were subsequently transcribed. The read-speech recordings therefore reflect a rich oral tradition rendered in text, offering valuable prosodic and lexical diversity for training and evaluating TTS and ASR models. The dataset is intended for research and scientific use in speech technology for Duala.

Visit

mozilladatacollective.com

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

Duala

Tags

mdcmozilla data collectiveTTSMP3TSV

Licenses

Nwulite Obodo Open Data Licence 1.0 (NOODL-1.0)

Similar

Laari-TTS-DatasetEwondo-TTS-DatasetTobydata Tts DatasetMbosi-TTS-DatasetKiswahili TTS DatasetRw Tts Dataset

Laari-TTS-Dataset

The dataset contains audio and text resources on Laari, a Bantu language spoken in the Congo. The re

Ewondo-TTS-Dataset

This dataset comprises high-quality audio recordings of read speech from a single female speaker of

Tobydata Tts Dataset

Luganda TTS dataset (Toby-data) collected by TericLab. Contains read speech in Luganda, primarily on

Mbosi-TTS-Dataset

The dataset consists of paired audio and text data on Mbosi (mdw), a language spoken in Congo. The a

Kiswahili TTS Dataset

The dataset contains Kiswahili text and audio files. The dataset contains 7,108 text files and audio files. The Kiswahili dataset was created from an open-source non-copyrighted material: Kiswahili audio Bible. The authors permit use for non-profit, educational, a

Rw Tts Dataset

Kinyarwanda (rw) text-to-speech dataset. Studio-recorded read speech aligned with transcriptions, co