Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Laari-TTS-Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Ins
Host:
The dataset contains audio and text resources on Laari, a Bantu language spoken in the Congo. The resources, which are suitable for TTS tasks and possibly ASR tasks, consist of the following: - 6,311 audio clips totalling 241 minutes and 44.97 seconds; - an audio mapping file with 5,321 lines, each beginning with the name of an audio file, followed by a tab and then the corresponding text excerpt; - two raw audio files totalling 120 minutes and 54.90 seconds; - two long audio files with their original, non-split transcription files, for a total duration of 120 minutes and 41.90 seconds.

Visit

mozilladatacollective.com

Tasks

speech processingtext to speech

Languages

Laari

Tags

mdcmozilla data collectiveASRWAVTRJSTSV

Licenses

Nwulite Obodo Open Data Licence 1.0 (NOODL-1.0)