Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

YodaSpeech

Domain:

natural language processing

Record type:

dataset
Creator:
Tho
Host:
YodaLingua is a high-quality speech dataset designed for training text-to-speech (TTS) systems, ASR models, and any application requiring clean, well-aligned audio–text pairs.This release contains the Swahili portion of the multilingual YodaLingua collection. Property Value Total clips 5,144 audio–transcription pairs Total duration 15 hours Speakers 443 distinct speakers Audio format

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

Swahili

Tags

TTSswahili

Licenses

cc-by-4.0