Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Somali TTS Corpus

Domain:

natural language processing

Record type:

dataset
Creator:
maa
Host:
Somali TTS Corpus is a cleaned Somali speech dataset designed for Text-to-Speech (TTS) and speech synthesis research. The dataset contains high-quality Somali audio recordings paired with text, processed with noise reduction and audio normalization techniques to improve training quality for speech synthesis models. Feature Type text string audio Audio Processing

Visit

huggingface.co

Tasks

speech processingtext to speech

Languages

Somali

Tags

somalittstext-to-speechspeech-synthesisspeech-corpusaudiofemale-voiceclean-audionoise-reductionlow-resource-language+1

Licenses

apache-2.0

Similar

Somali-tts/somali-tts-datasetsSomali TTSSomali-tts/tts_somali_finetunedSomali-tts/somali_speecht5_outputSomali-tts/Jelle_ttsSomali-tts/diirow_tts

Somali-tts/somali-tts-datasets

Somali TTS

This dataset contains high-quality Somali speech recordings with corresponding transcriptions.It is

Somali-tts/tts_somali_finetuned

Somali-tts/somali_speecht5_output

Somali-tts/Jelle_tts

Somali-tts/diirow_tts