Somali TTS Corpus is a cleaned Somali speech dataset designed for Text-to-Speech (TTS) and speech synthesis research. The dataset contains high-quality Somali audio recordings paired with text, processed with noise reduction and audio normalization techniques to improve training quality for speech synthesis models.
Feature
Type
text
string
audio
Audio
Processing