Logo Lanfrica

Efik-TTS-Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Ins
Hôte:
This dataset comprises audio recordings of Efik speech aligned with textual transcriptions. The dataset is structured into 10 folders, each containing audio files and a corresponding audio-text mapping file. The audio clips are short, typically ranging from 1 to 30 seconds, and are suitable for training and evaluating Text-to-Speech (TTS) systems. The dataset follows a structured format where each audio file is paired with its corresponding transcription in a tab-separated mapping file. The textual content used in this dataset originates from a variety of written sources in Efik, including encyclopaedic and informational texts, as well as everyday conversational phrases. These texts were segmented into short utterances suitable for read speech and TTS modelling.