This repository hosts a collection of Wolof speech datasets. The datasets, sourced from various contributors, are each stored in separate pickle files. Each pickle file contains the following columns:
audio: The audio data in WAV or MP3 format.
transcription: The corresponding transcriptions of the audio data.
length(duration(s)): The duration of each audio recording.
filename: The name of the audio file.
path: The path to the audio file.