mghana-st is a curated audio dataset intended for speech tasks (e.g., speech recognition, speaker identification, speech clustering, or speech translation). It contains short local language audio clips collected from various sources, annotated with English translations and tags for non-verbal events.
Dataset Description