48775 speech-text pairs split from long recordings.
Source audio from ghananlpcommunity/ewe-tts-bible-full-audio-text
Full-file CTC forced alignment (MMS-300M) for word-level timestamps
Words grouped into 16-word segments
Leading/trailing silence trimmed with VAD (-40 dBFS threshold)
Filtered: min 1.0s, max 15.0s
Original sample rate preserved (24kHz)
from datasets import load_dataset