This dataset contains the 150-sample Shona corpus stored in the Modal sna-data volume.
id: integer row id from the source metadata
filename: audio file name
speaker_id: speaker label
text: transcription
audio: WAV audio loaded as a Hugging Face audio feature
Rows: 150
WAV files: 150
Sample rates: {"48000": 150}
Channels: {"1": 150}
Duration min: 3.947 sec
Duration max: 14.229 sec