International audience Speech-to-Text (STT) systems, despite their stellar performanc
This paper presents Yankari, a large-scale monolingual dataset for the Yoruba language, aimed at add
Monolingual corpus for South African English. The data is given as a single UTF-8 text file, with ea
# 🗣️ Whisper ASR Fine-Tuning for Amharic # 🗣️ Whisper ASR Fine-Tuning for Amharic This project imp
Monolingual corpus for isiXhosa. The data is given as a single UTF-8 text file, with each segment on