v2605 · Tigrinya · 14.7 hours · CC-BY-4.0
Phonetico Speech is a speech corpus for automatic speech recognition (ASR) in Ethiopian languages. Each language is available as a separate config. Load only what you need. v2605 contains 14.7 hours of transcribed Tigrinya audio.
This dataset is part of a long-term effort to build foundational speech technology for Ethiopian languages.
Language
Tigrinya (tir, ISO 639-3)