This dataset is a verbatim archival upload of the Mozilla Common Voice 24.0 scripted speech data for two Ethiopian languages: Amharic (am) and Tigrinya (ti), sourced from the Mozilla Data Collective.
Subset
Language
Code
Clips
Total Hours
Validated Hours
Speakers
amharic
Amharic
am
1,632
2.85h
1.82h
46
tigrinya
Tigrinya
ti
451
0.65h
0.10h
16
Splits