Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Common Voice 24.0 — Ethiopian Languages

Domain:

natural language processing

Record type:

dataset
Creator:
had
Host:
This dataset is a verbatim archival upload of the Mozilla Common Voice 24.0 scripted speech data for two Ethiopian languages: Amharic (am) and Tigrinya (ti), sourced from the Mozilla Data Collective. Subset Language Code Clips Total Hours Validated Hours Speakers amharic Amharic am 1,632 2.85h 1.82h 46 tigrinya Tigrinya ti 451 0.65h 0.10h 16 Splits

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

AmharicTigrigna

Licenses

cc0-1.0

Similar

Common Voice 24.0 — Ethiopian Languages (v2)

Common Voice 24.0 — Ethiopian Languages (v2)

This is a cleaned and restructured version of hadamard-2/common-voice-24-ethiopian, which is a verba