Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Amharic Speech Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Sil
Host:
The most comprehensive Amharic speech dataset on HuggingFace - natural, real-world Amharic from native speakers in Ethiopia and the diaspora. Total audio samples: 51 recordings Total duration: ~23 minutes Primary region: Ethiopia (Addis Ababa) Context: Natural spontaneous speech (free_speech) Audio format: WAV files Sample rate: 48 kHz License: CC BY-NC 4.0 (free for research, non-commercial use)

Visit

huggingface.co

Languages

Amharic

Tags

amharicethiopian-languageseast-africaethiopiageez-scriptsemitic-languagesafrican-languageslow-resourcespeech-datavoice-ai+2

Licenses

cc-by-nc-4.0