Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

ALFFA Amharic Speech Corpus (v2)

Domain:

natural language processing

Record type:

dataset
Creator:
had
Host:
Read speech corpus for Amharic (አማርኛ) automatic speech recognition. Converted from the original ALFFA project and restructured to match the google/waxalnlp schema for interoperability. This is a restructured version of hadamard-2/alffa-amharic. utterance_id renamed to id transcript renamed to transcription

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

Amharic

Tags

audioautomatic-speech-recognition

Licenses

mit

Similar

ALFFA Amharic Speech CorpusAmharic Speech CorpusAmharic speech corpusALFFA Speech DatasetwubeZ/Custom-Amharic-Speech-CorpusFongbe Speech Dataset (ALFFA + Zenodo)

ALFFA Amharic Speech Corpus

Read speech corpus for Amharic (አማርኛ) automatic speech recognition, converted to HuggingFace Dataset

Amharic Speech Corpus

This is an Amharic speech corpus which is suitable for the development and evaluation of speech recognition and retrieval systems. The corpus contains 110 hours of speech data with syllable and grapheme-based transcriptions collected from public domain or resources

Amharic speech corpus

ALFFA Speech Dataset

The ALFFA dataset includes both audio files and original text transcriptions for Swahili, utilized f

wubeZ/Custom-Amharic-Speech-Corpus

an Amharic tts custom Dataset # Amharic TTS Custom Dataset ## Overview This repository contains a

Fongbe Speech Dataset (ALFFA + Zenodo)

This dataset is a unified, high-quality collection of Fongbe speech data, specifically curated to pr