Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

ALFFA Amharic Speech Corpus

Domaine:

natural language processing

Type de record:

dataset
Créateur:
had
Hôte:
Read speech corpus for Amharic (አማርኛ) automatic speech recognition, converted to HuggingFace Datasets format from the original ALFFA project. { 'audio': Audio(sampling_rate=16000), 'utterance_id': 'tr_10000_tr097082', 'transcript': 'ይህ አማርኛ ጽሑፍ ነው', 'speaker_id': '097', 'split': 'train' } from datasets import load_dataset dataset = load_dataset("hadamard-2/alffa-amharic")

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

Amharic

Tags

audioautomatic-speech-recognition

Licenses

mit

Similaires

ALFFA Amharic Speech Corpus (v2)Amharic speech corpusAmharic Speech CorpusALFFA Speech DatasetwubeZ/Custom-Amharic-Speech-CorpusFongbe Speech Dataset (ALFFA + Zenodo)

ALFFA Amharic Speech Corpus (v2)

Read speech corpus for Amharic (አማርኛ) automatic speech recognition. Converted from the original ALFF

Amharic speech corpus

Amharic Speech Corpus

This is an Amharic speech corpus which is suitable for the development and evaluation of speech recognition and retrieval systems. The corpus contains 110 hours of speech data with syllable and grapheme-based transcriptions collected from public domain or resources

ALFFA Speech Dataset

The ALFFA dataset includes both audio files and original text transcriptions for Swahili, utilized f

wubeZ/Custom-Amharic-Speech-Corpus

an Amharic tts custom Dataset # Amharic TTS Custom Dataset ## Overview This repository contains a

Fongbe Speech Dataset (ALFFA + Zenodo)

This dataset is a unified, high-quality collection of Fongbe speech data, specifically curated to pr