Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

yonas-g/Afaan-Oromo-Speech-to-Text-Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
yon
Hôte:
Afaan Oromo Speech to Text Dataset # Afaan Oromo Speech to Text Dataset This repository constains preprocessed audion mfcc and transcripts. The dataset is separated into train, dev and test sets. The Dataset statistics | | | | ------------------- | -----------| | Train | 979 | | Dev | 122 | | Test | 123 | Folder structure ``` \train: \mfcc: \transcript: \dev: \mfcc: \transcript: \test: \mfcc: \transcript: ``` The Dataset statistics | | | | ------------------- | -----------| | Total clips: | 1,224 | | Total Words: | 17,559 | | Total characters: | 116,439 | | Total Duration: | 03:11:13 | | Min clip length: | 1 sec | | max clip length: | 59 sec | | Unique words: | 5,040 | Dataset Source: Afaan Oromoo Text-to-Speech… ``` Girma, Birhanu Shimelis; Senbatu, Dereje Hinsermu (2022), “Afaan Oromoo Text-to-Speech Dataset”, Mendeley Data, V1, doi: 10.17632/hnvkvj589y.1 ```

Visit

github.com

Tasks

automatic speech recognitionspeech processing

Languages

OromoOromo, Borana-Arsi-GujiOromo, EasternOromo, West Central