Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Afaan Oromo Text to Speech Synthesis dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
sem
Éditeur:
AdaMet
Éditeur:
Men
Hôte:avatar
Afaan Oromo Text to Speech Synthesis dataset is a public domain speech dataset consisting of 8,076 short audio clips of a single male speaker reading sentences collected from legitimate sources such as News Media sources, Non-fiction books, and Afaan Oromo Holy bible. A transcription and its normalized text are provided for each clip. After two weeks of the audio recording process, a total of 17 hours of recorded speech data that corresponded to a total of 8076 recorded .wav files was created. File Format Metadata is provided in metadata.csv. This file consists of one record per line, delimited by the pipe character. The fields are: ID: this is the name of the corresponding .wav file Transcription: words spoken by the reader (UTF-8) Normalized Transcription: transcription with numbers, ordinals, and monetary units expanded into full words (UTF-8). Each audio file is a single-channel 16-bit PCM WAV with a sample rate of 22050 Hz.

Visit

doi.orgdata.mendeley.com

Tasks

speech processingtext to speech

Languages

OromoOromo, Borana-Arsi-Guji

Tags

Natural Language ProcessingText-to-SpeechSpeech Synthesis

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode