Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Afaan Oromoo Text-to-Speech Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Bir
Éditeur:
BirDer
Éditeur:
Men
Hôte:avatar
Afaan Oromo is one of the languages that have huge speakers in the horn of Africa. It is also one of the under-resourced languages like other Ethiopian languages. In this Dataset preparation, the soul purpose of the project was to include Afaan Oromo text-to-speech synthesis in our Final year Humanoid robot that can speak the Oromo language in addition to its vision capability to detect emotion, gender, and detect faces of humans. Currently, the natural language processing applications that use this language are in high demand. Furthermore, the linguists and researchers that work on these languages are contributing a lot of data to the growth of this language to make it an international and machine language. This dataset has been started by two Electrical Engineering students at Madda Walabu University, during their 4 months industry internship program at iCog-Labs, while working on machine learning tasks related to Natural language processing. The first phase of the audio was recorded at Addis Ababa University by the two students and the preprocessing was also done by them. The second phase of the audio recording was done at Madda Walabu university after their internship was completed. They have selected female students from the Electrical engineering and Afaan Oromo department to record the audio to get more corpus to train the machine learning model. The model used was Tacotron 1 and 2 with waveGlow and TensorFlow framework that was developed by NVIDIA company. The Corpus statistics Total clips: 1,224 Total Words: 17,559 Total characters: 116,439 Total Duration: 03:11:13 Min clip length: 1 sec max clip length: 59 sec Unique words: 5,040 Credits The volunteers that participated in the audio recording need to be appreciated and they were concerned about their language. They are Obsinet Asmare Motuma, Milko Wariyo Gobana, Roza Hailu Isho, Demitu Baye Boyosa.

Visit

doi.orgdata.mendeley.com

Tasks

speech processingtext to speech

Languages

AmharicMadaOromoOromo, Borana-Arsi-GujiOromo, EasternOromo, West Central

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Afaan Oromo Text to Speech Synthesis datasetAfaan Oromo Text to Speech Synthesis datasetyonas-g/Afaan-Oromo-Speech-to-Text-DatasetText To Speech Synthesis for Afaan Oromoo Language Using Deep Learning ApproachAfaan Oromoo-TTS-DatasetAfaan Oromoo Wikipedia Dataset

Afaan Oromo Text to Speech Synthesis dataset

Afaan Oromo Text to Speech Synthesis dataset is a public domain speech dataset consisting of 8,076 s

Afaan Oromo Text to Speech Synthesis dataset

yonas-g/Afaan-Oromo-Speech-to-Text-Dataset

Afaan Oromo Speech to Text Dataset # Afaan Oromo Speech to Text Dataset This repository constains

Text To Speech Synthesis for Afaan Oromoo Language Using Deep Learning Approach

Afaan Oromoo-TTS-Dataset

This dataset comprises 1,737 high-quality audio recordings of read speech produced by a single Afaan

Afaan Oromoo Wikipedia Dataset

Train a Language Model with Afaan Oromoo Articles