Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Ehugbo TTS: biblical text to speech dataset in Ehugbo Language

Domaine:

natural language processing

Type de record:

dataset

This dataset contains audio recordings of Bible verses in Ehugbo, a dialect of Igbo (a Niger-Congo language spoken in Nigeria). It contains 312 audio recordings of biblical text-to-speech data comprising 1 hour and 30 seconds of speech data.

This dataset contains verse-by-verse Bible translations, where each recording corresponds to a specific Bible verse. The recordings are segmented at the verse level, making this dataset particularly valuable for fine-grained linguistic analysis, verse-level alignment studies, and Bible translation research. The dataset covers six books from the New Testament: Acts of Apostles, Ephesians, Galatians, John, Revelations, and Romans.

Visit

mozilladatacollective.com

Tasks

speech processingautomatic speech recognitiontext to speech

Languages

Igbo

Tags

bible versesttsdialect of igbonaijavoices micro grantsmdcmozilla data collectiveehugbo

Licenses

CC-BY-NC-SA-4.0

Similaires

Kachiengineers/Ehugbo-Audioosinkolu/Ehugbo-Crosslingual-RetrievalYoruba TTS (text-to-speech) training datasetDesign of a Yoruba Language Speech Corpus for the Purposes of Text-to-Speech (TTS) SynthesisKinyarwanda Agricultural Text-to-Speech DatasetAlgerian Darija Speech-to-Text Dataset

Kachiengineers/Ehugbo-Audio

This is a repository of my Ehugbo project working on the first publicly available Ehugbo audio data

osinkolu/Ehugbo-Crosslingual-Retrieval

# Ehugbo-QA: A Multimodal Benchmark for Cross-Dialect Information Retrieval (CDIR) **Ehugbo-QA** is

Yoruba TTS (text-to-speech) training dataset

Textbook audio archive size: total 36M archive created: 8 July 2011 mp3 file size ======== ==== 01-

Design of a Yoruba Language Speech Corpus for the Purposes of Text-to-Speech (TTS) Synthesis

Kinyarwanda Agricultural Text-to-Speech Dataset

Kinyarwanda Agricultural Text-to-Speech Dataset

Algerian Darija Speech-to-Text Dataset

A speech recognition dataset of spoken Algerian Darija (Algerian Arabic dialect), containing approxi