Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

carlshizi/yoruba-speech-pipelines

Domain:

natural language processing

Record type:

paper
Creator:
car
Host:
A research writing sample proposing a data-centric speech pipeline for low-resource Yoruba ASR, TTS, and tone-aware evaluation. ## Building Low-Resource Speech Pipelines for Yoruba Using Data-Centric Methods This repository contains a methodological research paper that outlines a data-centric speech processing pipeline for Yoruba, a low-resource tonal language. The paper describes system design, data collection methodology, and an evaluation framework for Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and tone-aware assessment. Empirical results are planned as future work, and this document is intended as a research writing sample reflecting my research direction.

Visit

github.com

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

Yoruba

Similar

Yoruba Speech Datasetyoruba speech datasetshunyalabs/yoruba-speech-datasetReported speech in Yorubatemi/yoruba-speech-translationMaryAdelua/yoruba-speech-tools

Yoruba Speech Dataset

The most comprehensive Yoruba speech dataset on HuggingFace - natural, real-world Yoruba from native

yoruba speech dataset

Yoruba Language Audio Dataset (12 Speakers, 6 Hours of Recordings)

shunyalabs/yoruba-speech-dataset

Reported speech in Yoruba

temi/yoruba-speech-translation

MaryAdelua/yoruba-speech-tools

# Yoruba pronunciation resource and integration research This repository develops a training-ready