Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

ghananlpcommunity/ghana-english-corrected-transcriptions

Domain:

natural language processing

Record type:

dataset
Creator:
gha
Host:
A semi-synthetic dataset of original and corrected transcriptions from Ghanaian news media, designed for training automatic speech recognition (ASR) and text-to-speech (TTS) correction models. Dataset Description

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processingtext to speech

Similar

ghananlpcommunity/ghana-qaghananlpcommunity/ghana-chatghananlpcommunity/omnivoice-ghanaghananlpcommunity/english-twi_sentence-pairs-4mghananlpcommunity/twi-english-paragraph-dataset_newsghananlpcommunity/ghana-asr-audio

ghananlpcommunity/ghana-qa

The Ghana-QA dataset consists of question-answer pairs derived from Ghanaian news articles. It conta

ghananlpcommunity/ghana-chat

Ghana-Chat is a large-scale, high-quality synthetic conversational dataset designed specifically to

ghananlpcommunity/omnivoice-ghana

ghananlpcommunity/english-twi_sentence-pairs-4m

This dataset is made available because of Ghana NLP's volunteer driven research work. Please conside

ghananlpcommunity/twi-english-paragraph-dataset_news

This is a parallel Twi-English dataset designed for training machine translation models that underst

ghananlpcommunity/ghana-asr-audio