Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

michsethowusu/twi_multispeaker_audio_transcribed

Domain:

natural language processing

Record type:

dataset
Creator:
mic
Host:
The Twi Multispeaker Audio Transcribed dataset is a collection of speech recordings and their transcriptions in Asante Twi, a widely spoken dialect of the Akan language in Ghana. The dataset is designed for training and evaluating automatic speech recognition (ASR) models and other natural language processing (NLP) applications. Dataset Details

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

AkanAsanteBwamu, CwiDinka, SoutheasternTwi

Tags

speechaudioasrtwimultilingual

Licenses

mit

Similar

michsethowusu/luganda_TTS_v1michsethowusu/kasanomamichsethowusu/akuapem_multispeaker_audio_transcribedmichsethowusu/SHOLAmichsethowusu/afri-bigramsmichsethowusu/ghana-speech

michsethowusu/luganda_TTS_v1

michsethowusu/kasanoma

Offline-first TTS models for African languages # Kasanoma – Offline TTS Models for languages of Afr

michsethowusu/akuapem_multispeaker_audio_transcribed

The Akuapem Multispeaker Audio Transcribed dataset is a collection of speech recordings and their tr

michsethowusu/SHOLA

SHOLA (Share Your Language) — volunteers verify translations of everyday words in Twi, Ewe, Ga and D

michsethowusu/afri-bigrams

This dataset contains a large-scale collection of word bigrams extracted from text across 154 Africa

michsethowusu/ghana-speech

Language Subset Segments Duration Akuapem_Twi Akuapem_Twi_twi 52,650 63.25h Anyin Anyin_any 13,986