Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Twi Trigrams Speech-Text Parallel Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
gha
Host:
This dataset contains 166156 parallel speech-text pairs for Twi, a language spoken primarily in Ghana. The dataset consists of audio recordings of trigram segments (3-word sequences) paired with their corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Language: Twi - twi

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

AkanBwamu, CwiDinka, SoutheasternTwi

Tags

speechtwighanaafrican-languageslow-resourceparallel-corpustrigramsn-grams

Licenses

cc-by-4.0