Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Kasem Speech-Text Parallel Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
gha
Host:
This dataset contains 75990 parallel speech-text pairs for Kasem, a language spoken primarily in Ghana. The dataset consists of audio recordings paired with their corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Language: Kasem - xsm Task: Speech Recognition, Text-to-Speech

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

Kasem

Tags

speechkasemburkina-fasoghanaafrican-languageslow-resourceparallel-corpus

Licenses

cc-by-4.0

Similar

Yoruba Speech-Text Parallel DatasetGa Speech-Text Parallel DatasetVagla Speech-Text Parallel DatasetDeg Speech-Text Parallel DatasetYoruba Speech-Text Parallel DatasetVai Speech-Text Parallel Dataset

Yoruba Speech-Text Parallel Dataset

This dataset contains 1647022 parallel speech-text pairs for Yoruba, a language spoken primarily in

Ga Speech-Text Parallel Dataset

This dataset is made available because of Ghana NLP's volunteer driven research work. Please conside

Vagla Speech-Text Parallel Dataset

This dataset contains 48605 parallel speech-text pairs for Vagla, a language spoken primarily in Gha

Deg Speech-Text Parallel Dataset

This dataset contains 125958 parallel speech-text pairs for Deg, a language spoken primarily in Ghan

Yoruba Speech-Text Parallel Dataset

Speech recognition, text-to-speech synthesis, voice assistants, language modeling Notes / challenge

Vai Speech-Text Parallel Dataset

This dataset contains 23286 parallel speech-text pairs for Vai, a language spoken primarily in Ghana