Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Makhuwa Trigrams Speech-Text Parallel Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
mic
Host:
This dataset contains 154253 parallel speech-text pairs for Makhuwa, a language spoken primarily in Mozambique. The dataset consists of audio recordings of trigram segments (3-word sequences) paired with their corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Language: Makhuwa - vmw

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processingtext to speech

Languages

Makhuwa

Tags

speechmakhuwamozambiqueafrican-languageslow-resourceparallel-corpustrigramsn-grams

Licenses

cc-by-4.0