Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

yigagilbert/kinyarwanda-speech-trimmed

Domain:

natural language processing

Record type:

dataset
Creator:
yig
Host:
This dataset contains processed Kinyarwanda speech data with trimmed audio segments. The dataset contains two splits: dev_test: 9,263 samples test: 9,265 samples Each sample contains: id: Unique identifier for the sample audio: Audio data audio_language: Language of the audio (Kinyarwanda) text: Transcription of the audio prompt: Associated prompt or context

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

Kinyarwanda

Tags

speechaudiokinyarwandaautomatic-speech-recognitiontrimmed-audio

Similar

yigagilbert/luganda-english-speech-builderKinyarwanda-speech-to-text-ASR/Kinyarwanda-speech-to-text-ASRbadrex/kinyarwanda-speech-samplebadrex/kinyarwanda-speech-1000hKinyarwanda CommonVoice Speech Datasetbadrex/kinyarwanda-speech-500h

yigagilbert/luganda-english-speech-builder

Builds a clean Luganda–English paired speech dataset with translation, TTS synthesis, validation, an

Kinyarwanda-speech-to-text-ASR/Kinyarwanda-speech-to-text-ASR

Kinyarwanda Automatic Speech Recognition Model based on Whisper # KinyaWhisper - Kinyarwanda Automa

badrex/kinyarwanda-speech-sample

This dataset contains a sample from the 500 hours of Kinyarwanda speech data covering Health, Govern

badrex/kinyarwanda-speech-1000h

This dataset contains ~1000 hours of transcribed Kinyarwanda speech data covering Health, Government

Kinyarwanda CommonVoice Speech Dataset

Speech dataset including more than 2000 hours of annotated Kinyarwanda voice data from more than 1000 speakers.

badrex/kinyarwanda-speech-500h

This dataset contains 500 hours of transcribed Kinyarwanda speech data covering Health, Government,