Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

ddd-kenya-luhya-70hrs-asr

Domain:

natural language processing

Record type:

dataset
Creator:
Dig
Host:
A 70-hour subset of Luhya speech data collected by Digital Divide Data in Kenya. The dataset includes recorded sentences from native speakers and is intended to support research and development in Automatic Speech Recognition for low-resource African languages.

Visit

mozilladatacollective.com

Tasks

automatic speech recognitionspeech processing

Languages

Luhya

Tags

mdcmozilla data collectiveASRWAVXLSXTSV

Licenses

Creative Commons Attribution 4.0 International (CC-BY-4.0)

Similar

DDD-Kenya/Luhya-ASR-Data-subset-50hDDD-Kenya/Luhya-ASR-Data-subset-642HOdhuso/ddd-kenya-somali-asr-v1DDD-Kenya/Somali-ASR-Subset-68HOdhuso/ddd-kenya-gusii-asr-v1ddd-kenya-somali-68hrs-asr-part1

DDD-Kenya/Luhya-ASR-Data-subset-50h

DDD-Kenya/Luhya-ASR-Data-subset-642H

Odhuso/ddd-kenya-somali-asr-v1

DDD-Kenya/Somali-ASR-Subset-68H

Odhuso/ddd-kenya-gusii-asr-v1

ddd-kenya-somali-68hrs-asr-part1

This dataset, curated by Digital Divide Data (DDD), provides high-quality audio recordings and corre