Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

dialectra-hausa-speech-corpus-v1

Domain:

natural language processing

Record type:

dataset
Creator:
Dia
Host:
This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Curated by: Dialectra Funded by [optional]: Dialectra Shared by [optional]: Dialectra Language(s) (NLP): Hausa (hau) License: CC BY 4.0 Repository: [More Information Needed] Paper [optional]: [More Information Needed]

Visit

huggingface.co

Tasks

speech processing

Languages

Hausa

Licenses

cc-by-4.0

Similar

Hausa Speech Corpus0xlawal/hausa-ai-v1HausaHate: An Expert Annotated Corpus for Hausa Hate Speech DetectionSwahili Large Corpus (v1)Somali Web Corpus V1datawise-africa/sheria-corpus-v1

Hausa Speech Corpus

This is a Hausa Speech data set that was recorded as a baseline for Hausa Speech Recognition. The data sets can be used in building Automatic Speech recognition for Hausa language, Speech synthesis and speaker recognition.

0xlawal/hausa-ai-v1

This is v1(version 1) of Sannu AI. Sannu AI that was built to help and enhance communication between

HausaHate: An Expert Annotated Corpus for Hausa Hate Speech Detection

We introduce the first expert annotated corpus of Facebook comments for Hausa hate speech detection.

Swahili Large Corpus (v1)

The Swahili Large Corpus (v1) is one of the largest and most diverse open pretraining datasets for t

Somali Web Corpus V1

This dataset consists of clean, structured, and filtered Somali language text compiled from various

datawise-africa/sheria-corpus-v1

The Sheria Corpus v1 is a curated collection of Kenyan legal case summaries from both the High Court