Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

hausa-corpus

Domain:

natural language processing

Record type:

dataset
A collection of textual datasets in Hausa language and the corresponding translation in English language.

Visit

github.com

Connected records

paper

Tasks

machine translation

Languages

Hausa

Tags

large datasets from Lanfrica Insights

Similar

Hausa Corpushausa-text-corpusEnglish-Hausa CorpusHausa Speech Corpusijdutse/hausa-corpusHausa VOA NER Corpus

Hausa Corpus

Hausa datasets with stopwords

hausa-text-corpus

English-Hausa Corpus

English-Hausa parallel corpus.

Hausa Speech Corpus

This is a Hausa Speech data set that was recorded as a baseline for Hausa Speech Recognition. The data sets can be used in building Automatic Speech recognition for Hausa language, Speech synthesis and speaker recognition.

ijdutse/hausa-corpus

A collection of textual datasets in Hausa language and the corresponding translation in English lang

Hausa VOA NER Corpus

The Hausa VOA NER dataset is a labeled dataset for named entity recognition in Hausa. The texts were