Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Sepedi Speech Corpora

Domain:

natural language processing
Publisher:
University of Limpopo (Turfloop Campus)
Host:avatar
A corpus of Sesotho sa Leboa telephone speech data collected from mother tongue speakers of the standard version of Sesotho sa Leboa for the purpose of building a Sesotho sa Leboa ASR system.

Visit

hdl.handle.net

Tasks

automatic speech recognitionspeech processing

Languages

Sotho, NorthernSotho, Southern

Similar

Correction to: Two sepedi‑english code‑switched speech corporaNCHLT Sepedi Text CorporaAutshumato English-Sepedi Parallel CorporaNCHLT Sepedi Annotated Text CorporaNCHLT Speech Corpus -- SepediNCHLT Speech Corpus -- Sepedi

Correction to: Two sepedi‑english code‑switched speech corpora

NCHLT Sepedi Text Corpora

Collection of source text documents, genre classified text documents, raw corpus, clean corpus, lexi

Autshumato English-Sepedi Parallel Corpora

Aligned parallel corpora for the language pair English-Sepedi. The data is given as two separate UTF

NCHLT Sepedi Annotated Text Corpora

Lemmatised, part of speech tagged and morphologically analysed corpora developed during the NCHLT Te

NCHLT Speech Corpus -- Sepedi

This is the Sepedi language part of the NCHLT Speech Corpus of the South African languages. Language

NCHLT Speech Corpus -- Sepedi

This is the Sepedi language part of the NCHLT Speech Corpus of the South African languages. Language