Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

NCHLT Tshivenda Auxiliary Speech Corpus

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Febe de WetLaura MartinusJaco Badenhorst
Éditeur:
Charl van HeerderEtienne BarnardMarelie DavelAlta de Waal
Éditeur:
CSIR Meraka InstituteNorth-West University
Hôte:avatar
The corpus contains orthographically transcribed broadband speech in each of South Africa's eleven official languages. Transcriptions are provided in XML format.

Visit

hdl.handle.net

Tasks

speech processing

Languages

Venda

Tags

Tshivenda; Speech corpora; Transcribed

Licenses

Creative Commons Attribution 3.0 Unported (CC BY 3.0): https://creativecommons.org/licenses/by/3.0/legalcode

Similaires

NCHLT Speech Corpus -- TshivendaNCHLT Speech Corpus -- TshivendaNCHLT Tshivenda Speech CorpusNCHLT isiZulu Auxiliary Speech CorpusNCHLT Auxiliary Speech Corpus - MultilingualNCHLT Afrikaans Auxiliary Speech Corpus

NCHLT Speech Corpus -- Tshivenda

This is the Tshivenda language part of the NCHLT Speech Corpus of the South African languages. Langu

NCHLT Speech Corpus -- Tshivenda

This is the Tshivenda language part of the NCHLT Speech Corpus of the South African languages. Langu

NCHLT Tshivenda Speech Corpus

Orthographically transcribed broadband speech corpus of approximately 56 hours, including a test sui

NCHLT isiZulu Auxiliary Speech Corpus

This is the Zulu language split of the NCHLT speech corpus (nchlt-clean split). It containes 56 hour

NCHLT Auxiliary Speech Corpus - Multilingual

This is a combined multilingual version of the NCHLT Auxiliary Speech Corpus, compiled by the Data S

NCHLT Afrikaans Auxiliary Speech Corpus

The corpus contains orthographically transcribed broadband speech in each of South Africa's eleven o