Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

NCHLT Siswati Auxiliary Speech Corpus

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Febe de WetLaura MartinusJaco Badenhorst
Éditeur:
Charl van HeerderEtienne BarnardMarelie DavelAlta de Waal
Éditeur:
CSIR Meraka InstituteNorth-West University
Hôte:avatar
The corpus contains orthographically transcribed broadband speech in each of South Africa's eleven official languages. Transcriptions are provided in XML format.

Visit

hdl.handle.net

Tasks

speech processing

Languages

Swati

Tags

Siswati; Speech corpora; Transcribed

Licenses

Creative Commons Attribution 3.0 Unported (CC BY 3.0): https://creativecommons.org/licenses/by/3.0/legalcode

Similaires

NCHLT Speech Corpus -- siSwatiNCHLT Siswati Speech CorpusNCHLT Speech Corpus -- siSwatiNCHLT Sepedi Auxiliary Speech CorpusNCHLT Setswana Auxiliary Speech CorpusNCHLT isiZulu Auxiliary Speech Corpus

NCHLT Speech Corpus -- siSwati

This is the siSwati language part of the NCHLT Speech Corpus of the South African languages. Languag

NCHLT Siswati Speech Corpus

Orthographically transcribed broadband speech corpus of approximately 56 hours, including a test sui

NCHLT Speech Corpus -- siSwati

This is the siSwati language part of the NCHLT Speech Corpus of the South African languages. Languag

NCHLT Sepedi Auxiliary Speech Corpus

The corpus contains orthographically transcribed broadband speech in each of South Africa's eleven o

NCHLT Setswana Auxiliary Speech Corpus

The corpus contains orthographically transcribed broadband speech in each of South Africa's eleven o

NCHLT isiZulu Auxiliary Speech Corpus

This is the Zulu language split of the NCHLT speech corpus (nchlt-clean split). It containes 56 hour