Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Siswati NER Corpus

Domaine:

natural language processing

Type de record:

dataset
Créateur:
nwu
Hôte:
Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated with PERSON, LOCATION, ORGANISATION and MISCELLANEOUS tags.

Visit

huggingface.co

Tasks

information extractionnamed entity recognition

Languages

Swati

Licenses

other

Similaires

Siswati Ner CorpusMonolingual Siswati CorpusNCHLT Speech Corpus -- siSwatiNCHLT Speech Corpus -- siSwatiLwazi Siswati ASR corpusLwazi Siswati TTS corpus

Siswati Ner Corpus

Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated with PERSON, LOCATION, ORGANISATION and MISCELLANEOUS tags.

Monolingual Siswati Corpus

Monolingual corpus for SiSwati. The data is given as a single UTF-8 text file, with each segment on

NCHLT Speech Corpus -- siSwati

This is the siSwati language part of the NCHLT Speech Corpus of the South African languages. Languag

NCHLT Speech Corpus -- siSwati

This is the siSwati language part of the NCHLT Speech Corpus of the South African languages. Languag

Lwazi Siswati ASR corpus

Complete audio recordings and orthographic transcriptions used for Lwazi speech recognition systems.

Lwazi Siswati TTS corpus

Orthographic and phonemically aligned transcriptions