Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Afrikaans Ner Corpus

Domain:

natural language processing

Record type:

dataset
Creator:
nwu
Host:
The Afrikaans Ner Corpus is an Afrikaans dataset developed by The Centre for Text Technology (CTexT), North-West University, South Africa. The data is based on documents from the South African goverment domain and crawled from gov.za websites. It was created to support NER task for Afrikaans language. The dataset uses CoNLL shared task annotation standards. Supported Tasks and Leaderboards

Visit

huggingface.co

Tasks

information extractionnamed entity recognition

Languages

Afrikaans

Licenses

other

Similar

Afrikaans Ner CorpusMZSFighters/Dutch-Afrikaans-NERSiswati NER CorpusSesotho Ner CorpusSepedi NER CorpusSetswana NER Corpus

Afrikaans Ner Corpus

Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated with PERSON, LOCATION, ORGANISATION and MISCELLANEOUS tags.

MZSFighters/Dutch-Afrikaans-NER

# Dutch-Afrikaans-NER Please find the research report here: Named Entity Recognition in Afrikaans:

Siswati NER Corpus

Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated wi

Sesotho Ner Corpus

Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated with PERSON, LOCATION, ORGANISATION and MISCELLANEOUS tags.

Sepedi NER Corpus

The Sepedi Ner Corpus is a Sepedi dataset developed by The Centre for Text Technology (CTexT), North

Setswana NER Corpus

Named entity annotated data from the NCHLT Text Resource Development: Phase II Project, annotated wi