Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

ZA_LEX: lexical resources for South African languages

Domain:

natural language processing

Record type:

dataset
This repository contains lexical pronunciation resources and modules for use in text-to-speech (TTS) systems. Specifically, it was originally set up to track work on updating and enhancing existing resources for the NTTS project funded by the Department of Arts and Culture (DAC) of the Government of South Africa.

Visit

github.com

Tasks

text to speechspeech processing

Languages

AfrikaansSetswanaSotho, SouthernXhosaZulu

Licenses

Similar

A repository of free lexical resources for African languagesDeveloping Text Resources for Ten South African LanguagesTowards Augmenting Lexical Resources for Slang and African American EnglishDevelopment of linguistically annotated parallel language resources for four South African languagesMining Wikidata for Name Resources for African LanguagesMining Wikidata for Name Resources for African Languages

A repository of free lexical resources for African languages

Developing Text Resources for Ten South African Languages

The development of linguistic resources for use in natural language processing is of utmost importance for the continued growth of research and development in the field, especially for resource-scarce languages. In this paper we describe the process and challenges

Towards Augmenting Lexical Resources for Slang and African American English

Researchers in natural language processing have developed large, robust resources for understanding

Development of linguistically annotated parallel language resources for four South African languages

For this project, we collected and annotated data to develop language resources for the four official South African Nguni languages written with a conjunctive orthography. The data for these four languages is parallel to allow for comparative (computational) lingui

Mining Wikidata for Name Resources for African Languages

This work supports further development of language technology for the languages of Africa by providing a Wikidata-derived resource of name lists corresponding to common entity types (person, location, and organization). While we are not the first to mine Wikidata f

Mining Wikidata for Name Resources for African Languages

This work supports further development of language technology for the languages of Africa by providing a Wikidata-derived resource of name lists corresponding to common entity types (person, location, and organization). While we are not the first to mine Wikidata f