Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Datasets Collection Framework for Low-Resourced Languages in South Africa

Creator:
NonSkhMat
Publisher:
IEEE
Host:

Visit

doi.org

Licenses

https://doi.org/10.15223/policy-029https://doi.org/10.15223/policy-037

Similar

Building Text and Speech Datasets for Low Resourced Languages: A Case of Languages in East AfricaMachine Translation for Morphologically Rich Low-Resourced South African LanguagesAfri Code Datasets (A collection of datasets for code generation in African languages)Weakly-supervised Deep Cognate Detection Framework for Low-Resourced Languages Using Morphological Knowledge of Closely-Related LanguagesBuilding Text and Speech Benchmark Datasets and Models for Low‐Resourced East African Languages: Experiences and LessonsSpeech data collection system for KUI, a Low resourced tribal language

Building Text and Speech Datasets for Low Resourced Languages: A Case of Languages in East Africa

Africa has over 2000 languages; however, those languages are not well represented in the existing Natural Language Processing ecosystem. African languages lack essential digital resources to be engaged effectively in the advancing language technologies. This growin

Machine Translation for Morphologically Rich Low-Resourced South African Languages

Afri Code Datasets (A collection of datasets for code generation in African languages)

Training and evaluating Large Language Models (LLMs) for code generation, building AI-powered coding

Weakly-supervised Deep Cognate Detection Framework for Low-Resourced Languages Using Morphological Knowledge of Closely-Related Languages

Exploiting cognates for transfer learning in under-resourced languages is an exciting opportunity fo

Building Text and Speech Benchmark Datasets and Models for Low‐Resourced East African Languages: Experiences and Lessons

ABSTRACT Africa has over 2000 languages; however, those languages are not well represented in the e

Speech data collection system for KUI, a Low resourced tribal language

A new generation of speech translation technology is being developed to enable natural cros