Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Scaling Tonally-Sensitive African Language Models with NAIRR Resources

Domain:

natural language processing

Record type:

papermodel
Creator:
Odo
Publisher:
Zenodo
Host:avatar

Presented at the “2026 NAIRR Annual Meeting” (NSF Award #2536728), Arlington, VA, USA, March 10-13, 2026. https://zenodo.org/communit…

Visit

doi.org

Tasks

language modeling

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution No Derivatives 4.0 Internationalhttps://creativecommons.org/licenses/by-nd/4.0/legalcode

Similar

Glot500: Scaling Multilingual Corpora and Language Models to 500 LanguagesStory Generation with Large Language Models for African LanguagesQuestion Answering Research in African Languages with Large Language Models South African Language Resources: Phrase ChunkingScaling Multilingual Language Models for Low-Resource Cross-Lingual NER on the XTREME BenchmarkScaling Multilingual Language Models and CLCA Score Improvements via Optimal Transport Distillation in MIRACL

Glot500: Scaling Multilingual Corpora and Language Models to 500 Languages

The NLP community has mainly focused on scaling Large Language Models (LLMs) vertically, i.e., makin

Story Generation with Large Language Models for African Languages

Question Answering Research in African Languages with Large Language Models

Native African languages are grossly underrepresented in prevailing large language models (LLMs). Th

South African Language Resources: Phrase Chunking

Phrase chunking remains an important natural language processing (NLP) technique for intermediate syntactic processing. This paper describes the development of protocols, annotated phrase chunking data sets and automatic phrase chunkers for ten South African langua

Scaling Multilingual Language Models for Low-Resource Cross-Lingual NER on the XTREME Benchmark

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Scaling Multilingual Language Models and CLCA Score Improvements via Optimal Transport Distillation in MIRACL

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi