Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

AfroMAFT Corpus: Language Adaptation Corpus for African languages

Domaine:

natural language processing

Type de record:

datasetmodel
Créateur:
David Ifeoluwa AdelaniJesujoba O. Alabi
Éditeur:
Zenodo
Hôte:avatar

Language Adaptation Corpus for 17 African languages, English, French, and Arabic.

We used this corpus to train the following pre-trained language models:

  • AfroXLMR
  • AfriMT5
  • AfriByT5
  • AfriMBART

If you use this corpus, please cite the MAFAND paper and mC4 paper. 

Visit

doi.org

Tasks

language modeling

Licenses

info:eu-repo/semantics/openAccessNon-Commercial Government Licencehttps://github.com/spdx/license-list-XML/blob/master/src/Apache-2.0.xml