Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

afrolm-dataset

Domain:

natural language processing

Record type:

dataset
Creator:
bon
Host:
GitHub Repository of the Paper This repository contains the dataset for our paper AfroLM: A Self-Active Learning-based Multilingual Pretrained Language Model for 23 African Languages which will appear at the third Simple and Efficient Natural Language Processing, at EMNLP 2022. Languages Covered

Visit

huggingface.co

Tasks

language modeling

Languages

AmharicBamanankanChichewaDholuoÉwéFonGandaGhomálá’HausaIgbo+12

Tags

afrolmactive learninglanguage modelingresearch papersnatural language processingself-active learning

Licenses

cc-by-4.0

Similar

sato2ru/swahili-emotion-afrolmAfroLM: A Self-Active Learning-based Multilingual Pretrained Language Model for 23 African Languages

sato2ru/swahili-emotion-afrolm

AfroLM: A Self-Active Learning-based Multilingual Pretrained Language Model for 23 African Languages

In recent years, multilingual pre-trained language models have gained prominence due to their remark