Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

NaijaRC: A Multi-choice Reading Comprehension Dataset for Nigerian Languages

Domain:

natural language processing

Record type:

paperdataset
Creator:
AreAlaAboAgu
Host:avatar
In this paper, we create NaijaRC: a new multi-choice Reading Comprehension dataset for three native Nigeria languages that is based on high-school reading comprehension examination. We provide baseline results by performing cross-lingual transfer using existing English RACE and Belebele training dataset based on a pre-trained encoder-only model. Additionally, we provide results by prompting large language models (LLMs) like GPT-4. Accepted to AfricaNLP Workshop at ICLR 2024 (non-archival)

Visit

arxiv.org

Tasks

question answering

Tags

Computation and Language

Similar

DREAM: A Challenge Dataset and Models for Dialogue-Based Reading ComprehensionA Swahili Question-Answering Dataset for Machine Reading Comprehension in HorticultureThe Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language VariantsY-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension with Open-Ended QuestionsInvestigating the Comprehension Iceberg: Developing Empirical Benchmarks for Early-grade Reading in Agglutinating African LanguagesCan LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset

DREAM: A Challenge Dataset and Models for Dialogue-Based Reading Comprehension

We present DREAM, the first dialogue-based multiple-choice reading comprehension dataset. Collected

A Swahili Question-Answering Dataset for Machine Reading Comprehension in Horticulture

The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

We present Belebele, a multiple-choice machine reading comprehension (MRC) dataset spanning 122 lang

Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension with Open-Ended Questions

The purpose of this work is to share an English-Yorùbá evaluation dataset for openbook reading compr

Investigating the Comprehension Iceberg: Developing Empirical Benchmarks for Early-grade Reading in Agglutinating African Languages

Reading development in agglutinating African languages is a relatively under-researched

Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset

Open-ended questions, which require students to produce multi-word, nontrivial responses, are a popu