Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

AfriQA: Cross-lingual Open-Retrieval Question Answering for African Languages

Domain:

natural language processing

Record type:

paperdataset
Creator:
OguGwaRivCla
Host:avatar
African languages have far less in-language content available digitally, making it challenging for question answering systems to satisfy the information needs of users. Cross-lingual open-retrieval question answering (XOR QA) systems -- those that retrieve answer content from other languages while serving people in their native language -- offer a means of filling this gap. To this end, we create AfriQA, the first cross-lingual QA dataset with a focus on African languages. AfriQA includes 12,000+ XOR QA examples across 10 African languages. While previous datasets have focused primarily on languages where cross-lingual QA augments coverage from the target language, AfriQA focuses on languages where cross-lingual answer content is the only high-coverage source of answer content. Because of this, we argue that African languages are one of the most important and realistic use cases for XOR QA. Our experiments demonstrate the poor performance of automatic translation and multilingual retrieval methods. Overall, AfriQA proves challenging for state-of-the-art QA models. We hope that the dataset enables the development of more equitable QA technology.

Visit

arxiv.org

Tasks

question answering

Tags

Computation and LanguageArtificial IntelligenceInformation Retrieval

Similar

AfriCLIRMatrix: Enabling Cross-Lingual Information Retrieval for African LanguagesAbelAdissu/Cross-Lingual-Question-Answering-for-Amharic-Language-Using-Pretrained-LLMs-XLDA: Cross-Lingual Data Augmentation for Natural Language Inference and Question AnsweringAfri-MCQA: Multimodal Cultural Question Answering for African LanguagesAfri-MCQA: Multimodal Cultural Question Answering for African LanguagesUnsupervised Cross-Domain and Cross-Lingual Methods for Text Classification, Slot-Filling, and Question-Answering

AfriCLIRMatrix: Enabling Cross-Lingual Information Retrieval for African Languages

Language diversity in NLP is critical in enabling the development of tools for a wide range of users.However, there are limited resources for building such tools for many languages, particularly those spoken in Africa.For search, most existing datasets feature few

AbelAdissu/Cross-Lingual-Question-Answering-for-Amharic-Language-Using-Pretrained-LLMs-

## **INTRODUCTION** 📖 Welcome to the Amharic Text Generation project, a journey into the realm of na

XLDA: Cross-Lingual Data Augmentation for Natural Language Inference and Question Answering

While natural language processing systems often focus on a single language, multilingual transfer le

Afri-MCQA: Multimodal Cultural Question Answering for African Languages

Paper Afri-MCQA is the first multilingual cultural question-answering benchmark covering 8k Q&A pai

Afri-MCQA: Multimodal Cultural Question Answering for African Languages

Africa is home to over one-third of the world's languages, yet remains underrepresented in AI resear

Unsupervised Cross-Domain and Cross-Lingual Methods for Text Classification, Slot-Filling, and Question-Answering

Transfer learning has significantly revolutionized modern machine learning systems by instilling the