Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Cross-Encoder-Based Semantic Evaluation of Extractive and Generative Question Answering in Low-Resourced African Languages

Domain:

natural language processing

Record type:

paper
Creator:
FunYuaCheNob
Publisher:
MDP
Host:
Efficient language analysis techniques and models are crucial in the artificial intelligence age for enhancing cross-lingual question answering. Transfer learning with state-of-the-art models has been beneficial in this regard, but the performance of low-resource African languages with morphologically rich grammatical structures and unique typologies has shown deficiencies linkable to evaluation techniques and scarce training data. To enhance the former, this paper proposes an evaluation pipeline leveraging the semantic answer similarity method enhanced with automatic answer annotation. The pipeline uses the Language-agnostic BERT Sentence Embedding model integrated with an adapted vector measure to perform cross-lingual text analysis after answer prediction. Experimental results from the multilingual-T5 and AfroXLMR models on nine languages of the AfriQA dataset surpassed existing benchmarks deploying string-based methods for question answer evaluation. The results are also superior to the F1-score-based GPT4 and Llama-2 performances on the same downstream task. The automatic answer annotation technique effectively reduced the labelling time while maintaining a high performance. Thus, the proposed pipeline is more efficient than the prevailing string-based F1 and Exact Match metrics in mixed answer type question–answer evaluations, and it is a more natural performance estimator for models targeting real-world deployment.

Visit

doi.org

Tasks

question answering

Licenses

https://creativecommons.org/licenses/by/4.0/

Similar

EQ39/Extractive-Swahili-Question-Answering-NLP-SamAbr/Multilingual-Health-Question-Answering-in-Low-Resource-African-LanguagesOffei-op/Multi-Health-Question-Answering-In-Low-Resource-African-LanguagesAfriQA: Cross-lingual Open-Retrieval Question Answering for African LanguagesQuestion-Answering in a Low-resourced Language: Benchmark Dataset and Models for TigrinyaAmaSQuAD: A Benchmark for Amharic Extractive Question Answering

EQ39/Extractive-Swahili-Question-Answering-NLP-

# Extractive Swahili Question-Answering with DistilBERT focuses on developing a lightweight yet effe

SamAbr/Multilingual-Health-Question-Answering-in-Low-Resource-African-Languages

# Multilingual Health Question Answering in Low-Resource African Languages This repository contains

Offei-op/Multi-Health-Question-Answering-In-Low-Resource-African-Languages

# Lalang: multilingual health-QA retrieval and reranking This repository is the reproducible resear

AfriQA: Cross-lingual Open-Retrieval Question Answering for African Languages

African languages have far less in-language content available digitally, making it challenging for q

Question-Answering in a Low-resourced Language: Benchmark Dataset and Models for Tigrinya

Question-Answering (QA) has seen significant advances recently, achieving near human-level performan

AmaSQuAD: A Benchmark for Amharic Extractive Question Answering

This research presents a novel framework for translating extractive question-answering datasets into