Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

EQ39/Extractive-Swahili-Question-Answering-NLP-

Domain:

natural language processing

Record type:

project
Creator:
EQ39
Host:
# Extractive Swahili Question-Answering with DistilBERT focuses on developing a lightweight yet effective natural language processing (NLP) model for answering questions based on Swahili text. Using DistilBERT, a smaller and faster variant of BERT, this approach fine-tunes the model on Kencorpus Swahili Question Answering Dataset to extract precise answers from given passages. The project involves preprocessing Swahili text, handling tokenization challenges, and optimizing the model for accuracy and efficiency. By leveraging transfer learning and Swahili-specific linguistic adaptations, this work aims to improve accessibility to AI-driven information retrieval in Swahili, benefiting education, research, and automated customer support applications.

Visit

github.com

Tasks

question answering

Languages

Swahili

Similar

AmaSQuAD: A Benchmark for Amharic Extractive Question AnsweringSwahili Visual Question Answering DatasetSwahili Question-Answering Dataset for Horticulturesagwegeoffrey/Custom-Question-Answering-Using-NLP-BERT-ModelKenSwQuAD – A Question Answering Dataset for Swahili Low Resource LanguageSwahiliVQA: A Dataset for Visual Question Answering in Swahili Language

AmaSQuAD: A Benchmark for Amharic Extractive Question Answering

This research presents a novel framework for translating extractive question-answering datasets into

Swahili Visual Question Answering Dataset

Swahili Language VQA Dataset

Swahili Question-Answering Dataset for Horticulture

The dataset was created to contribute to Swahili language resources for natural language processing

sagwegeoffrey/Custom-Question-Answering-Using-NLP-BERT-Model

Low-resource Language Question Answering Using BERT-based Transformers for Kiswahili QnA Task.

KenSwQuAD – A Question Answering Dataset for Swahili Low Resource Language

This research developed a Kencorpus Swahili Question Answering Dataset KenSwQuAD from raw data of Swahili language, which is a low resource language predominantly spoken in Eastern African and also has speakers in other parts of the world. Question Answering datase

SwahiliVQA: A Dataset for Visual Question Answering in Swahili Language