Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

A comparative study of natural language inference in Swahili using monolingual and multilingual models

Domain:

natural language processing

Record type:

paper
Creator:
Hajra, Faki AliAdila, Alfa Krisnadhi
Publisher:
Zenodo
Host:avatar

Recent advancements in large language models (LLMs) have led to opportunities for improving applications across various domains. However, existing LLMs fine-tuned for Swahili or other African languages often rely on pre-trained multilingual models, resulting in a relatively small portion of training data dedicated to Swahili. In this study, we compare the performance of monolingual and multilingual models in Swahili natural language inference tasks using the cross-lingual natural language inference (XNLI) dataset. Our research demonstrates the superior effectiveness of dedicated Swahili monolingual models, achieving an accuracy rate of 69%. These monolingual models exhibit significantly enhanced precision, recall, and F1 scores, particularly in predicting contradiction and neutrality. Overall, the findings in this article emphasize the critical importance of using monolingual models in low-resource language processing contexts, providing valuable insights for developing more efficient and tailored natural language processing systems that benefit languages facing similar resource constraints.

Visit

doi.org

Tasks

natural language inference

Languages

Swahili

Tags

Monolingual modelMultilingual modelNatural language inferenceSwahiliXNLI dataset

Licenses

info:eu-repo/semantics/openAccessCreative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode