Leveraging Domain Specific Lexicons to Improve Language Preservation and Question Answering Tasks in Large Language Models: A Case of Swahili Poster presented at the Deep Learning Indaba 2023 by Kevin Omondi
Native African languages are grossly underrepresented in prevailing large language models (LLMs). Th
The rapid advancement of large language models (LLMs) has significantly propelled progress in natura
This research developed a Kencorpus Swahili Question Answering Dataset KenSwQuAD from raw data of Swahili language, which is a low resource language predominantly spoken in Eastern African and also has speakers in other parts of the world. Question Answering datase
Large Language Models (LLMs) have shown remarkable progress in medical question answering (QA), yet