Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

AI-based disease category prediction model using symptoms from low-resource Ethiopian language: Afaan Oromo text

Domain:

healthcarenatural language processing

Record type:

paper
Creator:
EtaMriTek
Publisher:
Spr
Host:
Abstract Automated disease diagnosis and prediction, powered by AI, play a crucial role in enabling medical professionals to deliver effective care to patients. While such predictive tools have been extensively explored in resource-rich languages like English, this manuscript focuses on predicting disease categories automatically from symptoms documented in the Afaan Oromo language, employing various classification algorithms. This study encompasses machine learning techniques such as support vector machines, random forests, logistic regression, and Naïve Bayes, as well as deep learning approaches including LSTM, GRU, and Bi-LSTM. Due to the unavailability of a standard corpus, we prepared three data sets with different numbers of patient symptoms arranged into 10 categories. The two feature representations, TF-IDF and word embedding, were employed. The performance of the proposed methodology has been evaluated using accuracy, recall, precision, and F1 score. The experimental results show that, among machine learning models, the SVM model using TF-IDF had the highest accuracy and F1 score of 94.7%, while the LSTM model using word2vec embedding showed an accuracy rate of 95.7% and F1 score of 96.0% from deep learning models. To enhance the optimal performance of each model, several hyper-parameter tuning settings were used. This study shows that the LSTM model verifies to be the best of all the other models over the entire dataset.

Visit

doi.org

Tasks

text classification

Languages

AmharicOromoOromo, Borana-Arsi-Guji

Licenses

https://creativecommons.org/licenses/by/4.0https://creativecommons.org/licenses/by/4.0

Similar

Deep Learning based Multilabel Hateful Speech Text Comments Recognition and Classification Model for Resource Scarce Ethiopian Language: The case of Afaan OromoAUTOMATIC THESAURUS CONSTRUCTION FROM AFAAN OROMO TEXT USING WORD EMBEDDINGMube2021/Afaan-Oromo-Text-Summarization-using-NLPSigned Language Translation into Afaan Oromo Text Using Deep-Learning ApproachRelation Extraction (RE) Model for Afaan Oromo Text Using Self-Attention MechanismsAFAAN OROMO TEXT SEMANTIC NETWORK ANALYSIS MODEL FOR CLASSIFICATION OF TEXT USING DEEP LEARNING APPROACH

Deep Learning based Multilabel Hateful Speech Text Comments Recognition and Classification Model for Resource Scarce Ethiopian Language: The case of Afaan Oromo

AUTOMATIC THESAURUS CONSTRUCTION FROM AFAAN OROMO TEXT USING WORD EMBEDDING

 Principal Advisor:  Gadisa Olani (PhD) ABSTRACT A Thesaurus is

Mube2021/Afaan-Oromo-Text-Summarization-using-NLP

Take Afaan Oromo text Generate a short summary Show evaluation (ROUGE score) # Afaan-Oromo-Text-Sum

Signed Language Translation into Afaan Oromo Text Using Deep-Learning Approach

Relation Extraction (RE) Model for Afaan Oromo Text Using Self-Attention Mechanisms

Abstract This study proposes a novel Relation Extraction (RE) model for Afaan Orom

AFAAN OROMO TEXT SEMANTIC NETWORK ANALYSIS MODEL FOR CLASSIFICATION OF TEXT USING DEEP LEARNING APPROACH

Major Advisor: Dr. M. Kumarasamy (PhD) Natural Language Processing is the intersection of computer