Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Bridging context gaps in low-resource language chatbots through multilevel attention and hybrid embedding approaches

Domain:

natural language processing

Record type:

paper
Creator:
GodIkeJulAlo
Publisher:
Nig
Host:
Conversational agents for low-resource languages (LRLs), such as Igbo, face major challenges, including limited annotated data, code-switching, and weak contextual coherence in multi-turn dialogue. This study proposes a multilevel-attention and hybrid-embedding framework that integrates FastText subword representations with multilingual BERT (mBERT) to improve semantic understanding and context retention. The architecture applies hierarchical attention at the word, utterance, and dialogue levels, enabling effective modeling of conversational dependencies and reducing context drift. The model was evaluated on a curated Igbo--English conversational dataset and benchmarked against long short-term memory (LSTM), Transformer, FastText, mBERT, and XLM-R baselines. For response generation, the proposed framework achieved a bilingual evaluation understudy (BLEU) score of 44.1%, a longest-common-subsequence recall-oriented understudy for gisting evaluation (ROUGE-L) score of 60.3%, and a context-retention accuracy (CRA) of 81.5%. For intent classification, it attained an F1-score of 87.3% and an area under the receiver operating characteristic curve (ROC-AUC) of 0.91; for context-dependency detection, it achieved an F1-score of 84.3%. The framework also reduced inference latency and was robust to code-switching and noisy conversational input. Human evaluation confirmed improvements in response clarity, cultural relevance, and multi-turn coherence. The findings show that hybrid embeddings combined with multilevel attention provide an effective and scalable approach to conversational AI for LRLs, with potential applicability to other African languages.

Visit

doi.org

Tasks

natural language generation

Languages

Igbo

Licenses

https://creativecommons.org/licenses/by/4.0

Similar

Optimal Transport Distillation for Low-Resource Language Embedding AlignmentBridging the Gap: Practical Approaches to Reproducible Bioinformatics Workflows in Low-Resource SettingsBRIDGING GAPS IN LOW-RESOURCE LANGUAGES: A MACHINE LEARNING APPROACH TO PRONOMINAL ANAPHORA RESOLUTIONBridging language gaps in multilingual large language modelsBridging Digital Inclusion Gaps in Rural South Africa: Strategic Approaches and InnovationsLGSE: Lexically Grounded Subword Embedding Initialization for Low-Resource Language Adaptation

Optimal Transport Distillation for Low-Resource Language Embedding Alignment

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Bridging the Gap: Practical Approaches to Reproducible Bioinformatics Workflows in Low-Resource Settings

In bioinformatics, reproducibility often assumes access to high-performance computing, stable intern

BRIDGING GAPS IN LOW-RESOURCE LANGUAGES: A MACHINE LEARNING APPROACH TO PRONOMINAL ANAPHORA RESOLUTION

This study proposes the Kazakh Coreference Adaptation (KCA) model, a hybrid framework for resolving

Bridging language gaps in multilingual large language models

Large language models (LLMs) have revolutionized natural language processing, yet significant perfor

Bridging Digital Inclusion Gaps in Rural South Africa: Strategic Approaches and Innovations

This study addresses a current research gap in Computer Science concerning Strategies for B

LGSE: Lexically Grounded Subword Embedding Initialization for Low-Resource Language Adaptation

Adapting pretrained language models to low-resource, morphologically rich languages remains a signif