Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

hauWE: Hausa Words Embedding for Natural Language Processing

Domaine:

natural language processing

Type de record:

model
Créateur:
Abdulmumin, IdrisGal
Éditeur:
arXiv
Hôte:avatar
Words embedding (distributed word vector representations) have become an essential component of many natural language processing (NLP) tasks such as machine translation, sentiment analysis, word analogy, named entity recognition and word similarity. Despite this, the only work that provides word vectors for Hausa language is that of Bojanowski et al. [1] trained using fastText, consisting of only a few words vectors. This work presents words embedding models using Word2Vec's Continuous Bag of Words (CBoW) and Skip Gram (SG) models. The models, hauWE (Hausa Words Embedding), are bigger and better than the only previous model, making them more useful in NLP tasks. To compare the models, they were used to predict the 10 most similar words to 30 randomly selected Hausa words. hauWE CBoW's 88.7% and hauWE SG's 79.3% prediction accuracy greatly outperformed Bojanowski et al. [1]'s 22.3%. In Proceedings of the 2019 2nd International Conference of the IEEE Nigeria Computer Chapter

Visit

doi.orgarxiv.org

Tasks

embeddings

Languages

Hausa

Tags

word embeddings

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

QCSE: A Pretrained Quantum Context-Sensitive Word Embedding for Natural Language ProcessingA Systematic Literature Review of Hausa Natural Language ProcessingHausaNLP: Current Status, Challenges and Future Directions for Hausa Natural Language ProcessingNatural language processing for African languagesReplication Data for Igbo Natural Language Processing Tasks I Igbo Synchronised Corpus for Natural Language Processing TasksReplication Data for Igbo Natural Language Processing Tasks II Igbo Synchronised Corpus for Natural Language Processing Tasks

QCSE: A Pretrained Quantum Context-Sensitive Word Embedding for Natural Language Processing

Quantum Natural Language Processing (QNLP) offers a novel approach to encoding and understanding the

A Systematic Literature Review of Hausa Natural Language Processing

The processing of natural languages is an area of computer science that has gained growing attention

HausaNLP: Current Status, Challenges and Future Directions for Hausa Natural Language Processing

Hausa Natural Language Processing (NLP) has gained increasing attention in recent years, yet remains

Natural language processing for African languages

Recent advances in pre-training of word embeddings and language models leverage large amounts of unlabelled texts and self-supervised learning to learn distributed representations that have significantly improved the performance of deep learning models on a large v

Replication Data for Igbo Natural Language Processing Tasks I Igbo Synchronised Corpus for Natural Language Processing Tasks

The Igbo synchronised corpus (IgboSynCorp) is an annotated corpus of spoken Igbo created by a team o

Replication Data for Igbo Natural Language Processing Tasks II Igbo Synchronised Corpus for Natural Language Processing Tasks

The Igbo synchronised corpus (IgboSynCorp) is an annotated corpus of spoken Igbo created by a team o