Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

A Statistically Validated Bidirectional Neural Machine Translation Framework for the Low-Resource Tigrigna–Kunama Language Pair

Domain:

natural language processing

Record type:

datasetmodel
Creator:
MekAbeGebNeg
Publisher:
Zenodo
Host:avatar
This repository contains the official dataset and experimental codebase for the study "A Statistically Validated Bidirectional Neural Machine Translation Framework for the Low-Resource Tigrigna–Kunama Language Pair." The dataset includes: An expert-verified parallel corpus of 4,712 sentence pairs. Original, authentic Ge'ez script morphology preserved without homophone normalization. The codebase includes: A unified joint-bidirectional Bi-LSTM architecture with Luong global attention. Preprocessing pipelines, BPE subword modeling (4,000 units), and automated training/evaluation scripts. Statistical validation protocols (paired bootstrap resampling, 5-fold cross-validation).

Visit

doi.org

Tasks

machine translation

Languages

KunamaTigrigna

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

Neural Machine Translation for Mooré, a Low-Resource LanguageLanguage-Family Adapters for Low-Resource Multilingual Neural Machine TranslationNeural Machine Translation for Low-Resource Languages: A SurveyNeural Machine Translation Models with Back-Translation for the Extremely Low-Resource Indigenous Language BribriData Augmentation for Low-Resource Neural Machine TranslationMultilingual Neural Machine Translation for Low Resource Languages

Neural Machine Translation for Mooré, a Low-Resource Language

Language-Family Adapters for Low-Resource Multilingual Neural Machine Translation

Large multilingual models trained with self-supervision achieve state-of-the-art results in a wide r

Neural Machine Translation for Low-Resource Languages: A Survey

Neural Machine Translation (NMT) has seen a tremendous spurt of growth in less than ten years, and h

Neural Machine Translation Models with Back-Translation for the Extremely Low-Resource Indigenous Language Bribri

This paper presents a neural machine translation model and dataset for the Chibchan language Bribri,

Data Augmentation for Low-Resource Neural Machine Translation

The quality of a Neural Machine Translation system depends substantially on the availability of siza

Multilingual Neural Machine Translation for Low Resource Languages

Neural Machine Translation (NMT) has been shown to be more effective in translation tasks compared t