Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

kambale/luganda-english-parallel-corpus

Domain:

natural language processing

Record type:

dataset
Creator:
kam
Host:
This dataset contains parallel sentences in English (en) and Luganda (lg), designed primarily for training and fine-tuning machine translation models. The data consists of sentence pairs extracted from a source document. English (en) Luganda (lg) - ISO 639-1 code: lg Data Format

Visit

huggingface.co

Tasks

machine translation

Languages

Ganda

Tags

translationparallel-corpuslow-resource

Licenses

apache-2.0

Similar

kambale/luganda-english-bible-corpusluganda-english-parallel-corpusAn English-Luganda parallel corpuskimrichies/English-Luganda-Parallel-corpusThe Makerere MT Corpus: English to Luganda parallel corpusBuilding a Parallel Corpus and Training Translation Models Between Luganda and English

kambale/luganda-english-bible-corpus

This dataset contains 32,291 parallel sentences in English (en) and Luganda (lg), derived from bibli

luganda-english-parallel-corpus

An English-Luganda parallel corpus

This English-Luganda parallel sentence corpus was created by a team of researchers fro

kimrichies/English-Luganda-Parallel-corpus

This is a bilingual corpus of English and Luganda for use in Neural Machine Translation tasks. I giv

The Makerere MT Corpus: English to Luganda parallel corpus

This English-Luganda parallel sentence corpus was created by a team of researchers fro

Building a Parallel Corpus and Training Translation Models Between Luganda and English

Neural machine translation (NMT) has achieved great successes with large datasets, so NMT is more premised on high-resource languages. This continuously underpins the low resource languages such as Luganda due to the lack of high-quality parallel corpora, so even ‘