Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

English-Kalenjin Translation Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
mut
Host:
This is an English-Kalenjin parallel text dataset prepared for machine translation research and model training. The dataset combines mined English-Kalenjin pairs, manually translated synthetic English from Swahili-Kalenjin candidates, and a small direct manual collection set. This dataset is intentionally kept gated for now. Upstream ANV data appears gated.

Visit

huggingface.co

Tasks

machine translation

Languages

KalenjinKipsigisSwahili

Tags

englishkalenjinlow-resource-languagemachine-translationparallel-corpuskenya

Licenses

other

Similar

Kiprop2020/English-Kalenjin-translationXhosa-English Translation DatasetKipropAmos/Kalenjin-Language-TranslationEnglish to Luo Translation Datasetzfermi/english-kalenjin-dictionaryBiatus/afrinllb-kikuyu-kalenjin-dataset

Kiprop2020/English-Kalenjin-translation

seq2seq model to translate English to Kalenjin # English-Kalenjin-translation seq2seq model to tran

Xhosa-English Translation Dataset

A High-Quality Parallel Corpus for Low-Resource Machine Translation

KipropAmos/Kalenjin-Language-Translation

# Kalenjin-Language-Translation ## Aim The project is aims at creating a translation alghorithim fro

English to Luo Translation Dataset

This dataset contains English to Luo translation pairs extracted and cleaned from the Opus corpus. T

zfermi/english-kalenjin-dictionary

This is a Next.js project bootstrapped with `create-next-app`. ## Getting Started First, run the d

Biatus/afrinllb-kikuyu-kalenjin-dataset