Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

innadark/topxgen-gemma-3-27b-and-nllb-3.3b

Domain:

natural language processing

Record type:

dataset
Creator:
inn
Host:
This dataset is a synthetic parallel dataset for 10 low-resource languages, created by applying the TopXGen pipeline with recent multilingual LLMs. It is designed for machine translation (MT) fine-tuning and few-shot experiments (as a selection pool).The pipeline works as follows:

Visit

huggingface.co

Tasks

machine translation

Languages

HausaIgboKinyarwandaSomaliSwahiliXhosa

Similar

EPFLiGHT/Gemma-3-27B-MeditronFO-Amharicmradermacher/Gemma-3-27B-MeditronFO-Amharic-GGUFmimech011/nllb-200-kabyle-3.3Bmogu98/nllb-3.3B-wolof-v8mimech011/nllb-200-kabyle-3.3B-8bitb1n1yam/gemma-2-27b-amharic-cpt

EPFLiGHT/Gemma-3-27B-MeditronFO-Amharic

mradermacher/Gemma-3-27B-MeditronFO-Amharic-GGUF

mimech011/nllb-200-kabyle-3.3B

mogu98/nllb-3.3B-wolof-v8

mimech011/nllb-200-kabyle-3.3B-8bit

b1n1yam/gemma-2-27b-amharic-cpt