Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

No Language Left Behind Seed Data (Tamasheq (Latin script))

Record type:

dataset
Creator:
Tam
Host:

Visit

huggingface.co

Languages

BerberTamasheq

Licenses

cc-by-sa-4.0

Similar

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))No Language Left Behind Seed Data (Standard Moroccan Tamazight)No Language Left Behind : DataNLLB: No Language Left Behind

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))

No Language Left Behind Seed Data (Standard Moroccan Tamazight)

No Language Left Behind : Data

NLLB project uses data from three sources : public bitext, mined bitext and data generated using backtranslation. Details of different datasets used and open source links are provided in details here.

NLLB: No Language Left Behind

No Language Left Behind (NLLB) is a first-of-its-kind, AI breakthrough project that open-sources models capable of delivering high-quality translations directly between any pair of 200+ languages — including low-resource languages like Asturian, Luganda, Urdu and m