Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

No Language Left Behind Seed Data (Standard Moroccan Tamazight)

Record type:

dataset
Creator:
Tam
Host:

Visit

huggingface.co

Languages

AmazighBerberGhomaraSenhaja BerberTamazight, Central AtlasTamazight, Standard MoroccanTarifit

Licenses

cc-by-sa-4.0

Similar

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))No Language Left Behind Seed Data (Tamasheq (Latin script))No Language Left Behind : DataNLLB: No Language Left BehindNo Student Left BehindNo Language Left Behind: Scaling Human-Centered Machine Translation

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))

No Language Left Behind Seed Data (Tamasheq (Latin script))

No Language Left Behind : Data

NLLB project uses data from three sources : public bitext, mined bitext and data generated using backtranslation. Details of different datasets used and open source links are provided in details here.

NLLB: No Language Left Behind

No Language Left Behind (NLLB) is a first-of-its-kind, AI breakthrough project that open-sources models capable of delivering high-quality translations directly between any pair of 200+ languages — including low-resource languages like Asturian, Luganda, Urdu and m

No Student Left Behind

The end of 2019 was punctuated by the emergence of an infectious disease spread through human-to-hum

No Language Left Behind: Scaling Human-Centered Machine Translation

Driven by the goal of eradicating language barriers on a global scale, machine translation has solidified itself as a key focus of artificial intelligence research today. However, such efforts have coalesced around a small subset of languages, leaving behind the va