Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

No Language Left Behind Seed Data (Standard Moroccan Tamazight)

Type de record:

dataset
Créateur:
Tam
Hôte:

Visit

huggingface.co

Languages

AmazighBerberGhomaraSenhaja BerberTamazight, Central AtlasTamazight, Standard MoroccanTarifit

Licenses

cc-by-sa-4.0

Similaires

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))No Language Left Behind Seed Data (Tamasheq (Latin script))No Language Left Behind : DataNLLB: No Language Left BehindNo Student Left BehindNo Language Left Behind: Scaling Human-Centered Machine Translation

No Language Left Behind Seed Data (Tamasheq (Tifinagh script))

No Language Left Behind Seed Data (Tamasheq (Latin script))

No Language Left Behind : Data

NLLB project uses data from three sources : public bitext, mined bitext and data generated using backtranslation. Details of different datasets used and open source links are provided in details here.

NLLB: No Language Left Behind

No Language Left Behind (NLLB) is a first-of-its-kind, AI breakthrough project that open-sources models capable of delivering high-quality translations directly between any pair of 200+ languages — including low-resource languages like Asturian, Luganda, Urdu and m

No Student Left Behind

The end of 2019 was punctuated by the emergence of an infectious disease spread through human-to-hum

No Language Left Behind: Scaling Human-Centered Machine Translation

Driven by the goal of eradicating language barriers on a global scale, machine translation has solidified itself as a key focus of artificial intelligence research today. However, such efforts have coalesced around a small subset of languages, leaving behind the va