Logo Lanfrica
fr
Accueil
Atlas
Analyses
Documentation
Sign in
Retour
No Language Left Behind : Data
Domaine:
natural language processing
Type de record:
dataset
Visit
Actions
Share
Report an issue
NLLB project uses data from three sources : public bitext, mined bitext and data generated using backtranslation. Details of different datasets used and open source links are provided in details here.
Visit
github.com
Connected records
project
model
paper
Tasks
machine translation
Languages
Afrikaans
Aka
Akan
Amazigh
Amharic
Arabic, Egyptian Spoken
Arabic, Moroccan Spoken
Bamanankan
Bemba
Bwamu, Cwi
+56
View more
Tags
nllb
Licenses
https://github.com/facebookresearch/fairseq/blob/nllb/LICENSE.model.md