Estructura dataset: El dataset está en formato de texto plano (.txt) con valores separados por tabul
This is a collection of translated sentences from Tatoeba 359 languages, 3,403 bitexts total number of files: 750 total number of tokens: 65.54M total number of sentence fragments: 8.96M
This dataset contains translations from Algerian Darja (Arabic dialect) to English. The dataset incl