130 883 aligned sentence pairs extracted from the open Tatoeba database.
Download – raw Tatoeba dumps
Gather – filter English & Kabyle sentences
Align – pair by sentence-id
Fix – normalise
All steps were performed with the kabyle-nlp-toolkit.
File
Lines
Size
Format
en-kab-parallel.jsonl
130 883
10.7 MiB
One JSON object per line: {"en": "…", "kab": "…"}
Example