This repository contains various Tamazight language datasets created by Col·lectivaT in collaboration with CIEMEN and with funding from Municipality of Barcelona and Government of Catalonia.
Under mono you can find monolingual sentences.
tc_wajdm_v1.txt - Texts from language learning material “tc wawjdm”
IRCAM-clean-tifinagh.txt - Tifinagh scripted sentences extracted from IRCAM's text corpus
Under parallel you can find sentences with translations.