Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

abdelhaqueidali/Kabyle-Latin-to-Tifinagh-Parallel-Corpus

Domain:

natural language processing

Record type:

dataset
Creator:
abd
Host:
This dataset provides a parallel corpus of the Kabyle language (Taqbaylit), pairing native Latin-based orthography with automated, context-aware Amazigh script transliterations. It is built by processing raw text data through a rule-based algorithmic pipeline designed to enforce strict orthographic purity, manage contextual phonetic mutations, and isolate foreign vocabulary.

Visit

huggingface.co

Tasks

text normalization

Languages

AmazighBerber

Tags

berberamazightamazightkabyletifinaghtransliterationtatoeba

Licenses

cc-by-2.0