Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

AMR-KELEG/PTCC

Domain:

natural language processing

Record type:

dataset
Creator:
AMR
Host:
The Parallel Tunisian Constitution Corpus (PTCC) corpus is a corpus of 149 articles written in Modern Standard Arabic and Tunisian Arabic. Tesseract was used to transform the constitution's pdf files into text files. Afterward, alignment of the parallel articles was achieved by a simple Python script. More details can be found in: amr-keleg.github.io Tunisian Arabic translation of the 2014 Tunisian Constitution

Visit

huggingface.co

Tasks

machine translation

Languages

Arabic, Tunisian Spoken

Tags

legal

Licenses

mit

Similar

AMR TriageAMR ChallengeAMR Sentinel — Vivli 2026 AMR Data Challenge Pre-RegistrationAMR Surveillance DatasetMR-AMR Datasetkbastug/nanopore-amr-nigeria

AMR Triage

Background With high burden of antimicrobial resistance and steep increase in watch antibiotics in l

AMR Challenge

Antimicrobial resistance is a global problem that has so far resisted modern therapeutic interventio

AMR Sentinel — Vivli 2026 AMR Data Challenge Pre-Registration

Pre-registered secondary analysis linking Vivli AMR Surveillance Register (SPIDAAR, ATLAS, SMART) wi

AMR Surveillance Dataset

A synthetic tabular dataset for antimicrobial resistance surveillance in African healthcare faciliti

MR-AMR Dataset

MR-AMR is a public dataset composed of 140000 images of the digits in full-state and mid-state. It c

kbastug/nanopore-amr-nigeria

Bioinformatics workflows supporting nanopore-based bacterial identification and AMR gene detection i