Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

AMR-KELEG/PTCC

Domaine:

natural language processing

Type de record:

dataset
Créateur:
AMR
Hôte:
The Parallel Tunisian Constitution Corpus (PTCC) corpus is a corpus of 149 articles written in Modern Standard Arabic and Tunisian Arabic. Tesseract was used to transform the constitution's pdf files into text files. Afterward, alignment of the parallel articles was achieved by a simple Python script. More details can be found in: amr-keleg.github.io Tunisian Arabic translation of the 2014 Tunisian Constitution

Visit

huggingface.co

Tasks

machine translation

Languages

Arabic, Tunisian Spoken

Tags

legal

Licenses

mit

Similaires

AMR TriageAMR ChallengeAMR Sentinel — Vivli 2026 AMR Data Challenge Pre-RegistrationAMR Surveillance DatasetMR-AMR Datasetkbastug/nanopore-amr-nigeria

AMR Triage

Background With high burden of antimicrobial resistance and steep increase in watch antibiotics in l

AMR Challenge

Antimicrobial resistance is a global problem that has so far resisted modern therapeutic interventio

AMR Sentinel — Vivli 2026 AMR Data Challenge Pre-Registration

Pre-registered secondary analysis linking Vivli AMR Surveillance Register (SPIDAAR, ATLAS, SMART) wi

AMR Surveillance Dataset

A synthetic tabular dataset for antimicrobial resistance surveillance in African healthcare faciliti

MR-AMR Dataset

MR-AMR is a public dataset composed of 140000 images of the digits in full-state and mid-state. It c

kbastug/nanopore-amr-nigeria

Bioinformatics workflows supporting nanopore-based bacterial identification and AMR gene detection i