Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Tunisian Text Data, TTD

Domain:

natural language processing

Record type:

dataset
Creator:
Azi
Host:
This dataset is a curated compilation of various Tunisian datasets, aimed at gathering as much Tunisian text data as possible in one place. It combines multiple sources of Tunisian language data, providing a rich resource for research, development of NLP models, and linguistic studies on Tunisian text. Dataset Description

Visit

huggingface.co

Licenses

cc-by-4.0

Similar

TunArTTS: Tunisian Arabic Text-To-Speech CorpusTunBERT: Pretrained Contextualized Text Representation for Tunisian DialectAutomatic diacritization of Tunisian dialect text using SMT modelYoruba Text DataTunisian Foreign Debt Dataiyedex-labs/tunisian-data

TunArTTS: Tunisian Arabic Text-To-Speech Corpus

TunBERT: Pretrained Contextualized Text Representation for Tunisian Dialect

Pretrained contextualized text representation models learn an effective representation of a natural

Automatic diacritization of Tunisian dialect text using SMT model

Yoruba Text Data

Yoruba Culture and Language Dataset for NLP Tasks

Tunisian Foreign Debt Data

Quarterly data of the Tunisian foreign debt and some other indicators (savings, exchange, current de

iyedex-labs/tunisian-data

Public repository containing Tunisian datasets for use in 3amali and other Tunisian AI projects # 🇹