Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

linagora/linto-dataset-text-ar-tn

Domain:

natural language processing

Record type:

dataset
Creator:
Lin
Host:
This is a collection of Tunisian dialect textual documents for Language Modeling. It was used to train the Linto ASR in Tunisian dialect with code-switching capabilities linagora/linto-asr-ar-tn-0.1. Dataset Summary Dataset composition Sources Data Table Data sources Content Types Languages and Dialects Example use (python) License Citations Dataset Summary

Visit

huggingface.co

Languages

Arabic, Tunisian Spoken

Licenses

apache-2.0

Similar

linagora/linto-dataset-audio-ar-tnlinagora/linto-dataset-audio-ar-tn-augmented

linagora/linto-dataset-audio-ar-tn

This is the first packaged version of the datasets used to train the Linto Tunisian dialect with cod

linagora/linto-dataset-audio-ar-tn-augmented

This is the augmented datasets used to train the Linto Tunisian dialect with code-switching STT lina