Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

TuniSpeech-AI/TuniSpeech-21h

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Tun
Hôte:
TuniSpeech-21h is a 21-hour speech corpus specifically designed for Tunisian Arabic (Derja). It was developed to address the underrepresentation of this dialect in the landscape of Automatic Speech Recognition (ASR). The dataset is compiled from social media (YouTube and Facebook) and broadcast materials, capturing a wide range of spontaneous speech and diverse linguistic characteristics. Feature

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

Arabic, Tunisian Spoken

Tags

tunisian-arabicderjaspeech-recognitiontunispeechasr

Licenses

cc-by-nc-sa-4.0

Similaires

TuniSpeech-AI/whisper-tunisian-dialect

TuniSpeech-AI/whisper-tunisian-dialect