Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

AymanMansour/New-Lisan-Sudanese-TTS-Dataset

Domain:

natural language processing

Record type:

dataset
Creator:
Aym
Host:
Lisan Sudanese TTS Dataset A synthetic Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) dataset specifically for Sudanese Arabic. 1,878 high-quality sentences featuring 20 synthetic speakers (10 male, 10 female). Reconstructed from the Lisan-Sudanese Morphological Dataset (52K manually annotated social media tokens from Facebook/X). Only sentences with a diacritic density of >=25% were kept to ensure enough phonetic information for accurate synthesis. model:

Visit

huggingface.co

Tasks

text to speechspeech processing

Languages

Arabic, Sudanese Spoken

Tags

sudanese-arabicdialectal-arabicsynthetic-audioresemble-aidiacritizedsocial-media

Similar

Lisan: Yemeni, Iraqi, Libyan, and Sudanese Arabic Dialect Copora with Morphological AnnotationsLisan: Yemeni, Iraqi, Libyan, and Sudanese Arabic Dialect Corpora with Morphological AnnotationsHausa TTS DatasetHausa TTS DatasetHausa TTS DatasetHausa TTS Dataset

Lisan: Yemeni, Iraqi, Libyan, and Sudanese Arabic Dialect Copora with Morphological Annotations

This article presents morphologically-annotated Yemeni, Sudanese, Iraqi, and Libyan Arabic dialects

Lisan: Yemeni, Iraqi, Libyan, and Sudanese Arabic Dialect Corpora with Morphological Annotations

Hausa TTS Dataset

This dataset contains 1,283 Hausa language audio recordings with transcriptions for Text-to-Speech (

Hausa TTS Dataset

This dataset contains Hausa language text-to-speech (TTS) recordings from multiple speakers. It incl

Hausa TTS Dataset

This dataset contains Hausa language text-to-speech (TTS) recordings from multiple speakers. It incl

Hausa TTS Dataset

This dataset contains Hausa language text-to-speech (TTS) recordings from multiple speakers. It incl