Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Arabic Speech Corpus

Domain:

natural language processing

Record type:

dataset
Creator:
tun
Host:
This Speech corpus has been developed as part of PhD work carried out by Nawar Halabi at the University of Southampton. The corpus was recorded in south Levantine Arabic (Damascian accent) using a professional studio. Synthesized speech as an output using this corpus has produced a high quality, natural voice. [Needs More Information] Languages

Visit

huggingface.co

Tasks

text to speechspeech processing

Licenses

cc-by-4.0

Similar

TunArTTS: Tunisian Arabic Text-To-Speech CorpusDarijaVoice-Dysarthria: A Moroccan Arabic Dysarthric Speech CorpusOMAN-SPEECH: A Multi-Layer Annotated Speech Corpus for Omani Arabic DialectsA Corpus and Phonetic Dictionary for Tunisian Arabic Speech RecognitionA New Tunisian Arabic Corpus and Benchmark for Automatic Speech RecognitionTuniFra: A Tunisian Arabic Speech Corpus with Orthographic Transcriptions and French Translations

TunArTTS: Tunisian Arabic Text-To-Speech Corpus

DarijaVoice-Dysarthria: A Moroccan Arabic Dysarthric Speech Corpus

DarijaVoice-Dysarthria is the first dysarthric speech corpus for Moroccan Arabic (Darija). It contai

OMAN-SPEECH: A Multi-Layer Annotated Speech Corpus for Omani Arabic Dialects

Automatic Speech Recognition (ASR) has achieved strong performance in high-resource languages; howev

A Corpus and Phonetic Dictionary for Tunisian Arabic Speech Recognition

A New Tunisian Arabic Corpus and Benchmark for Automatic Speech Recognition

TuniFra: A Tunisian Arabic Speech Corpus with Orthographic Transcriptions and French Translations