This repo contains diacritized transcription and wavs of the TunSwitch dataset.
This work builds on the existing TunSwitch dataset by providing diacritics for the original transcriptions.
If you use the diacritized transcripts, please cite these works:
@misc{talafha2025nadi2025multidialectalarabic,
title={NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task},