Logo Lanfrica

amkraneYoussef/TamazightTTS

Domaine:

natural language processing

Type de record:

modelsoftware
Créateur:
amk
Hôte:
FINE-TUNING SPEECHT5 FOR TAMAZIGHT TEXT-TO-SPEECH # Fine-Tuning SpeechT5 for Tamazight Text-to-Speech Companion Repository for: “Fine-Tuning SpeechT5 for Tamazight Text-to-Speech: A Foundation for Educational Technology in Low-Resource Languages” (Submitted to International Journal of Speech Technology) ## Overview This repository contains the resources, scripts, and configuration files used to fine-tune Microsoft’s SpeechT5 model for text-to-speech (TTS) synthesis in Tamazight, a low-resource language. The training pipeline uses Hugging Face Transformers and Datasets libraries, with detailed preprocessing, model training, and evaluation steps made publicly available. ## Quick Links - Preprocessed Training Set (HF-compatible): drive.google.com - Preprocessed Validation Set (HF-compatible): drive.google.com - Raw Audio Archive (with CSV files for train, val, and test): drive.google.com - Kaggle Preprocessing Notebook: kaggle.com - Training Pipeline (Colab): colab.research.google.com - Checkpoints + Logs + Best models: drive.google.com ## Dataset Sources This project was built using the following open-source speech datasets: | Dataset | Link | |-----------------------|---------------------------------------------------------------------| | Common Voice v13 | datacollective.mozillafound… | | tamazight_asr | TutlaytAI/tamazight_asr | | tamawalt-n-imZZyann | huggingface.co | …