# ๐ท๐ผ Parakeet Kinyarwanda TTS Fine-Tuning Pipeline
# ๐ท๐ผ Parakeet Kinyarwanda TTS Fine-Tuning Pipeline
This project implements a complete pipeline for fine-tuning the Parakeet CTC model (v0.6b) on the **Kinyarwanda TTS dataset** using NVIDIA NeMo and Hugging Face's ecosystem.
## ๐ Project Overview
- ๐ Loads and parses 5 hours of Kinyarwanda TTS data (CSV and audio)
- ๐ Computes dataset statistics (duration, WER, vocabulary, SNR)
- ๐ Generates train/val/test JSON manifests
- ๐ค Builds a Byte Pair Encoding (BPE) tokenizer
- ๐๏ธ Loads and adapts a pre-trained `parakeet-ctc-0.6b` ASR model
- ๐ง Fine-tunes the model using subword tokenization and frozen encoder
- ๐ Logs metrics and training progress to Weights & Biases (wandb)
- โ๏ธ Optionally pushes final `.nemo` model to the Hugging Face Hub
## ๐ File Structure