Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

abeladamushumet/Fluentian_STT_TTS_Project

Domain:

natural language processing

Record type:

modelsoftware
Creator:
abe
Host:
Amharic STT & TTS pipeline using Whisper-small and Coqui TTS # Fluentian STT/TTS Project **Fine-Tuning Speech-to-Text Model for Amharic Language** A comprehensive implementation of Speech-to-Text (STT) system using OpenAI Whisper, fine-tuned on Amharic speech data from the Leyu dataset. This project was developed as part of the Fluentian Internship Programme - Task Round 1. --- ## 📋 Table of Contents - Project Overview - Task Assignment - Goals - Model Selection - Dataset - Project Structure - Setup & Installation - Usage - Training Process - Results - Challenges & Solutions - Low-Compute Fine-Tuning - Deployment Considerations - Key Learnings - Author --- ## 🎯 Project Overview This project explores and experiments with Speech-to-Text (STT) systems by fine-tuning OpenAI's Whisper model on Amharic, an under-resourced language. The implementation demonstrates the complete ML pipeline from data preprocessing to model evaluation, with a focus on handling real-world challenges in low-resource language ASR systems. **Key Highlights:** - Fine-tuned Whisper-small model on 1000 Amharic speech samples - Achieved significant WER improvement: 1.4310 → 1.0300 - Handled Gojjam dialect variations for realistic ASR scenarios - Complete end-to-end pipeline with preprocessing, training, and evaluation --- ## 📝 Task Assignment **Fluentian Internship Programme - Task Round 1: AI Engineer Task** **Task Title:** Fine-Tuning STT/TTS Models **Objective:** Explore, experiment, and report on speech processing models (Speech-to-Text and/or Text-to-Speech) using open-source models and public datasets. **Requirements:** 1. Select at least one open-source STT or TTS model 2. Find a compatible public dataset (encouraged: under-resourced languages) 3. Run the model and perform minimal fine-tuning if feasible 4. Document the entire process with detailed analysis **Deliverables:** - Working code/notebook demonstrating STT/TTS - Comprehensive PDF report addressing all evaluation criteria - Public GitHub repository with code and documentation * …

Visit

github.com

Languages

Amharic

Licenses

Apache-2.0

Similar

abeladamushumet/begena_kiinit_ai_platformabeladamushumet/telegram_medical_pipelineabeladamushumet/Amharic-ASR-Summarizationabeladamushumet/Amharic-Ecommerce-Data-Extractor

abeladamushumet/begena_kiinit_ai_platform

AI system comparing ML and CNN models for Ethiopian Begena Kiñit classification with FastAPI backend

abeladamushumet/telegram_medical_pipeline

End‑to‑end medical data ETL pipeline extracts Amharic Telegram messages, preprocesses and cleans da

abeladamushumet/Amharic-ASR-Summarization

Amharic Speech-to-Text system powered by Wav2Vec 2.0 and Hugging Face Transformers, fine-tuned on th

abeladamushumet/Amharic-Ecommerce-Data-Extractor

Modular NLP pipeline that scrapes Amharic Telegram e-commerce posts, fine-tunes multilingual transfo