Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

NaomiMeseret/Fine-tuning-stt-amharic-whisper

Domain:

natural language processing

Record type:

project
Creator:
Nao
Host:
# Fine Tuning STT/TTS Models: Amharic Whisper Project This repository contains my AI Engineer task on speech processing. I focused on **speech-to-text (STT)** for **Amharic**, using the **Whisper** model family and a public **Common Voice** dataset. ## Project Summary - Task type: STT - Language: Amharic - Model family: Whisper - Main baseline: `openai/whisper-tiny` - Fine-tuned model: `openai/whisper-tiny` after a small Colab fine-tuning run - Dataset: `hadamard-2/common-voice-24-ethiopian-v2` - Main goal: compare a raw Whisper baseline with a minimally fine-tuned Whisper Tiny model on public Amharic speech ## Why I Chose This Setup I chose Whisper because it is open-source, multilingual, and widely used for speech recognition. I chose Amharic because the task encouraged under-resourced languages. I used Common Voice because it is public and provides audio-text pairs, which makes it suitable for speech-to-text experiments. I kept the notebook setup small enough to work in a Colab-style environment with limited compute. ## Repository Structure ```text Fine Tuning STT_TTS models/ ├── README.md ├── requirements.txt ├── .gitignore ├── notebooks/ │ └── amharic_whisper_common_voice_colab.ipynb │ ├── outputs/ │ ├── model_comparison.png │ └── sample_waveform.png │ └── data/ ├── raw/ │ └── README.md └── processed/ ├── README.md ├── predictions.csv └── summary.json ``` ## Deliverables This repository includes the three main deliverables requested in the task: 1. Code / Notebook `notebooks/amharic_whisper_common_voice_colab.ipynb` 2. PDF Report `reports/amharic_whisper_report.pdf` 3. GitHub Repository This project folder is organized to be uploaded as a public GitHub repository with the notebook, code, outputs, and report. ## What I Did - Loaded public Amharic speech clips from Common Voice - Ran a baseline Whisper model - Ran a minimally fine-tuned Whisper Tiny model - Compared both models on the same test samples - Saved metrics, predictions, and visua …

Visit

github.com

Languages

Amharic

Similar

JONAHKYAGABA/-Whisper-ASR-Fine-Tuning-for-Amharicsykeb21/Fine-Tuning-OpenAI-Whisper-for-Amharic-Speech-RecognitionWhispering in Amharic: Fine-tuning Whisper for Low-resource LanguageAbdelelta/Luganda-whisper-fine-tuningHasanovChE/Fine-Tuning-OpenAI-Whisper-modelaman3013/Fine-tuning-Amharic-NER

JONAHKYAGABA/-Whisper-ASR-Fine-Tuning-for-Amharic

# 🗣️ Whisper ASR Fine-Tuning for Amharic # 🗣️ Whisper ASR Fine-Tuning for Amharic This project imp

sykeb21/Fine-Tuning-OpenAI-Whisper-for-Amharic-Speech-Recognition

This repository contains a complete, memory-optimized pipeline for fine-tuning OpenAI’s Whisper auto

Whispering in Amharic: Fine-tuning Whisper for Low-resource Language

This work explores fine-tuning OpenAI's Whisper automatic speech recognition (ASR) model for Amharic

Abdelelta/Luganda-whisper-fine-tuning

HasanovChE/Fine-Tuning-OpenAI-Whisper-model

Azərbaycan dili üçün Avtomatik Nitq Tanıma (ASR) sisteminin qurulması və təkmilləşdirilməsi (fine-tu

aman3013/Fine-tuning-Amharic-NER

# Fine-tuning-Amharic-NER ## Overview EthioMart aims to become the primary hub for Telegram-based