Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Saadnasirbashir/asr_data_pipline_for_hausa_language

Domain:

natural language processing

Record type:

software
Creator:
Saa
Host:
A toolkit for preprocessing and organizing speech datasets for ASR, TTS, and voice AI projects. # ASR Data Pipeline for Hausa Language This project provides a modular data preprocessing pipeline designed for Automatic Speech Recognition (ASR) tasks, specifically for Hausa language audio datasets. It prepares audio data and corresponding transcripts into a format suitable for training speech recognition models. ## Features - 🔊 Audio normalization and resampling - 📜 Transcript preparation and alignment - 📁 Dataset organization for ASR model training - 🗣️ Supports Hausa language datasets ## Folder Structure

Visit

github.com

Tasks

automatic speech recognitionspeech processing

Languages

Hausa