Logo Lanfrica

Mwungere/KinyaWisper

Domaine:

natural language processing

Type de record:

software
Créateur:
Mwu
Hôte:
A Kinyarwanda voice assistant # Kinyarwanda Voice Assistant 🤖 An intelligent voice assistant for Kinyarwanda language interaction, developed as part of the Intelligent Robotics course. ## Features 🌟 - 🎙️ **Kinyarwanda ASR** using KinyaWhisper (16kHz optimized) - 🧠 **Contextual Understanding** with fuzzy logic matching - 📢 **Natural Responses** with Kinyarwanda TTS - 🔇 **Noise Reduction** using advanced audio cleaning - 🎚️ **Voice Activity Detection** for precise speech recognition - 🔄 **Anti-Repetition** transcription filters - 📊 **Conversation Analytics** with matching insights - 🌐 **Web Interface** with Gradio integration ## Tech Stack 🛠️ - **Core AI**: Hugging Face Transformers - **Audio Processing**: Librosa + Soundfile - **NLP**: FuzzyWuzzy + Python-Levenshtein - **Interface**: Gradio - **Optimization**: WebRTC VAD + Noisereduce ## Installation 💻 ### Prerequisites - Python 3.12 - FFmpeg (audio processing): ```bash # Ubuntu/Debian sudo apt-get install ffmpeg # macOS brew install ffmpeg # Windows (via chocolatey) choco install ffmpeg ``` ## Quick Start 🚀 - Clone repository ```bash git clone github.com cd KinyaWisper ``` - Set up virtual environment ```bash python -m venv .venv source .venv/bin/activate # Linux/macOS .\.venv\Scripts\activate # Windows ``` - Install dependencies ```bash pip install -r requirements.txt ``` ## Configuration ⚙️ #### QA Configuration in `nlp_mapping.json` ```json { "qa_pairs": [ { "question": "Mwaramuce neza?", "answer": "Mwaramutse! Amakuru yanyu?" } ], "default_response": "Vugurura ikibazo." } ``` #### Audio Files You can find sample Kinyarwanda recordings in the `/sample_inputs` folder Supported formats: `WAV`, `MP3`, `OGG` ## Usage 🚀 #### Start the application ```bash python main.py ``` #### Access the interface - Navigate to localhost ## Interface Guide 💡 1. Record using your microphone or upload an audio file 1. Click **Submit** to process (⏳ ~10–60 sec) 1. Response audio auto-pla …