Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

MulukenSholaye/amharic_englist_translation_hg

Domaine:

natural language processing

Type de record:

software
Créateur:
Mul
Hôte:
comparing performance of huggingface transformer models for amharic to english machine translation # Amharic-English Machine Translation Comparator Project ## Overview This project is a machine translation application designed to compare the performance of various transformer-based models for translating text from Amharic to English. The application provides a user-friendly interface to input Amharic text, receive translations from multiple models simultaneously, and visualize their performance metrics. The goal is to identify the most effective and efficient transformer architecture for this specific language pair. # Key Features * Multi-Model Translation: Translate a single Amharic text snippet using several different transformer models. * Performance Comparison: View and compare key metrics for each translation, such as BLEU score, model latency, and other qualitative assessments. * Intuitive UI: A clean and responsive user interface for easy text input, model selection, and result viewing. * Extensible Architecture: Designed to easily integrate new transformer models for future comparisons.ModelsThe application is built to support a variety of transformer models. Initial models included in this project are:Base * Transformer Model: A standard, vanilla transformer architecture trained from scratch on the Amharic-English dataset.Pre-trained Transformer Model (e.g., mBART-50): * A multilingual pre-trained model fine-tuned for the Amharic-English language pair. * Distilled Transformer Model: A smaller, more efficient version of a larger transformer model, optimized for faster inference with minimal loss in accuracy.DatasetThe models are trained and evaluated on a custom-curated Amharic-English parallel corpus. The dataset consists of parallel sentences sourced from various domains to ensure a broad coverage of vocabulary and grammar. The dataset is split into training, validation, and test sets to facilitate robust model training and evaluation.Getting StartedFollow these instructions to get a copy of the project up and running on your local machine. …

Visit

github.com

Tasks

machine translation

Languages

Amharic

Tags

amharic-nlptransformertranslation

Licenses

Apache-2.0