Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Dynamic Aggregation and Augmentation for Low-Resource Machine Translation using Federated Fine-tuning of Pretrained Transformer Models

Domain:

natural language processing

Record type:

datasetmodel
Creator:
EmmZhaAmaVic
Publisher:
MDP
Host:
Machine Translation (MT) for low-resource languages, such as Twi, remains a persistent challenge in Natural Language Processing (NLP) due to the scarcity of extensive parallel datasets. Due to their heavy reliance on high-resource data, traditional methods frequently fall short, underserving low-resource languages. To address this, we propose a fine-tuned T5 model trained using Cross-Lingual Optimization Framework (CLOF), a unique method that dynamically modifies gradient weights to balance low-resource (Twi) and high-resource (English) datasets. This cross-lingual learning framework leverages the strengths of federated training to improve translation performance while ensuring scalability for other low-resource languages. In order to maximize model input, the study makes use of a carefully selected parallel English-Twi corpus that has been aligned and tokenized. A thorough evaluation of translation quality is provided by the use of SPBLEU, ROUGE (ROUGE-1, ROUGE-2, and ROUGE-L) measures, and Word Error Rate (WER) metrics. A pretrained mT5 model is used to set baseline performance, which acts as a standard for the optimized model. The suggested method shows notable benefits, according to experimental results. The fine-tuned model achieves a remarkable increase in SPBLEU from 2.16% to 71.30%, a rise in ROUGE-1 from 15.23% to 65.24%, and a notable reduction in WER from 183.16% to 68.32%. These findings highlight the effectiveness of CLOF in addressing the challenges of low-resource MT and enhancing the quality of Twi translations. This work demonstrates the potential of combining cross-lingual learning and federated training to advance NLP for underrepresented languages, paving the way for more inclusive and scalable translation systems.

Visit

doi.org

Tasks

machine translationtransfer learning

Languages

AkanBwamu, CwiDinka, SoutheasternTwi

Licenses

http://creativecommons.org/licenses/by/4.0

Similar

Hausa Fake News Detection Using Lightweight Transformer Models with Adaptive Fine-Tuning in Low-Resource SettingsData Augmentation for Low-Resource Neural Machine TranslationBoosting English-Amharic machine translation using corpus augmentation and TransformerFine-Tuning LLMs for Low-Resource Dialect Translation: The Case of LebaneseKrishnateja244/Fine-tuning-of-ASR-models-on-low-resource-languagesA Diverse Data Augmentation Strategy for Low-Resource Neural Machine Translation

Hausa Fake News Detection Using Lightweight Transformer Models with Adaptive Fine-Tuning in Low-Resource Settings

The rapid growth of digital communication technologies and online news platforms has enhanced global

Data Augmentation for Low-Resource Neural Machine Translation

The quality of a Neural Machine Translation system depends substantially on the availability of siza

Boosting English-Amharic machine translation using corpus augmentation and Transformer

The Transformer-based neural machine translation (NMT) model has been very successful in recent year

Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese

This paper examines the effectiveness of Large Language Models (LLMs) in translating the low-resourc

Krishnateja244/Fine-tuning-of-ASR-models-on-low-resource-languages

Fine tuning ASR models such as Wave2Vec2.0, Whisper, Nemo and MMS models on low-resource languages

A Diverse Data Augmentation Strategy for Low-Resource Neural Machine Translation

One important issue that affects the performance of neural machine translation is the scale of avail