Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Investigating Massive Multilingual Pre-Trained Machine Translation Models for Clinical Domain via Transfer Learning

Domaine:

natural language processinghealthcare

Type de record:

papermodel
Créateur:
HanEroSorGla
Hôte:avatar
Massively multilingual pre-trained language models (MMPLMs) are developed in recent years demonstrating superpowers and the pre-knowledge they acquire for downstream tasks. This work investigates whether MMPLMs can be applied to clinical domain machine translation (MT) towards entirely unseen languages via transfer learning. We carry out an experimental investigation using Meta-AI's MMPLMs ``wmt21-dense-24-wide-en-X and X-en (WMT21fb)'' which were pre-trained on 7 language pairs and 14 translation directions including English to Czech, German, Hausa, Icelandic, Japanese, Russian, and Chinese, and the opposite direction. We fine-tune these MMPLMs towards English-\textit{Spanish} language pair which \textit{did not exist at all} in their original pre-trained corpora both implicitly and explicitly. We prepare carefully aligned \textit{clinical} domain data for this fine-tuning, which is different from their original mixed domain knowledge. Our experimental result shows that the fine-tuning is very successful using just 250k well-aligned in-domain EN-ES segments for three sub-task translation testings: clinical cases, clinical terms, and ontology concepts. It achieves very close evaluation scores to another MMPLM NLLB from Meta-AI, which included Spanish as a high-resource setting in the pre-training. To the best of our knowledge, this is the first work on using MMPLMs towards \textit{clinical domain transfer-learning NMT} successfully for totally unseen languages during pre-training. Accepted to ClinicalNLP-2023 WS@ACL-2023

Visit

arxiv.org

Tasks

machine translationtransfer learning

Languages

Hausa

Tags

Computation and LanguageArtificial Intelligence

Similaires

CLIPTrans: Transferring Visual Knowledge with Pre-trained Models for Multimodal Machine TranslationAdapting Pre-trained Language Models to African Languages via Multilingual Adaptive Fine-TuningNo Error Left Behind: Multilingual Grammatical Error Correction with Pre-trained Translation ModelsImpact of Pre-trained Multilingual Language Models on Zero-shot Cross-lingual NER Transfer PerformanceEvaluating Large Language Models for Low-Resource Multilingual Machine Translation in the Medical DomainFrom N-grams to Pre-trained Multilingual Models For Language Identification

CLIPTrans: Transferring Visual Knowledge with Pre-trained Models for Multimodal Machine Translation

There has been a growing interest in developing multimodal machine translation (MMT) systems that en

Adapting Pre-trained Language Models to African Languages via Multilingual Adaptive Fine-Tuning

Multilingual pre-trained language models (PLMs) have demonstrated impressive performance on several downstream tasks for both high-resourced and low-resourced languages. However, there is still a large performance drop for languages unseen during pre-training, espe

No Error Left Behind: Multilingual Grammatical Error Correction with Pre-trained Translation Models

Grammatical Error Correction (GEC) enhances language proficiency and promotes effective communicatio

Impact of Pre-trained Multilingual Language Models on Zero-shot Cross-lingual NER Transfer Performance

Multi-lingual language models (LM), such as mBERT, XLM-R, mT5, mBART, have been remarkably successfu

Evaluating Large Language Models for Low-Resource Multilingual Machine Translation in the Medical Domain

This dissertation explores neural machine translation (NMT) in multilingual medical domain, with

From N-grams to Pre-trained Multilingual Models For Language Identification

In this paper, we investigate the use of N-gram models and Large Pre-trained Multilingual models for