Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
TagMaiKurSak
Hôte:avatar
Multilingual machine translation (MT) benchmarks play a central role in evaluating the capabilities of modern MT systems. Among them, the FLORES+ benchmark is widely used, offering English-to-many translation data for over 200 languages, curated with strict quality control protocols. However, we study data in four languages (Asante Twi, Japanese, Jinghpaw, and South Azerbaijani) and uncover critical shortcomings in the benchmark's suitability for truly multilingual evaluation. Human assessments reveal that many translations fall below the claimed 90% quality standard, and the annotators report that source sentences are often too domain-specific and culturally biased toward the English-speaking world. We further demonstrate that simple heuristics, such as copying named entities, can yield non-trivial BLEU scores, suggesting vulnerabilities in the evaluation protocol. Notably, we show that MT models trained on high-quality, naturalistic data perform poorly on FLORES+ while achieving significant gains on our domain-relevant evaluation set. Based on these findings, we advocate for multilingual MT benchmarks that use domain-general and culturally neutral source texts rely less on named entities, in order to better reflect real-world translation challenges. 13 pages, 7 tables, 2 figures. Accepted at EMNLP Main 2025. Code and data released at github.com

Visit

arxiv.org

Tasks

machine translation

Languages

AkanAsante

Tags

Computation and LanguageArtificial Intelligence

Similaires

No Language Left Behind: Scaling Human-Centered Machine TranslationNo Error Left Behind: Multilingual Grammatical Error Correction with Pre-trained Translation ModelsLow Resource Neural Machine Translation: A Benchmark for Five African LanguagesMMTAfrica: Multilingual Machine Translation for African LanguagesMMTAfrica: Multilingual Machine Translation for African LanguagesWhat is driving Malawi's reproductive health gains, and who is still being left behind?

No Language Left Behind: Scaling Human-Centered Machine Translation

Driven by the goal of eradicating language barriers on a global scale, machine translation has solidified itself as a key focus of artificial intelligence research today. However, such efforts have coalesced around a small subset of languages, leaving behind the va

No Error Left Behind: Multilingual Grammatical Error Correction with Pre-trained Translation Models

Grammatical Error Correction (GEC) enhances language proficiency and promotes effective communicatio

Low Resource Neural Machine Translation: A Benchmark for Five African Languages

Recent advents in Neural Machine Translation (NMT) have shown improvements in low-resource language (LRL) translation tasks. In this work, we benchmark NMT between English and five African LRL pairs (Swahili, Amharic, Tigrigna, Oromo, Somali [SATOS]). We collected

MMTAfrica: Multilingual Machine Translation for African Languages

In this paper, we focus on the task of multilingual machine translation for African languages and describe our contribution in the 2021 WMT Shared Task: Large-Scale Multilingual Machine Translation. We introduce MMTAfrica, the first many-to-many multilingual transl

MMTAfrica: Multilingual Machine Translation for African Languages

This repository contains the official implementation of the MMTAfrica paper

What is driving Malawi's reproductive health gains, and who is still being left behind?

This two-page policy brief summarises findings from a peer-reviewed study published in BMC Public He