Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

SSA-COMET: Do LLMs Outperform Learned Metrics in Evaluating MT for Under-Resourced African Languages?

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
Li,WanAliChe
Hôte:avatar
Evaluating machine translation (MT) quality for under-resourced African languages remains a significant challenge, as existing metrics often suffer from limited language coverage and poor performance in low-resource settings. While recent efforts, such as AfriCOMET, have addressed some of the issues, they are still constrained by small evaluation sets, a lack of publicly available training data tailored to African languages, and inconsistent performance in extremely low-resource scenarios. In this work, we introduce SSA-MTE, a large-scale human-annotated MT evaluation (MTE) dataset covering 14 African language pairs from the News domain, with over 73,000 sentence-level annotations from a diverse set of MT systems. Based on this data, we develop SSA-COMET and SSA-COMET-QE, improved reference-based and reference-free evaluation metrics. We also benchmark prompting-based approaches using state-of-the-art LLMs like GPT-4o, Claude-3.7 and Gemini 2.5 Pro. Our experimental results show that SSA-COMET models significantly outperform AfriCOMET and are competitive with the strongest LLM Gemini 2.5 Pro evaluated in our study, particularly on low-resource languages such as Twi, Luo, and Yoruba. All resources are released under open licenses to support future research.

Visit

arxiv.org

Tasks

machine translation

Languages

Yoruba

Tags

Computation and LanguageArtificial Intelligence

Similaires

AfriMTE and AfriCOMET: Enhancing COMET to Embrace Under-resourced African LanguagesExamining the Cultural Encoding of Gender Bias in LLMs for Low-Resourced African LanguagesStrategies for building wordnets for under-resourced languages: The case of African languagesLeveraging LLMs for MT in Crisis Scenarios: a blueprint for low-resource languagesLLM Probe: Evaluating LLMs for Low-Resource LanguagesDatasheets for Under-resourced Languages: An Example

AfriMTE and AfriCOMET: Enhancing COMET to Embrace Under-resourced African Languages

International audience Despite the recent progress on scaling multilingual machine tr

Examining the Cultural Encoding of Gender Bias in LLMs for Low-Resourced African Languages

Strategies for building wordnets for under-resourced languages: The case of African languages

The African Wordnet Project (AWN) aims at building wordnets for five African languages: Setswana, is

Leveraging LLMs for MT in Crisis Scenarios: a blueprint for low-resource languages

In an evolving landscape of crisis communication, the need for robust and adaptable Machine Translat

LLM Probe: Evaluating LLMs for Low-Resource Languages

Despite rapid advances in large language models (LLMs), their linguistic abilities in low-resource a

Datasheets for Under-resourced Languages: An Example

The datasheet provides an example of how to use the Datasheet standard for describing and sharing un