Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

An Elicitation-Matrix Approach to Pragmatic Context Modeling in Low-Resource Machine Translation

Domaine:

natural language processing

Type de record:

dataset
Créateur:
KweGodKevNee
Éditeur:
Uni
Hôte:
Pragmatic ambiguity poses a major challenge for machine translation in low-resource languages like Akan, where a single English phrase may represent multiple pragmatic contexts and vice versa. To address this gap, we develop an elicitation matrix capturing key social and situational factors and use it to create a pragmatics‑focused Akan--English dataset of 863 annotated pairs. We then evaluate whether large language models (LLMs) can infer pragmatic context and whether explicit pragmatic tags improve translation selection choices. Across two models, three prompting strategies, and three experimental settings, human‑annotated pragmatic tags consistently yield the highest accuracy, with the largest gains on expansive (many‑to‑one) mappings. Chain‑of‑thought prompting further boosts performance. These findings indicate that pragmatic conditioning---rather than model size---is the primary driver of improvement, and they suggest that future models will benefit from incorporating pragmatic information during training and inference.

Visit

doi.org

Tasks

machine translation

Languages

Akan

Licenses

https://creativecommons.org/licenses/by-nc/4.0

Similaires

Low-resource neural machine translation with morphological modelingBeyond Many-Shot Translation: Scaling In-Context Demonstrations For Low-Resource Machine TranslationAn Empirical Study of Many-Shot In-Context Learning for Machine Translation of Low-Resource LanguagesIn-Context Example Selection via Similarity Search Improves Low-Resource Machine TranslationMachine Translation in Low-Resource Languages by an Adversarial Neural NetworkThe Low-Resource Double Bind: An Empirical Study of Pruning for Low-Resource Machine Translation

Low-resource neural machine translation with morphological modeling

Morphological modeling in neural machine translation (NMT) is a promising approach to achieving open

Beyond Many-Shot Translation: Scaling In-Context Demonstrations For Low-Resource Machine Translation

Building machine translation (MT) systems for low-resource languages is notably difficult due to the

An Empirical Study of Many-Shot In-Context Learning for Machine Translation of Low-Resource Languages

In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks from a few examp

In-Context Example Selection via Similarity Search Improves Low-Resource Machine Translation

The ability of generative large language models (LLMs) to perform in-context learning has given rise

Machine Translation in Low-Resource Languages by an Adversarial Neural Network

Existing Sequence-to-Sequence (Seq2Seq) Neural Machine Translation (NMT) shows strong capability wit

The Low-Resource Double Bind: An Empirical Study of Pruning for Low-Resource Machine Translation

A “bigger is better” explosion in the number of parameters in deep neural networks has made it increasingly challenging to make state-of-the-art networks accessible in compute-restricted environments. Compression techniques have taken on renewed importance as a way