Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Mitigating Translationese in Low-resource Languages: The Storyboard Approach

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
KuwUruAmuMuhammad, Shamsuddeen Hassan
Hôte:avatar
Low-resource languages often face challenges in acquiring high-quality language data due to the reliance on translation-based methods, which can introduce the translationese effect. This phenomenon results in translated sentences that lack fluency and naturalness in the target language. In this paper, we propose a novel approach for data collection by leveraging storyboards to elicit more fluent and natural sentences. Our method involves presenting native speakers with visual stimuli in the form of storyboards and collecting their descriptions without direct exposure to the source text. We conducted a comprehensive evaluation comparing our storyboard-based approach with traditional text translation-based methods in terms of accuracy and fluency. Human annotators and quantitative metrics were used to assess translation quality. The results indicate a preference for text translation in terms of accuracy, while our method demonstrates worse accuracy but better fluency in the language focused. published at LREC-COLING 2024

Visit

arxiv.org

Tags

Computation and LanguageI.2.7

Similaires

Mitigating Cross-Lingual Performance Degradation in Low-Resource Languages via Intermediate-Task Training on English ReasoningMitigating Annotation Projection Noise in Cross-Lingual NER via Source Dataset Scaling for Low-Resource LanguagesTowards a Crowdsourcing Platform for Low Resource Languages -- A Collectivist ApproachReusable Component Retrieval: A Semantic Search Approach for Low-Resource LanguagesAlligators All Around: Mitigating Lexical Confusion in Low-resource Machine TranslationBRIDGING GAPS IN LOW-RESOURCE LANGUAGES: A MACHINE LEARNING APPROACH TO PRONOMINAL ANAPHORA RESOLUTION

Mitigating Cross-Lingual Performance Degradation in Low-Resource Languages via Intermediate-Task Training on English Reasoning

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Mitigating Annotation Projection Noise in Cross-Lingual NER via Source Dataset Scaling for Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Towards a Crowdsourcing Platform for Low Resource Languages -- A Collectivist Approach

This work demonstrates how semi-supervised learning and human-in-the-loop crowdsourcing can help neu

Reusable Component Retrieval: A Semantic Search Approach for Low-Resource Languages

A common practice among programmers is to reuse existing code, accomplished by performing natural la

Alligators All Around: Mitigating Lexical Confusion in Low-resource Machine Translation

Current machine translation (MT) systems for low-resource languages have a particular failure mode:

BRIDGING GAPS IN LOW-RESOURCE LANGUAGES: A MACHINE LEARNING APPROACH TO PRONOMINAL ANAPHORA RESOLUTION

This study proposes the Kazakh Coreference Adaptation (KCA) model, a hybrid framework for resolving