Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

XLDA: Cross-Lingual Data Augmentation for Natural Language Inference and Question Answering

Domaine:

natural language processing

Type de record:

paper
Créateur:
SinMcCKesXio
Hôte:avatar
While natural language processing systems often focus on a single language, multilingual transfer learning has the potential to improve performance, especially for low-resource languages. We introduce XLDA, cross-lingual data augmentation, a method that replaces a segment of the input text with its translation in another language. XLDA enhances performance of all 14 tested languages of the cross-lingual natural language inference (XNLI) benchmark. With improvements of up to $4.8\%$, training with XLDA achieves state-of-the-art performance for Greek, Turkish, and Urdu. XLDA is in contrast to, and performs markedly better than, a more naive approach that aggregates examples in various languages in a way that each example is solely in one language. On the SQuAD question answering task, we see that XLDA provides a $1.0\%$ performance increase on the English evaluation set. Comprehensive experiments suggest that most languages are effective as cross-lingual augmentors, that XLDA is robust to a wide range of translation quality, and that XLDA is even more effective for randomly initialized models than for pretrained models.

Visit

arxiv.org

Tasks

natural language inferencequestion answering

Tags

Computation and LanguageArtificial IntelligenceMachine Learning

Similaires

Cross-lingual Natural Language InferenceNeural Network Models for Paraphrase Identification, Semantic Textual Similarity, Natural Language Inference, and Question AnsweringAbelAdissu/Cross-Lingual-Question-Answering-for-Amharic-Language-Using-Pretrained-LLMs-AfriQA: Cross-lingual Open-Retrieval Question Answering for African LanguagesUnsupervised Cross-Domain and Cross-Lingual Methods for Text Classification, Slot-Filling, and Question-AnsweringM2M-100 Zero-Shot Cross-Lingual Retrieval with Language-Family Data Augmentation

Cross-lingual Natural Language Inference

XNLI is a subset of a few thousand examples from MNLI which has been translated into a 14 different

Neural Network Models for Paraphrase Identification, Semantic Textual Similarity, Natural Language Inference, and Question Answering

In this paper, we analyze several neural network designs (and their variations) for sentence pair mo

AbelAdissu/Cross-Lingual-Question-Answering-for-Amharic-Language-Using-Pretrained-LLMs-

## **INTRODUCTION** 📖 Welcome to the Amharic Text Generation project, a journey into the realm of na

AfriQA: Cross-lingual Open-Retrieval Question Answering for African Languages

African languages have far less in-language content available digitally, making it challenging for q

Unsupervised Cross-Domain and Cross-Lingual Methods for Text Classification, Slot-Filling, and Question-Answering

Transfer learning has significantly revolutionized modern machine learning systems by instilling the

M2M-100 Zero-Shot Cross-Lingual Retrieval with Language-Family Data Augmentation

Information retrieval across different languages is an increasingly important challenge in natural l