Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

CONCRETE: Improving Cross-lingual Fact-checking with Cross-lingual Retrieval

Domaine:

natural language processing

Type de record:

paperdatasetsoftware
Créateur:
HuaZhaJi,
Hôte:avatar
Fact-checking has gained increasing attention due to the widespread of falsified information. Most fact-checking approaches focus on claims made in English only due to the data scarcity issue in other languages. The lack of fact-checking datasets in low-resource languages calls for an effective cross-lingual transfer technique for fact-checking. Additionally, trustworthy information in different languages can be complementary and helpful in verifying facts. To this end, we present the first fact-checking framework augmented with cross-lingual retrieval that aggregates evidence retrieved from multiple languages through a cross-lingual retriever. Given the absence of cross-lingual information retrieval datasets with claim-like queries, we train the retriever with our proposed Cross-lingual Inverse Cloze Task (X-ICT), a self-supervised algorithm that creates training instances by translating the title of a passage. The goal for X-ICT is to learn cross-lingual retrieval in which the model learns to identify the passage corresponding to a given translated title. On the X-Fact dataset, our approach achieves 2.23% absolute F1 improvement in the zero-shot cross-lingual setup over prior systems. The source code and data are publicly available at github.com. Accepted by COLING 2022

Visit

arxiv.org

Tasks

information retrieval

Tags

Computation and Language

Similaires

Improving Low-Resource Cross-lingual Document Retrieval by Reranking with Deep Bilingual RepresentationsData Efficient Dense Cross-Lingual Information RetrievalImproving Cross-lingual Information Retrieval on Low-Resource Languages via Optimal Transport DistillationXF2T: Cross-lingual Fact-to-Text Generation for Low-Resource LanguagesImproving Generative Cross-lingual Aspect-Based Sentiment Analysis with Constrained DecodingM2M-100 Zero-Shot Cross-Lingual Retrieval with Language-Family Data Augmentation

Improving Low-Resource Cross-lingual Document Retrieval by Reranking with Deep Bilingual Representations

In this paper, we propose to boost low-resource cross-lingual document retrieval performance with de

Data Efficient Dense Cross-Lingual Information Retrieval

Cross-Lingual Information Retrieval (CIR) remains challenging due to limited annotated data and ling

Improving Cross-lingual Information Retrieval on Low-Resource Languages via Optimal Transport Distillation

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

XF2T: Cross-lingual Fact-to-Text Generation for Low-Resource Languages

Multiple business scenarios require an automated generation of descriptive human-readable text from

Improving Generative Cross-lingual Aspect-Based Sentiment Analysis with Constrained Decoding

While aspect-based sentiment analysis (ABSA) has made substantial progress, challenges remain for lo

M2M-100 Zero-Shot Cross-Lingual Retrieval with Language-Family Data Augmentation

Information retrieval across different languages is an increasingly important challenge in natural l