Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Correlation between SWIM-IR Training Sample Volume and Zero-Shot Recall@50 Performance in Low-Resource Cross-Lingual Retrieval

Domaine:

natural language processing
Créateur:
Ass
Éditeur:
Zenodo
Hôte:avatar
Benefiting from transformer-based pre-trained language models, neural ranking models have made significant progress. More recently, the advent of multilingual pre-trained language models provides great support for designing neural cross-lingual retrieval models. However, due to unbalanced pre-training data in different languages, multilingual language models have already shown a performance gap between high and low-resource languages in many downstream tasks. And cross-lingual retrieval models built on such pre-trained models can inherit language bias, leading to suboptimal result for low-reso Research goal: What is the correlation between the volume of SWIM-IR training samples and zero-shot Recall@50 performance on very-low-resource languages in cross-lingual retrieval benchmarks? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.2/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 9.2/10.

Visit

doi.orgzenodo.org

Tasks

information retrieval

Tags

correlationvolumeSWIM-IRtrainingsampleszero-shotRecallperformance

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode