Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Performance of Zero-Shot Cross-Lingual Retrieval Models on Artificially Code-Switched Data vs. Multilingual Datasets in

Domaine:

natural language processing

Type de record:

paper
Créateur:
Ass
Éditeur:
Zenodo
Hôte:avatar
Benefiting from transformer-based pre-trained language models, neural ranking models have made significant progress. More recently, the advent of multilingual pre-trained language models provides great support for designing neural cross-lingual retrieval models. However, due to unbalanced pre-training data in different languages, multilingual language models have already shown a performance gap between high and low-resource languages in many downstream tasks. And cross-lingual retrieval models built on such pre-trained models can inherit language bias, leading to suboptimal result for low-reso Research goal: How does the performance of zero-shot cross-lingual retrieval models trained on artificially code-switched data compare to models fine-tuned on multilingual datasets like mC4 or OSCAR across a broader range of low-resource languages in XNLI? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.3/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 9.3/10.

Visit

doi.orgzenodo.org

Tasks

information retrieval

Tags

performancezero-shotcross-lingualretrievalmodelstrainedartificiallycode-switched

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Performance of Zero-Shot Cross-Lingual Retrieval Models with Artificially Code-Switched Training DataPerformance of Artificially Code-Switched Models in Zero-Shot Cross-Lingual Retrieval Across Resource LevelsCross-Lingual Embeddings for Zero-Shot Retrieval on Artificially Code-Switched Low-Resource DataCode-switched Training vs Multilingual Fine-tuning for Zero-shot Cross-lingual RetrievalArtificially Code-Switched Training for Zero-Shot Cross-Lingual Retrieval in Low-Resource LanguagesScaling Behavior of Zero-Shot Cross-Lingual Retrieval on Artificially Code-Switched Versus Native Data Across Low-Resource

Performance of Zero-Shot Cross-Lingual Retrieval Models with Artificially Code-Switched Training Data

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Performance of Artificially Code-Switched Models in Zero-Shot Cross-Lingual Retrieval Across Resource Levels

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Cross-Lingual Embeddings for Zero-Shot Retrieval on Artificially Code-Switched Low-Resource Data

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Code-switched Training vs Multilingual Fine-tuning for Zero-shot Cross-lingual Retrieval

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Artificially Code-Switched Training for Zero-Shot Cross-Lingual Retrieval in Low-Resource Languages

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Scaling Behavior of Zero-Shot Cross-Lingual Retrieval on Artificially Code-Switched Versus Native Data Across Low-Resource

Transferring information retrieval (IR) models from a high-resource language (typically English) to