Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Scaling Multilingual Language Models and CLCA Score Improvements via Optimal Transport Distillation in MIRACL

Domaine:

natural language processing

Type de record:

paper
Créateur:
Ass
Éditeur:
Zenodo
Hôte:avatar
Benefiting from transformer-based pre-trained language models, neural ranking models have made significant progress. More recently, the advent of multilingual pre-trained language models provides great support for designing neural cross-lingual retrieval models. However, due to unbalanced pre-training data in different languages, multilingual language models have already shown a performance gap between high and low-resource languages in many downstream tasks. And cross-lingual retrieval models built on such pre-trained models can inherit language bias, leading to suboptimal result for low-reso Research goal: How does scaling the size of the multilingual language model affect the CLCA score improvements achieved by optimal transport distillation for low-resource languages in the MIRACL benchmark compared to high-resource languages? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 9.2/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 9.2/10.

Visit

doi.orgzenodo.org

Tasks

information retrieval

Tags

scalingsizemultilinguallanguagemodelaffectCLCAscore

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Scaling Multilingual Language Models and Zero-Shot R@1 Gaps via Optimal Transport Distillation on Flickr30k-EntitiesPerformance Variation in Multilingual Pre-trained Language Models with Optimal Transport Distillation for AdversarialScaling of Cross-Lingual Retrieval Performance with Model Size via Optimal Transport DistillationOptimal Transport Distillation for Cross-Lingual Alignment Robustness in MIRACLOptimal Transport Distillation for mBERT Performance Variance Reduction in MIRACL RetrievalOptimal Transport Distillation for Low-Resource Language Alignment in Multilingual Retrieval

Scaling Multilingual Language Models and Zero-Shot R@1 Gaps via Optimal Transport Distillation on Flickr30k-Entities

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Performance Variation in Multilingual Pre-trained Language Models with Optimal Transport Distillation for Adversarial

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Scaling of Cross-Lingual Retrieval Performance with Model Size via Optimal Transport Distillation

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Optimal Transport Distillation for Cross-Lingual Alignment Robustness in MIRACL

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Optimal Transport Distillation for mBERT Performance Variance Reduction in MIRACL Retrieval

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi

Optimal Transport Distillation for Low-Resource Language Alignment in Multilingual Retrieval

Benefiting from transformer-based pre-trained language models, neural ranking models have made signi