Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

BUFFET: Benchmarking Large Language Models for Few-shot Cross-lingual Transfer

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
AssAsaBleGon
Éditeur:
Und
Hôte:avatar
Despite remarkable advancements in few-shot generalization in natural language processing, most models are developed and evaluated primarily in English. To establish a rigorous and equitable evaluation framework for few-shot cross-lingual transfer, we introduce a new benchmark, called BUFFET, which unifies 15 diverse tasks across 54 languages in a sequence-to-sequence format and provides a fixed set of few-shot examples and instructions. Using BUFFET, we perform thorough evaluations of ten state-of-the-art multilingual large language models with different transfer methods, namely in-context learning and fine-tuning. Our findings reveal significant room for improvement in few-shot in-context cross-lingual transfer. Strong multilingual pre-trained or instruction-tuned models such as BLOOM or ChatGPT often lag behind much smaller mT5-base models given the same number of few-shot samples, particularly in low-resource languages. Our analysis suggests avenues for future research in few-shot cross-lingual transfer.

Visit

doi.orgunderline.io

Tasks

language modelingtransfer learning

Tags

Computational LinguisticsNatural Language ProcessingArtificial Intelligence

Similaires

Few-Shot Cross-Lingual Transfer for Prompting Large Language Models in Low-Resource LanguagesVicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language DetectionCross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on RomanianVicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language Detection | VIDEOIntermediate-Task Training and Few-Shot Learning for Cross-Domain Robustness in Zero-Shot Cross-Lingual TransferZero-Shot Cross-Lingual Reranking with Large Language Models for Low-Resource Languages

Few-Shot Cross-Lingual Transfer for Prompting Large Language Models in Low-Resource Languages

Large pre-trained language models (PLMs) are at the forefront of advances in Natural Language Proces

Vicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language Detection

Cross-lingual transfer learning from high-resource to medium and low-resource languages has shown en

Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian

Relation extraction (RE) for low-resource languages is typically constrained by the lack of annotate

Vicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language Detection | VIDEO

Cross-lingual transfer learning from high-resource to medium and low-resource languages has shown en

Intermediate-Task Training and Few-Shot Learning for Cross-Domain Robustness in Zero-Shot Cross-Lingual Transfer

Pre-trained multilingual language encoders, such as multilingual BERT and XLM-R, show great potentia

Zero-Shot Cross-Lingual Reranking with Large Language Models for Low-Resource Languages

Large language models (LLMs) have shown impressive zero-shot capabilities in various document rerank