Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Few-Shot Cross-Lingual Stance Detection with Sentiment-Based Pre-Training

Domain:

natural language processing

Record type:

paper
Creator:
HarAroNakAug
Host:avatar
The goal of stance detection is to determine the viewpoint expressed in a piece of text towards a target. These viewpoints or contexts are often expressed in many different languages depending on the user and the platform, which can be a local news outlet, a social media platform, a news forum, etc. Most research in stance detection, however, has been limited to working with a single language and on a few limited targets, with little work on cross-lingual stance detection. Moreover, non-English sources of labelled data are often scarce and present additional challenges. Recently, large multilingual language models have substantially improved the performance on many non-English tasks, especially such with limited numbers of examples. This highlights the importance of model pre-training and its ability to learn from few examples. In this paper, we present the most comprehensive study of cross-lingual stance detection to date: we experiment with 15 diverse datasets in 12 languages from 6 language families, and with 6 low-resource evaluation settings each. For our experiments, we build on pattern-exploiting training, proposing the addition of a novel label encoder to simplify the verbalisation procedure. We further propose sentiment-based generation of stance data for pre-training, which shows sizeable improvement of more than 6% F1 absolute in low-shot settings compared to several strong baselines. Accepted to AAAI 2022 (Preprint version)

Visit

arxiv.org

Tasks

text classificationtransfer learning

Tags

Computation and LanguageMachine Learning

Similar

Few-shot Cross-lingual Aspect-Based Sentiment Analysis with Sequence-to-Sequence ModelsMulti-Source Cross-Lingual Pre-Training for Few-Shot NER in Low-Resource LanguagesIntermediate-Task Training and Few-Shot Learning for Cross-Domain Robustness in Zero-Shot Cross-Lingual TransferVicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language DetectionFew-shot text-based emotion detectionSelf-training vs Few-shot Learning in Cross-lingual NER for Low-resource Languages

Few-shot Cross-lingual Aspect-Based Sentiment Analysis with Sequence-to-Sequence Models

Aspect-based sentiment analysis (ABSA) has received substantial attention in English, yet challenges

Multi-Source Cross-Lingual Pre-Training for Few-Shot NER in Low-Resource Languages

Multi-lingual language models (LM), such as mBERT, XLM-R, mT5, mBART, have been remarkably successfu

Intermediate-Task Training and Few-Shot Learning for Cross-Domain Robustness in Zero-Shot Cross-Lingual Transfer

Pre-trained multilingual language encoders, such as multilingual BERT and XLM-R, show great potentia

Vicinal Risk Minimization for Few-Shot Cross-lingual Transfer in Abusive Language Detection

Cross-lingual transfer learning from high-resource to medium and low-resource languages has shown en

Few-shot text-based emotion detection

This paper describes the approach of the Unibuc - NLP team in tackling the SemEval 2025 Workshop, Ta

Self-training vs Few-shot Learning in Cross-lingual NER for Low-resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident