Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Detecting Urgency Status of Crisis Tweets: A Transfer Learning Approach for Low Resource Languages

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
IntDiaKayMcK
Éditeur:
Und
Hôte:avatar
We release an urgency dataset that consists of English tweets relating to natural crises. The set is annotated along with annotations of their corresponding urgency status. Additionally, we release evaluation datasets for two low-resource languages, i.e. Sinhala and Odia, and demonstrate an effective zero-shot transfer from English to these two languages by training cross-lingual classifiers. We adopt cross-lingual embeddings constructed using different methods to extract features of the tweets, including a few state-of-the-art contextual embeddings such as BERT, RoBERTa and XLM-R. We train a variety of classifier architectures, supervised and semi supervised, on the extracted features. We also further experiment with ensembling the various classifiers. With very limited amounts of labeled data in English and zero data in the low resource languages, we show a successful framework of training monolingual and cross-lingual classifiers using deep learning methods which are known to be data hungry. Specifically, we show that the recent deep contextual embeddings are also helpful when dealing with very small-scale datasets. Classifiers that incorporate RoBERTa yield the best performance for the English urgency detection task, with 25% F1 score absolute improvement over the baselines. For the zero-shot transfer to low resource languages, classifiers that use LASER features perform the best for Sinhala transfer while XLM-R features benefit the Odia transfer the most.

Visit

doi.orgunderline.io

Tasks

text classificationtransfer learning

Tags

Computer and Information ScienceNatural Language ProcessingNeural Network

Similaires

Multilingual NLP for Low-Resource Languages Using Transfer LearningMulti-Round Transfer Learning for Low-Resource NMT Using Multiple High-Resource Languagesnossamchakri05/Cross-Lingual-Transfer-Learning-Based-Sentiment-Analysis-for-Low-Resource-LanguagesEnhancing Cross-Lingual Transfer through Reversible Transliteration: A Huffman-Based Approach for Low-Resource LanguagesFine-Tuning mBART-50 for Akkadian-to-English Translation: A Transfer Learning Approach for Low-ResourceModel Transfer for Tagging Low-resource Languages using a Bilingual Dictionary

Multilingual NLP for Low-Resource Languages Using Transfer Learning

Abstract: Despite the emergence of large-scale multilingual pre-trained models like mBERT, XLM-RoBER

Multi-Round Transfer Learning for Low-Resource NMT Using Multiple High-Resource Languages

Neural machine translation (NMT) has made remarkable progress in recent years, but the performance o

nossamchakri05/Cross-Lingual-Transfer-Learning-Based-Sentiment-Analysis-for-Low-Resource-Languages

# Cross-Lingual Transfer Learning-Based Sentiment Analysis for Low-Resource Languages ## 📋 Overview

Enhancing Cross-Lingual Transfer through Reversible Transliteration: A Huffman-Based Approach for Low-Resource Languages

As large language models (LLMs) are trained on increasingly diverse and extensive multilingual corpo

Fine-Tuning mBART-50 for Akkadian-to-English Translation: A Transfer Learning Approach for Low-Resource

Overview The research paper "Fine-Tuning mBART-50 for Akkadian-to-English Translation" by Frank Mor

Model Transfer for Tagging Low-resource Languages using a Bilingual Dictionary

Cross-lingual model transfer is a compelling and popular method for predicting annotations in a low-