Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Model Size Influence on Cross-Lingual Euphemism Detection Performance

Domain:

natural language processing
Creator:
Ass
Publisher:
Zenodo
Host:avatar
Euphemisms are culturally variable and often ambiguous, posing challenges for language models, especially in low-resource settings. This paper investigates how cross-lingual transfer via sequential fine-tuning affects euphemism detection across five languages: English, Spanish, Chinese, Turkish, and Yoruba. We compare sequential fine-tuning with monolingual and simultaneous fine-tuning using XLM-R and mBERT, analyzing how performance is shaped by language pairings, typological features, and pretraining coverage. Results show that sequential fine-tuning with a high-resource L1 improves L2 perfo Research goal: What is the impact of model size (e.g., XLM-R-base vs. XLM-R-large) on cross-lingual euphemism detection performance when using sequential fine-tuning, measured by F1 scores across languages with different typological features? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 8.9/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 8.9/10.

Visit

doi.org

Tasks

text classificationtransfer learning

Languages

Yoruba

Tags

impactmodelsizeXLM-R-baseXLM-R-largecross-lingualeuphemismdetection

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode