Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Hybrid Multilingual Models with Vocabulary Augmentation and Script Transliteration for Low-Resource POS Tagging

Domaine:

natural language processing

Type de record:

paper
Créateur:
Ass
Éditeur:
Zenodo
Hôte:avatar
Pretrained multilingual language models have become a common tool in transferring NLP capabilities to low-resource languages, often with adaptations. In this work, we study the performance, extensibility, and interaction of two such adaptations: vocabulary augmentation and script transliteration. Our evaluations on part-of-speech tagging, universal dependency parsing, and named entity recognition in nine diverse low-resource languages uphold the viability of these approaches while raising new questions around how to optimally adapt multilingual models to low-resource settings. Research goal: How does the combination of vocabulary augmentation and script transliteration impact the cross-domain robustness of hybrid multilingual models when evaluated on low-resource language part-of-speech tagging across different text domains (e.g., social media, legal, medical)? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 7.6/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 7.6/10.

Visit

doi.orgzenodo.org

Tasks

part of speech tagging

Tags

combinationvocabularyaugmentationscripttransliterationimpactcross-domainrobustness

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Synergistic Effects of Vocabulary Augmentation and Script Transliteration on POS Tagging Accuracy in Low-Resource LanguagesImpact of Vocabulary Augmentation and Script Transliteration on Cross-Lingual Dependency Parsing in Low-Resource LanguagesImproving Data Augmentation for Low-Resource NMT Guided by POS-Tagging and Paraphrase Embeddingdhgarrette/low-resource-pos-tagging-2013Script Transliteration Effects on Multilingual NER Robustness in Low-Resource Non-Latin Languages

Synergistic Effects of Vocabulary Augmentation and Script Transliteration on POS Tagging Accuracy in Low-Resource Languages

Pretrained multilingual language models have become a common tool in transferring NLP capabilities t

Impact of Vocabulary Augmentation and Script Transliteration on Cross-Lingual Dependency Parsing in Low-Resource Languages

Pretrained multilingual language models have become a common tool in transferring NLP capabilities t

Improving Data Augmentation for Low-Resource NMT Guided by POS-Tagging and Paraphrase Embedding

Data augmentation is an approach for several text generation tasks. Generally, in the machine transl

dhgarrette/low-resource-pos-tagging-2013

# ANNOUNCEMENT: New Version Available *The code here has been completely rewritten to be significan

Script Transliteration Effects on Multilingual NER Robustness in Low-Resource Non-Latin Languages

Pretrained multilingual language models have become a common tool in transferring NLP capabilities t