Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Pretraining Corpus Domain and Hybrid Multilingual Model Performance in Low-Resource Universal Dependency Parsing

Domain:

natural language processing

Record type:

paper
Creator:
Ass
Publisher:
Zenodo
Host:avatar
Pretrained multilingual language models have become a common tool in transferring NLP capabilities to low-resource languages, often with adaptations. In this work, we study the performance, extensibility, and interaction of two such adaptations: vocabulary augmentation and script transliteration. Our evaluations on part-of-speech tagging, universal dependency parsing, and named entity recognition in nine diverse low-resource languages uphold the viability of these approaches while raising new questions around how to optimally adapt multilingual models to low-resource settings. Research goal: To what extent does the choice of pretraining corpus (domain-specific vs. general) influence the performance of hybrid multilingual models in universal dependency parsing for low-resource languages, as measured by labeled attachment score (LAS)? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 7.7/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 7.7/10.

Visit

doi.orgzenodo.org

Tasks

dependency parsinginformation extractionnamed entity recognitionparsingpart of speech tagging

Tags

extentchoicepretrainingcorpusdomain-specificgeneralinfluenceperformance

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similar

Multilingual Dependency Parsing for Low-Resource African Languages: Case Studies on Bambara, Wolof, and YorubaDomain-adaptive pretraining and hybrid batch training for robust zero-shot retrieval in low-resource languagesA Little Pretraining Goes a Long Way: A Case Study on Dependency Parsing Task for Low-resource Morphologically Rich LanguagesUDapter: Language Adaptation for Truly Universal Dependency ParsingDependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource LanguagesTypological Features for Multilingual Delexicalised Dependency Parsing

Multilingual Dependency Parsing for Low-Resource African Languages: Case Studies on Bambara, Wolof, and Yoruba

Domain-adaptive pretraining and hybrid batch training for robust zero-shot retrieval in low-resource languages

Information retrieval across different languages is an increasingly important challenge in natural l

A Little Pretraining Goes a Long Way: A Case Study on Dependency Parsing Task for Low-resource Morphologically Rich Languages

Neural dependency parsing has achieved remarkable performance for many domains and languages. The bo

UDapter: Language Adaptation for Truly Universal Dependency Parsing

Recent advances in multilingual dependency parsing have brought the idea of a truly universal parser

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages

Transformer-based models achieve state-of-the-art dependency parsing for high-resource languages, ye

Typological Features for Multilingual Delexicalised Dependency Parsing

International audience The existence of universal models to describe the syntax of la