Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

CSSL: Contrastive Self-Supervised Learning for Dependency Parsing on Relatively Free Word Ordered and Morphologically Rich Low Resource Languages

Domaine:

natural language processing

Type de record:

paper
Créateur:
RaySanKriGoy
Hôte:avatar
Neural dependency parsing has achieved remarkable performance for low resource morphologically rich languages. It has also been well-studied that morphologically rich languages exhibit relatively free word order. This prompts a fundamental investigation: Is there a way to enhance dependency parsing performance, making the model robust to word order variations utilizing the relatively free word order nature of morphologically rich languages? In this work, we examine the robustness of graph-based parsing architectures on 7 relatively free word order languages. We focus on scrutinizing essential modifications such as data augmentation and the removal of position encoding required to adapt these architectures accordingly. To this end, we propose a contrastive self-supervised learning method to make the model robust to word order variations. Furthermore, our proposed modification demonstrates a substantial average gain of 3.03/2.95 points in 7 relatively free word order languages, as measured by the UAS/LAS Score metric when compared to the best performing baseline. Accepted at EMNLP 2024 Main (Short), 9 pages, 3 figures, 4 Tables

Visit

arxiv.org

Tasks

dependency parsingparsing

Tags

Computation and Language

Similaires

A Little Pretraining Goes a Long Way: A Case Study on Dependency Parsing Task for Low-resource Morphologically Rich LanguagesIntermediate-Task Training on English Syntax for Zero-Shot Parsing in Morphologically Rich Low-Resource LanguagesDependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource LanguagesConLID: Supervised Contrastive Learning for Low-Resource Language IdentificationNon-Contrastive Self-Supervised Speech Representations vs. Wav2Vec 2.0 in Low-Resource LanguagesMultilingual Dependency Parsing for Low-Resource African Languages: Case Studies on Bambara, Wolof, and Yoruba

A Little Pretraining Goes a Long Way: A Case Study on Dependency Parsing Task for Low-resource Morphologically Rich Languages

Neural dependency parsing has achieved remarkable performance for many domains and languages. The bo

Intermediate-Task Training on English Syntax for Zero-Shot Parsing in Morphologically Rich Low-Resource Languages

Intermediate-task training---fine-tuning a pretrained model on an intermediate task before fine-tuni

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages

Transformer-based models achieve state-of-the-art dependency parsing for high-resource languages, ye

ConLID: Supervised Contrastive Learning for Low-Resource Language Identification

Language identification (LID) is a critical step in curating multilingual LLM pretraining corpora fr

Non-Contrastive Self-Supervised Speech Representations vs. Wav2Vec 2.0 in Low-Resource Languages

This report synthesises findings from 13 peer-reviewed papers addressing the following research ques

Multilingual Dependency Parsing for Low-Resource African Languages: Case Studies on Bambara, Wolof, and Yoruba