Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Synthetic Data Diversity vs. Back-Translation for Multilingual NER in Low-Resource Languages

Domaine:

natural language processing

Type de record:

paper
Créateur:
Ass
Éditeur:
Zenodo
Hôte:avatar
Named Entity Recognition(NER) for low-resource languages aims to produce robust systems for languages where there is limited labeled training data available, and has been an area of increasing interest within NLP. Data augmentation for increasing the amount of low-resource labeled data is a common practice. In this paper, we explore the role of synthetic data in the context of multilingual, low-resource NER, considering 11 languages from diverse language families. Our results suggest that synthetic data does in fact hold promise for low-resource language NER, though we see significant variatio Research goal: How does the use of synthetic data diversity compare to back-translation in improving the F1 score of multilingual NER models on low-resource target languages evaluated on the FLEx benchmark? Autonomous synthesis report generated by Assignee Research. Tribunal consensus score: 8.5/10. This report was generated autonomously by Assignee Research, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 8.5/10.

Visit

doi.org

Tasks

named entity recognitioninformation extraction

Tags

usesyntheticdatadiversityback-translationimprovingscoremultilingual

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Synthetic Data Diversity and Robustness in Teacher-Student NER Models for Low-Resource LanguagesProjection-based Cross-lingual NER vs. Zero-shot Multilingual Models in Low-resource LanguagesCross-lingual NER Performance via Annotation Projection vs. Multilingual Language Models in Low-Resource Languages13Aluminium/Synthetic-Multilingual-Code-Switch-Data-Generation-A-Pipeline-for-Low-Resource-LanguagesSynthetic Data and Annotation Projection for Low-Resource NER PerformanceMassively Multilingual Text Translation For Low-Resource Languages

Synthetic Data Diversity and Robustness in Teacher-Student NER Models for Low-Resource Languages

Named Entity Recognition(NER) for low-resource languages aims to produce robust systems for language

Projection-based Cross-lingual NER vs. Zero-shot Multilingual Models in Low-resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Cross-lingual NER Performance via Annotation Projection vs. Multilingual Language Models in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

13Aluminium/Synthetic-Multilingual-Code-Switch-Data-Generation-A-Pipeline-for-Low-Resource-Languages

# Synthetic-Multilingual-Code-Switch-Data-Generation-A-Pipeline-for-Low-Resource-Languages ## Resea

Synthetic Data and Annotation Projection for Low-Resource NER Performance

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Massively Multilingual Text Translation For Low-Resource Languages

Translation into severely low-resource languages has both the cultural goal of saving and reviving t