Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Open-Domain Response Generation in Low-Resource Settings using Self-Supervised Pre-Training of Warm-Started Transformers

Domaine:

natural language processing

Type de record:

paper
Créateur:
TarZahBasHaz
Éditeur:
Ass
Hôte:
Learning response generation models constitute the main component of building open-domain dialogue systems. However, training open-domain response generation models requires large amounts of labeled data and pre-trained language generation models that are often nonexistent for low-resource languages. In this article, we propose a framework for training open-domain response generation models in low-resource settings. We consider Dialectal Arabic (DA) as a working example. The framework starts by warm-starting a transformer-based encoder-decoder with pre-trained language model parameters. Next, the resultant encoder-decoder model is adapted to DA by employing self-supervised pre-training on large-scale unlabeled data in the desired dialect. Finally, the model is fine-tuned on a very small labeled dataset for open-domain response generation. The results show significant performance improvements on three spoken Arabic dialects after adopting the framework’s three stages, highlighted by higher BLEU and lower Perplexity scores compared with multiple baseline models. Specifically, our models are capable of generating fluent responses in multiple dialects with an average human-evaluated fluency score above 4. Our data is made publicly available.

Visit

doi.org

Tasks

language modelingnatural language generation

Licenses

https://www.acm.org/publications/policies/copyright_policy#Background

Similaires

Lacuna Reconstruction: Self-supervised Pre-training for Low-Resource Historical Document TranscriptionDomain Adaptation in Low-Resource Perso-Arabic ASR with XLSR-53 Pre-TrainingIntegrating Unsupervised Data Generation into Self-Supervised Neural Machine Translation for Low-Resource LanguagesComparing Self-Supervised Pre-Training and Semi-Supervised Training for Speech Recognition in Languages with Weak Language ModelsImproving Low-Resource Morphological Inflection via Self-Supervised ObjectivesDomain Specific Specialization in Low-Resource Settings: The Efficacy of Offline Response-Based Knowledge Distillation in Large Language Models

Lacuna Reconstruction: Self-supervised Pre-training for Low-Resource Historical Document Transcription

We present a self-supervised pre-training approach for learning rich visual language representations

Domain Adaptation in Low-Resource Perso-Arabic ASR with XLSR-53 Pre-Training

Self-supervised pre-training could effectively improve the performance of low-resource automatic spe

Integrating Unsupervised Data Generation into Self-Supervised Neural Machine Translation for Low-Resource Languages

For most language combinations, parallel data is either scarce or simply unavailable. To address thi

Comparing Self-Supervised Pre-Training and Semi-Supervised Training for Speech Recognition in Languages with Weak Language Models

International audience This paper investigates the potential of improving a hybrid au

Improving Low-Resource Morphological Inflection via Self-Supervised Objectives

Self-supervised objectives have driven major advances in NLP by leveraging large-scale unlabeled dat

Domain Specific Specialization in Low-Resource Settings: The Efficacy of Offline Response-Based Knowledge Distillation in Large Language Models

Large Language Models (LLMs) excel in general tasks but often struggle with hallucinations when hand