Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Dialogue Pidgin Text Adaptation via Contrastive Fine-Tuning

Domain:

natural language processing

Record type:

paper
The surging demand for multilingual dialogue systems often requires a costly labeling process for each language addition. For low resource languages, human annotators are continuously tasked with the adaptation of resource-rich language utterances for each new domain. However, this prohibitive and impractical process can often be a bottleneck for low resource languages that are still without proper translation systems nor parallel corpus. In particular, it is difficult to obtain task-specific low resource language annotations for the English-derived creoles (e.g. Nigerian and Cameroonian Pidgin). To address this issue, we utilize the pretrained language models i.e. BART which has shown great potential in language generation/understanding – we propose to finetune the BART model to generate utterances in Pidgin by leveraging the proximity of the source and target languages, and utilizing positive and negative examples in contrastive training objectives. We collected and released the first parallel Pidgin-English conversation corpus in two dialogue domains and showed that this simple and effective technique is sufficient to yield impressive results for English-to-Pidgin generation, which are two closely-related languages.

Visit

openreview.net

Tasks

natural language generation

Languages

Pidgin, CameroonPidgin, Nigerian

Tags

africanlp2language generationdialogue

Similar

raphaelguideke/TTS-System-for-Fulfulde-via-Fine-Tuning-BERT Fine-tuning For Arabic Text SummarizationmEdIT: Multilingual Text Editing via Instruction TuningFine-Tuning Without Forgetting via Loss-Adaptive Learning RatesCompletely Modular Fine-tuning for Dynamic Language AdaptationAmangtt/Fine-Tuning-NER-Models-on-Amharic-Text

raphaelguideke/TTS-System-for-Fulfulde-via-Fine-Tuning-

Speech Synthesis for Low-Resource Languages: A TTS System for Fulfulde via Fine-Tuning of a Foundati

BERT Fine-tuning For Arabic Text Summarization

Fine-tuning a pretrained BERT model is the state of the art method for extractive/abstractive text summarization, in this paper we showcase how this fine-tuning method can be applied to the Arabic language to both construct the first documented model for abstractiv

mEdIT: Multilingual Text Editing via Instruction Tuning

We introduce mEdIT, a multi-lingual extension to CoEdIT -- the recent state-of-the-art text editing

Fine-Tuning Without Forgetting via Loss-Adaptive Learning Rates

Fine-tuning large language models on new data improves task performance but degrades capabilities le

Completely Modular Fine-tuning for Dynamic Language Adaptation

Multilingual Fine-tuning of Large Language Models (LLMs) has achieved great advancements in machine

Amangtt/Fine-Tuning-NER-Models-on-Amharic-Text

# Entity Extraction for Amharic E-commerce Telegram Channels using LLM Fine-Tuning This project fo