Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Unsupervised Word Segmentation from Discrete Speech Units in Low-Resource Settings

Domaine:

natural language processing

Type de record:

paper
Créateur:
BoiYusOndVil
Hôte:avatar
Documenting languages helps to prevent the extinction of endangered dialects, many of which are otherwise expected to disappear by the end of the century. When documenting oral languages, unsupervised word segmentation (UWS) from speech is a useful, yet challenging, task. It consists in producing time-stamps for slicing utterances into smaller segments corresponding to words, being performed from phonetic transcriptions, or in the absence of these, from the output of unsupervised speech discretization models. These discretization models are trained using raw speech only, producing discrete speech units that can be applied for downstream (text-based) tasks. In this paper we compare five of these models: three Bayesian and two neural approaches, with regards to the exploitability of the produced units for UWS. For the UWS task, we experiment with two models, using as our target language the Mboshi (Bantu C25), an unwritten language from Congo-Brazzaville. Additionally, we report results for Finnish, Hungarian, Romanian and Russian in equally low-resource settings, using only 4 hours of speech. Our results suggest that neural models for speech discretization are difficult to exploit in our setting, and that it might be necessary to adapt them to limit sequence length. We obtain our best UWS results by using Bayesian models that produce high quality, yet compressed, discrete representations of the input speech signal. Accepted to SIGUL 2022

Visit

arxiv.org

Languages

Mbosi

Tags

Computation and LanguageSoundAudio and Speech Processing

Similaires

Unsupervised Word Segmentation from Speech with Attention Unsupervised Morphological Segmentation and Part-of-Speech Tagging for Low-Resource ScenariosMultilingual unsupervised sequence segmentation transfers to extremely low-resource languagesClinically-Informed Preprocessing Improves Stroke Segmentation in Low-Resource SettingsUnsupervised word segmentation for Sesotho using Adaptor GrammarsVisually grounded few-shot word learning in low-resource settings

Unsupervised Word Segmentation from Speech with Attention

We present a first attempt to perform attentional word segmentation directly from the speech signal,

Unsupervised Morphological Segmentation and Part-of-Speech Tagging for Low-Resource Scenarios

With the high cost of manually labeling data and the increasing interest in low-resource languages,

Multilingual unsupervised sequence segmentation transfers to extremely low-resource languages

We show that unsupervised sequence-segmentation performance can be transferred to extremely low-reso

Clinically-Informed Preprocessing Improves Stroke Segmentation in Low-Resource Settings

Stroke is among the top three causes of death worldwide, and accurate identification of ischemic str

Unsupervised word segmentation for Sesotho using Adaptor Grammars

Visually grounded few-shot word learning in low-resource settings

We propose a visually grounded speech model that learns new words and their visual depictions from j