Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Using articulatory feature detectors in progressive networks for multilingual low-resource phone recognition

Domaine:

natural language processing

Type de record:

paper
Créateur:
MahMar
Éditeur:
Aco
Hôte:
Systems inspired by progressive neural networks, transferring information from end-to-end articulatory feature detectors to similarly structured phone recognizers, are described. These networks, connecting the corresponding recurrent layers of pre-trained feature detector stacks and newly introduced phone recognizer stacks, were trained on data from four Asian languages, with experiments testing the system on those languages and four African languages. Later adjustments of these networks include the use of contrastive predictive coding layers at the inputs to those networks' recurrent portions. Such adjustments allow for performance differences to be attributed to the presence or absence of individual feature detectors (for consonant place/manner and vowel height/backness). Some of these differences manifest after feature-level comparisons of recognizer outputs, as well as through considering variations and ablations in architecture and training setup. These differences encourage further exploration of methods to reduce errors with phones having specific articulatory features as well as further architectural modifications.

Visit

doi.org

Tasks

automatic speech recognitionspeech processing

Similaires

Adversarial Meta Sampling for Multilingual Low-Resource Speech RecognitionAdaptive Activation Network For Low Resource Multilingual Speech RecognitionMultilingual Representations for Low Resource Speech Recognition and Keyword SearchPhone inventory optimization for multilingual automatic speech recognitionTask-based Meta Focal Loss for Multilingual Low-resource Speech RecognitionCross-lingual Embedding Clustering for Hierarchical Softmax in Low-Resource Multilingual Speech Recognition

Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition

Low-resource automatic speech recognition (ASR) is challenging, as the low-resource target language

Adaptive Activation Network For Low Resource Multilingual Speech Recognition

Low resource automatic speech recognition (ASR) is a useful but thorny task, since deep learning ASR

Multilingual Representations for Low Resource Speech Recognition and Keyword Search

This paper examines the impact of multilingual (ML) acoustic representations on Automatic Speech Rec

Phone inventory optimization for multilingual automatic speech recognition

This paper describes a phone inventory optimization procedure for application in multilingual automa

Task-based Meta Focal Loss for Multilingual Low-resource Speech Recognition

Low-resource automatic speech recognition is a challenging task due to a lack of labeled training da

Cross-lingual Embedding Clustering for Hierarchical Softmax in Low-Resource Multilingual Speech Recognition

We present a novel approach centered on the decoding stage of Automatic Speech Recognition (ASR) tha