Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Acoustic modeling for under-resourced languages based on vectorial HMM-states representation using Subspace Gaussian Mixture Models

Domain:

natural language processing

Record type:

paper
Creator:
BouFerMatLin
Editor:
Lab
Publisher:
CCSD
Host:avatar
International audience This paper explores a novel method for context-dependent models in automatic speech recognition (ASR), in the context of under-resourced languages. We present a simple way to realize a tying states approach, based on a new vectorial representation of the HMM states. This vectorial representation is considered as a vector of a low number of parameters obtained by the Subspace Gaussian Mixture Models paradigm (SGMM). The proposed method does not require phonetic knowledge or a large amount of data, which represent the major problems of acoustic modeling for under-resourced languages. This paper shows how this representation can be obtained and used for tying states. Our experiments, applied on Vietnamese, show that this approach achieves a stable gain compared to the classical approach which is based on decision trees. Furthermore, this method appears to be portable to other languages, as shown in the preliminary study conducted on Berber.

Visit

hal.science

Tasks

automatic speech recognitionspeech processing

Languages

Berber

Tags

Subspace Gaussian Mixture Modelsstate-tyingHMM-state vector representationunder-resourced languagesIndex Terms— Acoustic Modelling[INFO]Computer Science [cs]

Similar

Indigenuous Vocabulary Reformulation for Continuousyorùbá Speech Recognition In M-Commerce Using Acoustic Nudging-Based Gaussian Mixture ModelNon-Parametric Bayesian Subspace Models for Acoustic Unit DiscoveryUsing different acoustic, lexical and language modeling units for ASR of an under-resourced language - AmharicSpeech recognition system based on deep neural network acoustic modeling for low resourced language-AmharicImproved ASR for Under-Resourced Languages Through Multi-Task Learning with Acoustic LandmarksDetermining classes of food items for health requirements and nutrition guidelines using Gaussian mixture models

Indigenuous Vocabulary Reformulation for Continuousyorùbá Speech Recognition In M-Commerce Using Acoustic Nudging-Based Gaussian Mixture Model

Abstract One of the current research areas is speech recognition by aiding in the recognit

Non-Parametric Bayesian Subspace Models for Acoustic Unit Discovery

This work investigates subspace non-parametric models for the task of learning a set of acoustic uni

Using different acoustic, lexical and language modeling units for ASR of an under-resourced language - Amharic

State-of-the-art large vocabulary continuous speech recognition systems use mostly phone based acoustic models (AMs) and word based lexical and language models. However, phone based AMs are not efficient in modeling long-term temporal dependencies and the use of wo

Speech recognition system based on deep neural network acoustic modeling for low resourced language-Amharic

Improved ASR for Under-Resourced Languages Through Multi-Task Learning with Acoustic Landmarks

Furui first demonstrated that the identity of both consonant and vowel can be perceived from the C-V

Determining classes of food items for health requirements and nutrition guidelines using Gaussian mixture models

Introduction The identification of classes of nutritionally similar food items is important for cre