Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Language independent and unsupervised acoustic models for speech recognition and keyword spotting

Domaine:

natural language processing

Type de record:

paper
Créateur:
KniGalRagRat
Éditeur:
ApoUni
Éditeur:
Apo
Hôte:avatar
Copyright © 2014 ISCA. Developing high-performance speech processing systems for low-resource languages is very challenging. One approach to address the lack of resources is to make use of data from multiple languages. A popular direction in recent years is to train a multi-language bottleneck DNN. Language dependent and/or multi-language (all training languages) Tandem acoustic models (AM) are then trained. This work considers a particular scenario where the target language is unseen in multi-language training and has limited language model training data, a limited lexicon, and acoustic training data without transcriptions. A zero acoustic resources case is first described where a multilanguage AM is directly applied, as a language independent AM (LIAM), to an unseen language. Secondly, in an unsupervised approach a LIAM is used to obtain hypotheses for the target language acoustic data transcriptions which are then used in training a language dependent AM. 3 languages from the IARPA Babel project are used for assessment: Vietnamese, Haitian Creole and Bengali. Performance of the zero acoustic resources system is found to be poor, with keyword spotting at best 60% of language dependent performance. Unsupervised language dependent training yields performance gains. For one language (Haitian Creole) the Babel target is achieved on the in-vocabulary data.

Visit

doi.orgwww.repository.cam.ac.uk

Tasks

automatic speech recognitionkeywordsspeech processing

Tags

speech recognitionlow resourcemultilingual

Similaires

Low-Resource Speech Recognition and Keyword-SpottingCombining tandem and hybrid systems for improved speech recognition and keyword spotting on low resource languagesSyllable-Based and Hybrid Acoustic Models for Amharic Speech RecognitionFirst automatic fongbe continuous speech recognition system: Development of acoustic models and language modelsA Deep Learning Framework for Arabic Continuous Speech Keyword Spotting in Low-Resource Settings Using Isolated-Word Keyword Spotting and Posterior Probability FunctionsLow-resource keyword spotting using contrastively trained transformer acoustic word embeddings

Low-Resource Speech Recognition and Keyword-Spotting

The IARPA Babel program ran from March 2012 to November 2016. The aim of the program was to develop

Combining tandem and hybrid systems for improved speech recognition and keyword spotting on low resource languages

Copyright © 2014 ISCA. In recent years there has been significant interest in Automatic Speech Recog

Syllable-Based and Hybrid Acoustic Models for Amharic Speech Recognition

International audience no abstract

First automatic fongbe continuous speech recognition system: Development of acoustic models and language models

This paper reports our efforts toward an ASR system for a new under-resourced language (Fongbe). The aim of this work is to build acoustic models and language models for continuous speech decoding in Fongbe. The problem encountered with Fongbe (an African language

A Deep Learning Framework for Arabic Continuous Speech Keyword Spotting in Low-Resource Settings Using Isolated-Word Keyword Spotting and Posterior Probability Functions

Continuous Speech Keyword Spotting (CSKWS) presents a challenging paradigm shift from isolated-word

Low-resource keyword spotting using contrastively trained transformer acoustic word embeddings

We introduce a new approach, the ContrastiveTransformer, that produces acoustic word embeddings (AWE