Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Lattice-Based Unsupervised Test-Time Adaptation of Neural Network Acoustic Models

Domaine:

natural language processing

Type de record:

paper
Créateur:
KleFaiBelRen
Hôte:avatar
Acoustic model adaptation to unseen test recordings aims to reduce the mismatch between training and testing conditions. Most adaptation schemes for neural network models require the use of an initial one-best transcription for the test data, generated by an unadapted model, in order to estimate the adaptation transform. It has been found that adaptation methods using discriminative objective functions - such as cross-entropy loss - often require careful regularisation to avoid over-fitting to errors in the one-best transcriptions. In this paper we solve this problem by performing discriminative adaptation using lattices obtained from a first pass decoding, an approach that can be readily integrated into the lattice-free maximum mutual information (LF-MMI) framework. We investigate this approach on three transcription tasks of varying difficulty: TED talks, multi-genre broadcast (MGB) and a low-resource language (Somali). We find that our proposed approach enables many more parameters to be adapted without over-fitting being observed, and is successful even when the initial transcription has a WER in excess of 50%.

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Languages

Somali

Tags

Computation and LanguageSoundAudio and Speech Processing

Similaires

Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic ModelsAdvanced Convolutional Neural Network-Based Hybrid Acoustic Models for Low-Resource Speech RecognitionSimplified Neural Network-Based Models for Oil Flow Rate PredictionLanguage independent and unsupervised acoustic models for speech recognition and keyword spottingSpeech recognition system based on deep neural network acoustic modeling for low resourced language-AmharicBiomedical Event Trigger Identification Using Bidirectional Recurrent Neural Network Based Models

Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models

In this work, we propose lattice-free MMI (LFMMI) for supervised adaptation of self-supervised pretr

Advanced Convolutional Neural Network-Based Hybrid Acoustic Models for Low-Resource Speech Recognition

Deep neural networks (DNNs) have shown a great achievement in acoustic modeling for speech recogniti

Simplified Neural Network-Based Models for Oil Flow Rate Prediction

Available neural network-based models for predicting the oil flow rate (q<sub>o<

Language independent and unsupervised acoustic models for speech recognition and keyword spotting

Copyright © 2014 ISCA. Developing high-performance speech processing systems for low-resource langua

Speech recognition system based on deep neural network acoustic modeling for low resourced language-Amharic

Biomedical Event Trigger Identification Using Bidirectional Recurrent Neural Network Based Models

Biomedical events describe complex interactions between various biomedical entities. Event trigger i