Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Multilingual Representations for Low Resource Speech Recognition and Keyword Search

Domaine:

natural language processing

Type de record:

paper
Créateur:
CuiKinRamSet
Éditeur:
ApoUni
Éditeur:
Ins
Hôte:avatar
This paper examines the impact of multilingual (ML) acoustic representations on Automatic Speech Recognition (ASR) and keyword search (KWS) for low resource languages in the context of the OpenKWS15 evaluation of the IARPA Babel program. The task is to develop Swahili ASR and KWS systems within two weeks using as little as 3 hours of transcribed data. Multilingual acoustic representations proved to be crucial for building these systems under strict time constraints. The paper discusses several key insights on how these representations are derived and used. First, we present a data sampling strategy that can speed up the training of multilingual representations without appreciable loss in ASR performance. Second, we show that fusion of diverse multilingual representations developed at different LORELEI sites yields substantial ASR and KWS gains. Speaker adaptation and data augmentation of these representations improves both ASR and KWS performance (up to 8.7% relative). Third, incorporating un-transcribed data through semi-supervised learning, improves WER and KWS performance. Finally, we show that these multilingual representations significantly improve ASR and KWS performance (relative 9% for WER and 5% for MTWV) even when forty hours of transcribed audio in the target language is available. Multilingual representations significantly contributed to the LORELEI KWS systems winning the OpenKWS 15 evaluation.

Visit

doi.orgwww.repository.cam.ac.uk

Tasks

automatic speech recognitionkeywordsspeech processing

Languages

Swahili

Tags

46 Information and Computing Sciences47 Language, Communication and Culture4704 Linguistics

Similaires

Low-Resource Speech Recognition and Keyword-SpottingMultilingual self-supervised speech representations improve the speech recognition of low-resource African languages with codeswitchingAdversarial Meta Sampling for Multilingual Low-Resource Speech RecognitionAdaptive Activation Network For Low Resource Multilingual Speech RecognitionCombining tandem and hybrid systems for improved speech recognition and keyword spotting on low resource languagesTask-based Meta Focal Loss for Multilingual Low-resource Speech Recognition

Low-Resource Speech Recognition and Keyword-Spotting

The IARPA Babel program ran from March 2012 to November 2016. The aim of the program was to develop

Multilingual self-supervised speech representations improve the speech recognition of low-resource African languages with codeswitching

While many speakers of low-resource languages regularly code-switch between their languages and othe

Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition

Low-resource automatic speech recognition (ASR) is challenging, as the low-resource target language

Adaptive Activation Network For Low Resource Multilingual Speech Recognition

Low resource automatic speech recognition (ASR) is a useful but thorny task, since deep learning ASR

Combining tandem and hybrid systems for improved speech recognition and keyword spotting on low resource languages

Copyright © 2014 ISCA. In recent years there has been significant interest in Automatic Speech Recog

Task-based Meta Focal Loss for Multilingual Low-resource Speech Recognition

Low-resource automatic speech recognition is a challenging task due to a lack of labeled training da