Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

DOMAIN AND LANGUAGE ADAPTATION USING HETEROGENEOUS DATASETS FOR WAV2VEC2.0-BASED SPEECH RECOGNITION OF LOW-RESOURCE LANGUAGE

Domain:

natural language processing

Record type:

paper
Creator:
Kak
Publisher:
IEEE
Host:avatar
IEEE ICASSP 2023 Conference, Hybrid Event, 4-10 June 2023, Rhodes Island, Greece We address the effective finetuning of a large-scale pre-trained model for automatic speech recognition of low-resource languages with only a one-hour matched dataset. The finetuning is composed of domain adaptation and language adaptation, and they are conducted by using heterogeneous datasets, which are matched with either domain or language. For effective adaptation, we incorporate auxiliary tasks of domain identification and language identification with multi-task learning. Moreover, the embedding result of the auxiliary tasks is fused to the encoder output of the pre-trained model for ASR. Experimental evaluations on the Khmer ASR using the corpus of ECCC (the Extraordinary Chambers in the Courts of Cambodia) demonstrates that first conducting domain adaption and then language adaption is effective. In addition, multi-tasking with domain embedding gives the best performance, which reduces the baseline CER by 6.47%.

Visit

doi.orgrc.signalprocessingsociety.org

Tasks

automatic speech recognitionspeech processingtransfer learning

Similar

SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech RecognitionRapid Language Adaptation for Multilingual E2E Speech Recognition Using Encoder PromptingDonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech RecognitionFake News Classification in Urdu: A Domain Adaptation Approach for a Low-Resource LanguageFine-Tuning SLAM-ASR for Low-Resource Language Speech Recognition with High-Resource AlignmentDialect recognition for low resource language using an adaptive filter bank

SMILE: Speech Meta In-Context Learning for Low-Resource Language Automatic Speech Recognition

Automatic Speech Recognition (ASR) models demonstrate outstanding performance on high-resource langu

Rapid Language Adaptation for Multilingual E2E Speech Recognition Using Encoder Prompting

End-to-end multilingual speech recognition models handle multiple languages through a single model,

DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech Recognition

Low-resource automatic speech recognition (ASR) commonly relies on cross-lingual transfer, where mod

Fake News Classification in Urdu: A Domain Adaptation Approach for a Low-Resource Language

Misinformation on social media is a widely acknowledged issue, and researchers worldwide are activel

Fine-Tuning SLAM-ASR for Low-Resource Language Speech Recognition with High-Resource Alignment

Large language models (LLMs) have demonstrated potential in handling spoken inputs for high-resource

Dialect recognition for low resource language using an adaptive filter bank

Dialect recognition of low resource languages is the next stage in the technological advancement in