Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Towards Inclusive ASR: Investigating Voice Conversion for Dysarthric Speech Recognition in Low-Resource Languages

Domain:

natural language processinghealthcare

Record type:

paper
Creator:
Li,YeoChoPér
Host:avatar
Automatic speech recognition (ASR) for dysarthric speech remains challenging due to data scarcity, particularly in non-English languages. To address this, we fine-tune a voice conversion model on English dysarthric speech (UASpeech) to encode both speaker characteristics and prosodic distortions, then apply it to convert healthy non-English speech (FLEURS) into non-English dysarthric-like speech. The generated data is then used to fine-tune a multilingual ASR model, Massively Multilingual Speech (MMS), for improved dysarthric speech recognition. Evaluation on PC-GITA (Spanish), EasyCall (Italian), and SSNCE (Tamil) demonstrates that VC with both speaker and prosody conversion significantly outperforms the off-the-shelf MMS performance and conventional augmentation techniques such as speed and tempo perturbation. Objective and subjective analyses of the generated data further confirm that the generated speech simulates dysarthric characteristics. 5 pages, 1 figure, Proceedings of Interspeech

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Computation and LanguageSoundAudio and Speech Processing

Similar

Voice Conversion Can Improve ASR in Very Low-Resource SettingsAutomatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature ReviewRobust speech recognition for low-resource languagesDEFI-COLaF/Speech-Recognition-for-Low-Resource-LanguagesEnhancing Automatic Speech Recognition for Child Speech in Low-Resource Languagessashakhaf/speech-recognition-for-3-low-resource-african-languages

Voice Conversion Can Improve ASR in Very Low-Resource Settings

Voice conversion (VC) could be used to improve speech recognition systems in low-resource languages

Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature Review

ASR has achieved remarkable global progress, yet African low-resource languages remain rigorously un

Robust speech recognition for low-resource languages

Process of human-machine interaction is an integral part of everyday human life in a modern world. T

DEFI-COLaF/Speech-Recognition-for-Low-Resource-Languages

# Speech-Recognition-for-Low-Resource-Languages This repository contains code to fine-tune Whisper

Enhancing Automatic Speech Recognition for Child Speech in Low-Resource Languages

Automatic speech recognition (ASR) for children is demanding because their speech differs c

sashakhaf/speech-recognition-for-3-low-resource-african-languages

# speech-recongition-for-3-low-resource-african-languages This project aims to build an automatic s