Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Multilingual self-supervised speech representations improve the speech recognition of low-resource African languages with codeswitching

Domain:

natural language processing

Record type:

paper
Creator:
ÒgúManJur
Host:avatar
While many speakers of low-resource languages regularly code-switch between their languages and other regional languages or English, datasets of codeswitched speech are too small to train bespoke acoustic models from scratch or do language model rescoring. Here we propose finetuning self-supervised speech representations such as wav2vec 2.0 XLSR to recognize code-switched data. We find that finetuning self-supervised multilingual representations and augmenting them with n-gram language models trained from transcripts reduces absolute word error rates by up to 20% compared to baselines of hybrid models trained from scratch on code-switched data. Our findings suggest that in circumstances with limited training data finetuning self-supervised representations is a better performing and viable solution. 5 pages, 1 figure. Computational Approaches to Linguistic Code-Switching, CALCS 2023 (co-located with EMNLP 2023)

Visit

arxiv.org

Tasks

automatic speech recognitioncode switchingspeech processing

Tags

Computation and Language

Similar

Non-Contrastive Self-Supervised Speech Representations vs. Wav2Vec 2.0 in Low-Resource LanguagesMultilingual Representations for Low Resource Speech Recognition and Keyword SearchFine-Tuned Self-Supervised Speech Representations for Language Diarization in Multilingual Code-Switched SpeechSelf-supervised Speech Representations Still Struggle with African American Vernacular EnglishLarge vocabulary speech recognition for languages of Africa: multilingual modeling and self-supervised learningsashakhaf/speech-recognition-for-3-low-resource-african-languages

Non-Contrastive Self-Supervised Speech Representations vs. Wav2Vec 2.0 in Low-Resource Languages

This report synthesises findings from 13 peer-reviewed papers addressing the following research ques

Multilingual Representations for Low Resource Speech Recognition and Keyword Search

This paper examines the impact of multilingual (ML) acoustic representations on Automatic Speech Rec

Fine-Tuned Self-Supervised Speech Representations for Language Diarization in Multilingual Code-Switched Speech

Annotating a multilingual code-switched corpus is a painstaking process requiring specialist linguis

Self-supervised Speech Representations Still Struggle with African American Vernacular English

Underperformance of ASR systems for speakers of African American Vernacular English (AAVE) and other

Large vocabulary speech recognition for languages of Africa: multilingual modeling and self-supervised learning

Almost none of the 2,000+ languages spoken in Africa have widely available automatic speech recognition systems, and the required data is also only available for a few languages. We have experimented with two techniques which may provide pathways to large vocabular

sashakhaf/speech-recognition-for-3-low-resource-african-languages

# speech-recongition-for-3-low-resource-african-languages This project aims to build an automatic s