Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Investigations on Speech Recognition Systems for Low-Resource Dialectal Arabic-English Code-Switching Speech

Domain:

natural language processing

Record type:

paperdatasetmodel
Creator:
HamDenLi,Elm
Host:avatar
Code-switching (CS), defined as the mixing of languages in conversations, has become a worldwide phenomenon. The prevalence of CS has been recently met with a growing demand and interest to build CS ASR systems. In this paper, we present our work on code-switched Egyptian Arabic-English automatic speech recognition (ASR). We first contribute in filling the huge gap in resources by collecting, analyzing and publishing our spontaneous CS Egyptian Arabic-English speech corpus. We build our ASR systems using DNN-based hybrid and Transformer-based end-to-end models. In this paper, we present a thorough comparison between both approaches under the setting of a low-resource, orthographically unstandardized, and morphologically rich language pair. We show that while both systems give comparable overall recognition results, each system provides complementary sets of strength points. We show that recognition can be improved by combining the outputs of both systems. We propose several effective system combination approaches, where hypotheses of both systems are merged on sentence- and word-levels. Our approaches result in overall WER relative improvement of 4.7%, over a baseline performance of 32.1% WER. In the case of intra-sentential CS sentences, we achieve WER relative improvement of 4.8%. Our best performing system achieves 30.6% WER on ArzEn test set. To be published in Computer Speech and Language Journal

Visit

arxiv.org

Tasks

automatic speech recognitioncode switchingspeech processing

Tags

Computation and Language

Similar

Effects of Dialectal Code-Switching on Speech Modules: A Study using Egyptian Arabic Broadcast SpeechSAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech RecognitionBenchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and GermanNeural-based NLP systems for code-switched Arabic-English speechDialectal Arabic Code-Switching Dataset.MELANZ: A Trilingual Kreol Morisien–English–French Code-Switching SpeechCorpus for Automatic Speech Recognition

Effects of Dialectal Code-Switching on Speech Modules: A Study using Egyptian Arabic Broadcast Speech

The Intra-utterance code-switching (CS) is defined as the alternation between two or more languages within the same utterance. Despite the fact that spoken dialectal code-switching (DCS) is more challenging than CS, it remains largely unexplored. In this study, we

SAGE: Spliced-Audio Generated Data for Enhancing Foundational Models in Low-Resource Arabic-English Code-Switched Speech Recognition

This paper investigates the performance of various speech SSL models on dialectal Arabic (DA) and Ar

Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German

Code-switching -- the natural alternation between two languages within a single utterance -- remains

Neural-based NLP systems for code-switched Arabic-English speech

In the ever-evolving language landscape, code-switching has emerged as an interesting linguistic phe

Dialectal Arabic Code-Switching Dataset.

Dialectal Arabic Code-Switching Dataset: includes the annotated two-hours Egyptian dataset from the ADI-5 development split in the MGB-3 challenge. The first Dialectal Arabic Code Switching - DACS corpus from broadcast speech. Annotated at the token-level, conside

MELANZ: A Trilingual Kreol Morisien–English–French Code-Switching SpeechCorpus for Automatic Speech Recognition

MELANZ is the first code-switching (CS) speech corpus for Kreol Morisien (KM)