Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Linguistic disparities in cross-language automatic speech recognition transfer from Arabic to Tashlhiyt

Domain:

natural language processing

Record type:

paper
Creator:
GeoMoh
Publisher:
Spr
Host:
Abstract Tashlhiyt is a low-resource language with respect to acoustic databases, language corpora, and speech technology tools, such as Automatic Speech Recognition (ASR) systems. This study investigates whether a method of cross-language re-use of ASR is viable for Tashlhiyt from an existing commercially-available system built for Arabic. The source and target language in this case have similar phonological inventories, but Tashlhiyt permits typologically rare phonological patterns, including vowelless words, while Arabic does not. We find systematic disparities in ASR transfer performance (measured as word error rate (WER) and Levenshtein distance) for Tashlhiyt across word forms and speaking style variation. Overall, performance was worse for casual speaking modes across the board. In clear speech, performance was lower for vowelless than for voweled words. These results highlight systematic speaking mode- and phonotactic-disparities in cross-language ASR transfer. They also indicate that linguistically-informed approaches to ASR re-use can provide more effective ways to adapt existing speech technology tools for low resource languages, especially when they contain typologically rare structures. The study also speaks to issues of linguistic disparities in ASR and speech technology more broadly. It can also contribute to understanding the extent to which machines are similar to, or different from, humans in mapping the acoustic signal to discrete linguistic representations.

Visit

doi.org

Tasks

automatic speech recognitionspeech processing

Languages

Tachelhit

Licenses

https://creativecommons.org/licenses/by/4.0https://creativecommons.org/licenses/by/4.0

Similar

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognitionassermosa/Tunisian-Arabic-Automatic-Speech-Recognition-ASR-On Developing An Automatic Speech Recognition System For Standard Arabic LanguageAutomatic Code-switched Academic Tunisian Arabic Speech RecognitionAutomatic Speech Recognition for the Ika LanguageAutomatic speech recognition of the isiZulu language

Evaluating the Effect of Linguistic Relatedness on Cross-Lingual Transfer in Large Multilingual Automatic Speech Recognition

Extending automatic speech recognition (ASR) to low-resource African languages is constrained by the

assermosa/Tunisian-Arabic-Automatic-Speech-Recognition-ASR-

project combines multiple Tunisian speech datasets, applies audio augmentation techniques, and achie

On Developing An Automatic Speech Recognition System For Standard Arabic Language

The Automatic Speech Recognition (ASR) applied to Arabic language is a challenging task. This is mainly related to the language specificities which make the researchers facing multiple difficulties such as the insufficient linguistic resources and the very limited

Automatic Code-switched Academic Tunisian Arabic Speech Recognition

Automatic Speech Recognition for the Ika Language

We present a cost-effective approach for developing Automatic Speech Recognition (ASR) models for lo

Automatic speech recognition of the isiZulu language

A key component of artificial intelligence is human-to-machine communication. Such communication has