Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Linguistically Informed Evaluation of Multilingual ASR for African Languages

Domain:

natural language processing

Record type:

paper
Creator:
CheAdeDow
Host:avatar
Word Error Rate (WER) mischaracterizes ASR models' performance for African languages by combining phonological, tone, and other linguistic errors into a single lexical error. By contrast, Feature Error Rate (FER) has recently attracted attention as a viable metric that reveals linguistically meaningful errors in models' performance. In this paper, we evaluate three speech encoders on two African languages by complementing WER with CER, and FER, and add a tone-aware extension (TER). We show that by computing errors on phonological features, FER and TER reveal linguistically-salient error patterns even when word-level accuracy remains low. Our results reveal that models perform better on segmental features, while tones (especially mid and downstep) remain the most challenging features. Results on Yoruba show a striking differential in metrics, with WER=0.788, CER=0.305, and FER=0.151. Similarly for Uneme (an endangered language absent from pretraining data) a model with near-total WER and 0.461 CER achieves the relatively low FER of 0.267. This indicates model error is often attributable to individual phonetic feature errors, which is obscured by all-or-nothing metrics like WER. To appear at AfricaNLP 2026

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Languages

UnemeYoruba

Tags

Computation and Language

Similar

Linguistically Informed Tokenization Improves ASR for Underresourced LanguagesNSL-MT: Linguistically Informed Negative Samples for Efficient Machine Translation in African Low-Resource LanguagesMini But Mighty: Efficient Multilingual Pretraining with Linguistically-Informed Data SelectionNSL-MT: Linguistically Informed Negative Samples for Efficient Machine Translation in Low-Resource LanguagesFrom Monolingual to Multilingual: Evaluating Mamba for ASR in South African LanguagesLinguistically enriched corpora for conjunctively written South African languages

Linguistically Informed Tokenization Improves ASR for Underresourced Languages

Automatic speech recognition (ASR) is a crucial tool for linguists aiming to perform a variety of la

NSL-MT: Linguistically Informed Negative Samples for Efficient Machine Translation in African Low-Resource Languages

Mini But Mighty: Efficient Multilingual Pretraining with Linguistically-Informed Data Selection

With the prominence of large pretrained language models, low-resource languages are rarely modelled

NSL-MT: Linguistically Informed Negative Samples for Efficient Machine Translation in Low-Resource Languages

We introduce negative space learning machine translation (NSL-MT), a training method for underresour

From Monolingual to Multilingual: Evaluating Mamba for ASR in South African Languages

Recent advances in automatic speech recognition (ASR) have explored different sequence models, inclu

Linguistically enriched corpora for conjunctively written South African languages

This resource contains linguistically annotated data for four official South African languages with