Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

OpenWER: Improving Cross-Lingual ASR Evaluation and Enabling Token-Based Accuracy Metrics

Domaine:

natural language processing

Type de record:

papersoftware
Créateur:
KuhZim
Hôte:avatar
Advances in deep learning and end-to-end Automatic Speech Recognition (ASR) have enabled robust multilingual models, but evaluation metrics remain limited in assessing accuracy. Efforts to improve or replace the common metric Word Error Rate (WER) often focus on English, leaving evaluations for low-resource languages under-explored and hindering fair cross-lingual comparisons. We present OpenWER, an open-source implementation that improves WER robustness through language-specific normalisation and compound word detection. A token-based Levenshtein alignment preserves complementary metrics and allows metadata embedding for granular accuracy scores. Our analysis of 52 languages shows absolute WER reductions of up to 25% compared to common libraries. OpenWER contributes to fairness in ASR research by increasing the reliability of WER across diverse languages and enabling more comprehensive accuracy evaluations. 5 pages, 2 figures

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Computation and LanguageSoundI.2.7

Similaires

Discrete vs Continuous Audio Token Representations in Cross-Lingual Transfer Accuracy on CommonVoice Low-Resource BenchmarkProjection-based Method Scalability and Cross-lingual NER Accuracy in Low-Resource LanguagesCONCRETE: Improving Cross-lingual Fact-checking with Cross-lingual RetrievalImproving Generative Cross-lingual Aspect-Based Sentiment Analysis with Constrained DecodingCode-Switched Token Proportion Effects on Zero-Shot Cross-Lingual Dense RetrieversAfriCLIRMatrix: Enabling Cross-Lingual Information Retrieval for African Languages

Discrete vs Continuous Audio Token Representations in Cross-Lingual Transfer Accuracy on CommonVoice Low-Resource Benchmark

This paper presents XLSR which learns cross-lingual speech representations by pretraining a single m

Projection-based Method Scalability and Cross-lingual NER Accuracy in Low-Resource Languages

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

CONCRETE: Improving Cross-lingual Fact-checking with Cross-lingual Retrieval

Fact-checking has gained increasing attention due to the widespread of falsified information. Most f

Improving Generative Cross-lingual Aspect-Based Sentiment Analysis with Constrained Decoding

While aspect-based sentiment analysis (ABSA) has made substantial progress, challenges remain for lo

Code-Switched Token Proportion Effects on Zero-Shot Cross-Lingual Dense Retrievers

Transferring information retrieval (IR) models from a high-resource language (typically English) to

AfriCLIRMatrix: Enabling Cross-Lingual Information Retrieval for African Languages

Language diversity in NLP is critical in enabling the development of tools for a wide range of users.However, there are limited resources for building such tools for many languages, particularly those spoken in Africa.For search, most existing datasets feature few