Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
AshSanAwoGic
Hôte:avatar
Recent advances in speech-enabled AI, including Google's NotebookLM and OpenAI's speech-to-speech API, are driving widespread interest in voice interfaces globally. Despite this momentum, there exists no publicly available application-specific model evaluation that caters to Africa's linguistic diversity. We present AfriSpeech-MultiBench, the first domain-specific evaluation suite for over 100 African English accents across 10+ countries and seven application domains: Finance, Legal, Medical, General dialogue, Call Center, Named Entities and Hallucination Robustness. We benchmark a diverse range of open, closed, unimodal ASR and multimodal LLM-based speech recognition systems using both spontaneous and non-spontaneous speech conversation drawn from various open African accented English speech datasets. Our empirical analysis reveals systematic variation: open-source ASR models excels in spontaneous speech contexts but degrades on noisy, non-native dialogue; multimodal LLMs are more accent-robust yet struggle with domain-specific named entities; proprietary models deliver high accuracy on clean speech but vary significantly by country and domain. Models fine-tuned on African English achieve competitive accuracy with lower latency, a practical advantage for deployment, hallucinations still remain a big problem for most SOTA models. By releasing this comprehensive benchmark, we empower practitioners and researchers to select voice technologies suited to African use-cases, fostering inclusive voice applications for underserved communities. Accepted As a Conference Paper IJCNLP-AACL 2025

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Computation and Language

Similaires

AfriSpeech-200: Pan-African Accented Speech Dataset for Clinical and General Domain ASRAccoustic Modeling for Development of Accented Indian English ASRAfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech RecognitionAfrispeech-Dialog: A Benchmark Dataset for Spontaneous English Conversations in Healthcare and BeyondAdvancing African-Accented English Speech Recognition: Epistemic Uncertainty-Driven Data Selection for Generalizable ASR ModelsInnocent-ICS/asr-afrispeech

AfriSpeech-200: Pan-African Accented Speech Dataset for Clinical and General Domain ASR

Africa has a very poor doctor-to-patient ratio. At very busy clinics, doctors could see 30+ patients

Accoustic Modeling for Development of Accented Indian English ASR

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

Recent large language models (LLMs) show strong speech recognition and translation capabilities for

Afrispeech-Dialog: A Benchmark Dataset for Spontaneous English Conversations in Healthcare and Beyond

Speech technologies are transforming interactions across various sectors, from healthcare to call ce

Advancing African-Accented English Speech Recognition: Epistemic Uncertainty-Driven Data Selection for Generalizable ASR Models

Accents play a pivotal role in shaping human communication, enhancing our ability to convey and comp

Innocent-ICS/asr-afrispeech

This project is an RNN-based ASR system built to help African doctors auto-transcribe their consulta