Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
AwoAshOla
Hôte:avatar
Recent large language models (LLMs) show strong speech recognition and translation capabilities for high-resource languages. However, African languages remain dramatically underrepresented in benchmarks, limiting their practical use in low-resource settings. While early benchmarks tested African languages and accents, they lacked exhaustive real-world noise and granular domain evaluations. We present AfriVox-v2, a comprehensive benchmark designed to test speech models under realistic African deployment conditions. AfriVox-v2 introduces "in the wild" unscripted audio for all supported languages. We also introduce strict domain verticalization, evaluating model accuracy across ten sectors including government, finance, health, and agriculture and conducting targeted tests on numbers and named entities. Finally, we benchmark a new generation of speech models, including Sahara-v2, Gemini 3 Flash, and the Omnilingual CTC models. Our results expose the true generalization gap of modern speech models in specialized, noisy African contexts and provide a reliable blueprint for developers building localized voice AI.

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Computation and LanguageSound

Similaires

AfriVox: An African benchmark dataset for Automatic Speech Translation and Speech RecognitionAfriSwitch: A Benchmark for In-the-Wild African Code-Switched Speech RecognitionAfriVox-Translate: An African benchmark dataset for Automatic Speech TranslationA New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call DomainAfriVox-v2AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR

AfriVox: An African benchmark dataset for Automatic Speech Translation and Speech Recognition

This project creates a benchmark dataset for evaluating Automatic Speech Translation and Speech reco

AfriSwitch: A Benchmark for In-the-Wild African Code-Switched Speech Recognition

Code-switching is pervasive in bilingual African conversation, yet most ASR systems assume monolingu

AfriVox-Translate: An African benchmark dataset for Automatic Speech Translation

This project creates a benchmark dataset for evaluating Automatic Speech Translation models on Afric

A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain

This work is an attempt to introduce a comprehensive benchmark for Arabic speech recognition, specif

AfriVox-v2

AfriVox-v2 is a comprehensive multilingual speech recognition benchmark designed to evaluate ASR sys

AfriSpeech-MultiBench: A Verticalized Multidomain Multicountry Benchmark Suite for African Accented English ASR

Recent advances in speech-enabled AI, including Google's NotebookLM and OpenAI's speech-to-speech AP