Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

AfriVox-v2

Domain:

natural language processing

Record type:

dataset
Creator:
int
Host:
AfriVox-v2 is a comprehensive multilingual speech recognition benchmark designed to evaluate ASR systems under realistic African deployment conditions. It covers 24 African languages across 10 application domains, with a strong emphasis on spontaneous, unscripted "in the wild" audio. Dataset Summary

Visit

huggingface.co

Tasks

automatic speech recognitionspeech processing

Languages

AfrikaansAkanAmharicBembaFulaGaGandaHausaIgboKinyarwanda+11

Tags

audiospeechafrican-languagesasrbenchmarkconversationalin-the-wild

Licenses

cc-by-4.0

Similar

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognitionintronhealth/afrivox-transcribeAfriVox: Probing Multilingual and Accent Robustness of Speech LLMsAfriVox-Translate: An African benchmark dataset for Automatic Speech TranslationAfriVox: An African benchmark dataset for Automatic Speech Translation and Speech RecognitionSKDrepaData-v2

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

Recent large language models (LLMs) show strong speech recognition and translation capabilities for

intronhealth/afrivox-transcribe

This project creates a benchmark dataset for evaluating Automatic Speech recognition models on Afric

AfriVox: Probing Multilingual and Accent Robustness of Speech LLMs

Recent advances in multimodal and speech-native large language models (LLMs) have delivered impressi

AfriVox-Translate: An African benchmark dataset for Automatic Speech Translation

This project creates a benchmark dataset for evaluating Automatic Speech Translation models on Afric

AfriVox: An African benchmark dataset for Automatic Speech Translation and Speech Recognition

This project creates a benchmark dataset for evaluating Automatic Speech Translation and Speech reco

SKDrepaData-v2

SKDrepaData-v2 is a biomedical image dataset composed of 1,489 microscopic images of blood smears co