Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

A New Benchmark for Evaluating Automatic Speech Recognition in the Arabic Call Domain

Domain:

natural language processing

Record type:

paperdatasetmodel
Creator:
ObaZa'JalMah
Host:avatar
This work is an attempt to introduce a comprehensive benchmark for Arabic speech recognition, specifically tailored to address the challenges of telephone conversations in Arabic language. Arabic, characterized by its rich dialectal diversity and phonetic complexity, presents a number of unique challenges for automatic speech recognition (ASR) systems. These challenges are further amplified in the domain of telephone calls, where audio quality, background noise, and conversational speech styles negatively affect recognition accuracy. Our work aims to establish a robust benchmark that not only encompasses the broad spectrum of Arabic dialects but also emulates the real-world conditions of call-based communications. By incorporating diverse dialectical expressions and accounting for the variable quality of call recordings, this benchmark seeks to provide a rigorous testing ground for the development and evaluation of ASR systems capable of navigating the complexities of Arabic speech in telephonic contexts. This work also attempts to establish a baseline performance evaluation using state-of-the-art ASR technologies.

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processing

Tags

Artificial IntelligenceComputation and Language

Similar

A New Tunisian Arabic Corpus and Benchmark for Automatic Speech RecognitionKARAKALPAK SPEECH CORPUS: THE FIRST BENCHMARK DATASET FOR AUTOMATIC SPEECH RECOGNITIONAfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech RecognitionAfriVox: An African benchmark dataset for Automatic Speech Translation and Speech Recognitionassermosa/Tunisian-Arabic-Automatic-Speech-Recognition-ASR-Automatic Code-switched Academic Tunisian Arabic Speech Recognition

A New Tunisian Arabic Corpus and Benchmark for Automatic Speech Recognition

KARAKALPAK SPEECH CORPUS: THE FIRST BENCHMARK DATASET FOR AUTOMATIC SPEECH RECOGNITION

While large-scale pre-trained models have significantly advanced multilingual Automatic Speech Recog

AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognition

Recent large language models (LLMs) show strong speech recognition and translation capabilities for

AfriVox: An African benchmark dataset for Automatic Speech Translation and Speech Recognition

This project creates a benchmark dataset for evaluating Automatic Speech Translation and Speech reco

assermosa/Tunisian-Arabic-Automatic-Speech-Recognition-ASR-

project combines multiple Tunisian speech datasets, applies audio augmentation techniques, and achie

Automatic Code-switched Academic Tunisian Arabic Speech Recognition