Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

A Deep Learning Framework for Arabic Continuous Speech Keyword Spotting in Low-Resource Settings Using Isolated-Word Keyword Spotting and Posterior Probability Functions

Domaine:

natural language processing

Type de record:

paper
Créateur:
OsaAssOum
Éditeur:
Adv
Hôte:
Continuous Speech Keyword Spotting (CSKWS) presents a challenging paradigm shift from isolated-word Keyword Spotting (KWS), focusing on discovering the occurrences of predefined keywords within continuous speech streams. This paper addresses the prevalent issue of data scarcity in CSKWS for low-resource languages by introducing the innovative Posterior Probability Function approach (PPF-CSKWS). Utilizing unsupervised method, this approach leverages a few-shot KWS system to derive posterior probability estimates of keyword occurrences as a discrete-time functions. Contrary to the data-intensive training procedures typically associated with CSKWS system development, this method requires only 15 isolated audio samples per keyword, significantly reducing the data bottleneck. The generated posterior probability functions provide crucial temporal information, facilitating both keyword identification and localization. This characteristic allows for the reformulation of the CSKWS problem as a detection task, with evaluation metrics such as mean Average Precision (mAP) and Maximum Term Weighted Value (MTWV) being applicable. To evaluate the proposed method, a dedicated Arabic speech corpus was constructed. Experimental results demonstrated the achievement of mAP = 0.613 and MTWV = 0.641. This performance meets or exceeds that of established supervised techniques requiring significantly larger quantities of labelled training data, thereby demonstrating the potential of PPF-CSKWS in data scarcity scenarios.

Visit

doi.org

Tasks

keywordsspeech processing

Similaires

Low-Resource Speech Recognition and Keyword-SpottingLow-resource keyword spotting using contrastively trained transformer acoustic word embeddingsSynth4Kws: Synthesized Speech for User Defined Keyword Spotting in Low Resource EnvironmentsFeature learning for efficient ASR-free keyword spotting in low-resource languagesUsing web text to improve keyword spotting in speechHalimatou10/Fulfulde-Keyword-Spotting

Low-Resource Speech Recognition and Keyword-Spotting

The IARPA Babel program ran from March 2012 to November 2016. The aim of the program was to develop

Low-resource keyword spotting using contrastively trained transformer acoustic word embeddings

We introduce a new approach, the ContrastiveTransformer, that produces acoustic word embeddings (AWE

Synth4Kws: Synthesized Speech for User Defined Keyword Spotting in Low Resource Environments

One of the challenges in developing a high quality custom keyword spotting (KWS) model is the length

Feature learning for efficient ASR-free keyword spotting in low-resource languages

We consider feature learning for efficient keyword spotting that can be applied in severely under-re

Using web text to improve keyword spotting in speech

For low resource languages, collecting sufficient training data to build acoustic and language model

Halimatou10/Fulfulde-Keyword-Spotting

Keyword Spotting system for Fulfulde livestock vocabulary # KWS Fulfulde ## Description Ce projet