This paper examines the impact of multilingual (ML) acoustic representations on Automatic Speech Rec
The IARPA Babel program ran from March 2012 to November 2016. The aim of the program was to develop
This study investigates the use of Visually Grounded Speech (VGS) models for keyword localisation in
International audience For languages with limited training resources, out-of-vocabula
We consider feature learning for efficient keyword spotting that can be applied in severely under-re
One of the challenges in developing a high quality custom keyword spotting (KWS) model is the length