Visually prompted keyword localisation Poster presented at the Deep Learning Indaba 2022 by Aletta S E Nortje
Given an image query, visually prompted keyword localisation (VPKL) aims to find occurrences of the
This study investigates the use of Visually Grounded Speech (VGS) models for keyword localisation in