Fetching the paper…

Semantic speech retrieval with a visually grounded model of untranscribed speech · Around