Fetching the paper…

YFACC: A Yor\`ub\'a speech-image dataset for cross-lingual keyword localisation through visual grounding · Around