Fetching the paper…

Weakly-Supervised Multi-Level Attentional Reconstruction Network for Grounding Textual Queries in Videos · Around