Fetching the paper…

Learning Visual Relation Priors for Image-Text Matching and Image Captioning with Neural Scene Graph Generators · Around