Fetching the paper…
Reading the bibliography…
We propose an efficient and interpretable scene graph generator.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. Lawrence Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, M. Bernstein, and L. Fei-Fei · 2016
Earlier work this paper cites.
Visual relationship detection with language priors
C. Lu, R. Krishna, M. Bernstein, and L. Fei-Fei · 2016
Earlier work this paper cites.
Detecting visual relationships with deep relational networks
B. Dai, Y. Zhang, and D. Lin · 2017
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Earlier work this paper cites.
Openimages: A public dataset for large-scale multi-label and multi-class image classification
I. Krasin, T. Duerig, N. Alldrin, V. Ferrari, S. Abu-El-Haija, A. Kuznetsova, H. Rom, J. Uijlings, S. Popov, S. Kamali, M. Malloci, J. Pont-Tuset, A. Veit, S. Belongie, V. Gomes, A. Gupta, C. Sun, G. Chechik, D. Cai, Z. Feng, D. Narayanan, and K. Murphy · 2017
Earlier work this paper cites.
Vip-cnn: A visual phrase reasoning convolutional neural network for visual relationship detection
Y. Li, W. Ouyang, and X. Wang · 2017
Earlier work this paper cites.
Deep variation-structured reinforcement learning for visual relationship and attribute detection
X. Liang, L. Lee, and E. P. Xing · 2017
Earlier work this paper cites.
Pixels to graphs by associative embedding
A. Newell and J. Deng · 2017
Cited alongside, same era.
Weakly-supervised learning of visual relations
J. Peyre, I. Laptev, C. Schmid, and J. Sivic · 2017
Cited alongside, same era.
Phrase localization and visual relationship detection with comprehensive image-language cues
B. A. Plummer, A. Mallya, C. M. Cervantes, J. Hockenmaier, and S. Lazebnik · 2017
Cited alongside, same era.
Scene graph generation by iterative message passing
D. Xu, Y. Zhu, C. Choy, and L. Fei-Fei · 2017
Cited alongside, same era.
Scene graph generation by iterative message passing
D. Xu, Y. Zhu, C. B. Choy, and L. Fei-Fei · 2017
Cited alongside, same era.
Visual relationship detection with internal and external linguistic knowledge distillation
R. Yu, A. Li, V. I. Morariu, and L. S. Davis · 2017
Cited alongside, same era.
Towards context-aware interaction recognition for visual relationship detection
B. Zhuang, L. Liu, C. Shen, and I. Reid · 2017
Later among the works it cites.
Detecting and recognizing human-object intaractions
G. Gkioxari, R. Girshick, P. Dollár, and K. He · 2018
Closest in time.
Referring relationships
R. Krishna, I. Chami, M. Bernstein, and L. Fei-Fei · 2018
Closest in time.
Neural baby talk
J. Lu, J. Yang, D. Batra, and D. Parikh · 2018
Closest in time.
Shuffle-then-assemble: Learning object-agnostic visual relationship features
X. Yang, H. Zhang, and J. Cai · 2018
Closest in time.
Exploring visual relationship for image captioning
T. Yao, Y. Pan, Y. Li, and T. Mei · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Visual translation embedding network for visual relation detection
H. Zhang, Z. Kyaw, S.-F. Chang, and T.-S. Chua · 2017
Cited alongside, same era.
Ppr-fcn: Weakly supervised visual relation detection via parallel pairwise r-fcn
H. Zhang, Z. Kyaw, J. Yu, and S.-F. Chang · 2017
Cited alongside, same era.
Relationship proposal networks
J. Zhang, M. Elhoseiny, S. Cohen, W. Chang, and A. Elgammal · 2017
Cited alongside, same era.
Neural motifs: Scene graph parsing with global context
R. Zellers, M. Yatskar, S. Thomson, and Y. Choi
Cited in the paper.
Zoom-net: Mining deep feature interactions for visual relationship recognition
G. Yin, L. Sheng, B. Liu, N. Yu, X. Wang, J. Shao, and C. Change Loy · 2018
Closest in time.
Large-scale visual relationship understanding
J. Zhang, Y. Kalantidis, M. Rohrbach, M. Paluri, A. Elgammal, and M. Elhoseiny · 2019
Closest in time.