Fetching the paper…
Reading the bibliography…
Understanding a scene by decoding the visual relationships depicted in an image has been a long studied problem.
“Scene graph generation with external knowledge and image reconstruction”
Jiuxiang Gu et al · 1978
Earlier work this paper cites.
“Policy gradient methods for reinforcement learning with function approximation”
Richard Sutton, David McAllester, Satinder Singh and Yishay Mansour · 2000
Earlier work this paper cites.
“Diffusion-convolutional neural networks”
James Atwood and Don Towsley · 2001
Earlier work this paper cites.
“Direct and indirect effects”
Judea Pearl · 2001
Earlier work this paper cites.
“BLEU: a method for automatic evaluation of machine translation”
Kishore Papineni, Salim Roukos, Todd Ward and Wei-Jing Zhu · 2002
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng et al · 2009
Earlier work this paper cites.
“Neural network for graphs: A contextual constructive approach”
Alessio Micheli · 2009
Earlier work this paper cites.
“Word sense disambiguation: A survey”
Roberto Navigli · 2009
Earlier work this paper cites.
“The Graph Neural Network Model”
F. Scarselli et al · 2009
Earlier work this paper cites.
“Recognition using visual phrases”
Ali Farhadi and Amin Sadeghi · 2011
Earlier work this paper cites.
“Imagenet classification with deep convolutional neural networks”
Alex Krizhevsky, Ilya Sutskever and Geoffrey Hinton · 2012
Earlier work this paper cites.
“An introduction to conditional random fields”
Charles Sutton and Andrew McCallum · 2012
Earlier work this paper cites.
“Spectral Networks and Locally Connected Networks on Graphs”, 2013
Joan Bruna, Wojciech Zaremba, Arthur Szlam and Yann LeCun · 2013
Earlier work this paper cites.
“Auto-encoding variational bayes”
Diederik Kingma and Max Welling · 2013
Earlier work this paper cites.
“Combining generative and discriminative model scores for distant supervision”
Benjamin Roth and Dietrich Klakow · 2013
Earlier work this paper cites.
“ConceptNet 5: A large semantic network for relational knowledge”
Robert Speer and Catherine Havasi · 2013
Earlier work this paper cites.
“Rich feature hierarchies for accurate object detection and semantic segmentation”
Ross Girshick, Jeff Donahue, Trevor Darrell and Jitendra Malik · 2014
Earlier work this paper cites.
“Generative Adversarial Nets”
Ian Goodfellow et al · 2014
Earlier work this paper cites.
“Microsoft coco: Common objects in context”
Tsung-Yi Lin et al · 2014
Earlier work this paper cites.
“Very deep convolutional networks for large-scale image recognition”
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
“Object Detectors Emerge in Deep Scene CNNs”, 2014
Bolei Zhou et al · 2014
Earlier work this paper cites.
“Vqa: Visual question answering”
Stanislaw Antol et al · 2015
Earlier work this paper cites.
“Activitynet: A large-scale video benchmark for human activity understanding”
Fabian Caba, Victor Escorcia, Bernard Ghanem and Juan Carlos · 2015
Earlier work this paper cites.
“Fast r-cnn”
Ross Girshick · 2015
Earlier work this paper cites.
“Image retrieval using scene graphs”
Justin Johnson et al · 2015
Earlier work this paper cites.
“Deep visual-semantic alignments for generating image descriptions”
Andrej Karpathy and Li Fei-Fei · 2015
Earlier work this paper cites.
“Gated graph sequence neural networks”
Yujia Li, Daniel Tarlow, Marc Brockschmidt and Richard Zemel · 2015
Earlier work this paper cites.
“Faster r-cnn: Towards real-time object detection with region proposal networks”
Shaoqing Ren, Kaiming He, Ross Girshick and Jian Sun · 2015
Earlier work this paper cites.
“U-net: Convolutional networks for biomedical image segmentation”
Olaf Ronneberger, Philipp Fischer and Thomas Brox · 2015
Earlier work this paper cites.
“Imagenet large scale visual recognition challenge”
Olga Russakovsky et al · 2015
Earlier work this paper cites.
“Generating semantically precise scene graphs from textual descriptions for improved image retrieval”
Sebastian Schuster et al · 2015
Earlier work this paper cites.
“Going deeper with convolutions”
Christian Szegedy et al · 2015
Earlier work this paper cites.
“Improved semantic representations from tree-structured long short-term memory networks”
Kai Tai, Richard Socher and Christopher Manning · 2015
Cited alongside, same era.
“Cider: Consensus-based image description evaluation”
Ramakrishna Vedantam, C Lawrence and Devi Parikh · 2015
Cited alongside, same era.
“Learning from massive noisy labeled data for image classification”
Tong Xiao et al · 2015
Cited alongside, same era.
“Learning Deep Features for Discriminative Localization”, 2015
Bolei Zhou et al · 2015
Cited alongside, same era.
“Spice: Semantic propositional image caption evaluation”
Peter Anderson, Basura Fernando, Mark Johnson and Stephen Gould · 2016
Cited alongside, same era.
“Bottom-up and top-down attention for image captioning and visual question answering”
Peter Anderson et al · 2018
Later among the works it cites.
“Encoder-decoder with atrous separable convolution for semantic image segmentation”
Liang-Chieh Chen et al · 2018
Later among the works it cites.
“Visual Graphs from Motion (VGfM): Scene understanding with object geometry reasoning”
Paul Gay, James Stuart and Alessio Del · 2018
Later among the works it cites.
“Image generation from scene graphs”
Justin Johnson, Agrim Gupta and Li Fei-Fei · 2018
Later among the works it cites.
“Factorizable net: an efficient subgraph-based framework for scene graph generation”
Yikang Li et al · 2018
Later among the works it cites.
“Neural baby talk”
Jiasen Lu, Jianwei Yang, Dhruv Batra and Devi Parikh · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jifeng Dai, Yi Li, Kaiming He and Jian Sun · 2016
Cited alongside, same era.
“Convolutional neural networks on graphs with fast localized spectral filtering”
Michaël Defferrard, Xavier Bresson and Pierre Vandergheynst · 2016
Cited alongside, same era.
“Density estimation using real nvp”
Laurent Dinh, Jascha Sohl-Dickstein and Samy Bengio · 2016
Cited alongside, same era.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Cited alongside, same era.
“Semi-supervised classification with graph convolutional networks”
Thomas Kipf and Max Welling · 2016
Cited alongside, same era.
“Visual relationship detection with language priors”
Cewu Lu, Ranjay Krishna, Michael Bernstein and Li Fei-Fei · 2016
Cited alongside, same era.
“Hierarchical question-image co-attention for visual question answering”
Jiasen Lu, Jianwei Yang, Dhruv Batra and Devi Parikh · 2016
Cited alongside, same era.
“Yolov3: An incremental improvement”
Joseph Redmon and Ali Farhadi · 2018
Later among the works it cites.
“Graph r-cnn for scene graph generation”
Jianwei Yang et al · 2018
Later among the works it cites.
“Exploring visual relationship for image captioning”
Ting Yao, Yingwei Pan, Yehao Li and Tao Mei · 2018
Later among the works it cites.
“Neural motifs: Scene graph parsing with global context”
Rowan Zellers, Mark Yatskar, Sam Thomson and Yejin Choi · 2018
Later among the works it cites.
“Gaan: Gated attention networks for learning on large and spatiotemporal graphs”
Jiani Zhang et al · 2018
Later among the works it cites.
“3D Scene Graph: A Structure for Unified Semantics, 3D Space, and Camera”
Iro Armeni et al · 2019
Later among the works it cites.
“Specifying object attributes and relations in interactive scene generation”
Oron Ashual and Lior Wolf · 2019
Later among the works it cites.
“A short note on the kinetics-700 human action dataset”
Joao Carreira, Eric Noland, Chloe Hillier and Andrew Zisserman · 2019
Later among the works it cites.
“An attentive survey of attention models”
Sneha Chaudhari, Gungor Polatkan, Rohan Ramanath and Varun Mithal · 2019
Later among the works it cites.
“Counterfactual critic multi-agent training for scene graph generation”
Long Chen et al · 2019
Later among the works it cites.
“Knowledge-embedded routing network for scene graph generation”
Tianshui Chen, Weihao Yu, Riquan Chen and Liang Lin · 2019
Later among the works it cites.
“Scene graph prediction with limited labels”
Vincent Chen et al · 2019
Later among the works it cites.
“Visual Relationships as Functions: Enabling Few-Shot Scene Graph Prediction”
Apoorva Dornadula et al · 2019
Later among the works it cites.
Shalini Ghosh, Giedrius Burachas, Arijit Ray and Avi Ziskind · 2019
Later among the works it cites.
“Unpaired image captioning via scene graph alignments”
Jiuxiang Gu et al · 2019
Later among the works it cites.
“Action Genome: Actions as Composition of Spatio-temporal Scene Graphs”
Jingwei Ji, Ranjay Krishna, Li Fei-Fei and Juan Niebles · 2019
Later among the works it cites.
“Decoupling representation and classifier for long-tailed recognition”
Bingyi Kang et al · 2019
Later among the works it cites.
“Large-scale long-tailed recognition in an open world”
Ziwei Liu et al · 2019
Later among the works it cites.
“Learning to compose dynamic tree structures for visual contexts”
Kaihua Tang et al · 2019
Later among the works it cites.
“A comprehensive survey on graph neural networks”
Zonghan Wu et al · 2019
Later among the works it cites.
“Auto-encoding scene graphs for image captioning”
Xu Yang, Kaihua Tang, Hanwang Zhang and Jianfei Cai · 2019
Later among the works it cites.
“Graphical contrastive losses for scene graph parsing”
Ji Zhang et al · 2019
Later among the works it cites.
“Hacs: Human action clips and segments dataset for recognition and temporal localization”
Hang Zhao, Antonio Torralba, Lorenzo Torresani and Zhicheng Yan · 2019
Later among the works it cites.
“Unbiased Scene Graph Generation from Biased Training”
Kaihua Tang et al · 2020
Closest in time.
“Show, attend and tell: Neural image caption generation with visual attention”
Kelvin Xu et al · 2057
Closest in time.