Fetching the paper…
Reading the bibliography…
Understanding a visual scene goes beyond recognizing individual objects in isolation.
Contextual priming for object detection
A. Torralba · 2003
Earlier work this paper cites.
The role of context in object recognition
A. Oliva and A. Torralba · 2007
Earlier work this paper cites.
Objects in context
A. Rabinovich, A. Vedaldi, C. Galleguillos, E. Wiewiora, and S. Belongie · 2007
Earlier work this paper cites.
Statistics of 3d object locations in images
R. Baur, A. Efros, and M. Hebert · 2008
Earlier work this paper cites.
Object categorization using co-occurrence, location and appearance
C. Galleguillos, A. Rabinovich, and S. Belongie · 2008
Earlier work this paper cites.
Beyond nouns: Exploiting prepositions and comparative adjectives for learning visual classifiers
A. Gupta and L. S. Davis · 2008
Earlier work this paper cites.
Discriminative models for static human-object interactions
C. Desai, D. Ramanan, and C. Fowlkes · 2010
Earlier work this paper cites.
Graph cut based inference with co-occurrence statistics
L. Ladicky, C. Russell, P. Kohli, and P. H. Torr · 2010
Earlier work this paper cites.
Modeling mutual context of object and human pose in human-object interaction activities
B. Yao and L. Fei-Fei · 2010
Earlier work this paper cites.
Discriminative models for multi-class object layout
C. Desai, D. Ramanan, and C. C. Fowlkes · 2011
Earlier work this paper cites.
Characterizing structural relationships in scenes using graph kernels
M. Fisher, M. Savva, and P. Hanrahan · 2011
Earlier work this paper cites.
Efficient inference in fully connected crfs with gaussian edge potentials
P. Krähenbühl and V. Koltun · 2011
Earlier work this paper cites.
Recognition using visual phrases
M. A. Sadeghi and A. Farhadi · 2011
Earlier work this paper cites.
Learning to share visual appearance for multiclass object detection
R. Salakhutdinov, A. Torralba, and J. Tenenbaum · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
P. K. Nathan Silberman, Derek Hoiem and R. Fergus · 2012
Cited alongside, same era.
3d-based reasoning with blocks, support, and stability
Z. Jia, A. Gallagher, A. Saxena, and T. Chen · 2013
Cited alongside, same era.
Holistic scene understanding for 3d object detection with rgbd cameras
D. Lin, S. Fidler, and R. Urtasun · 2013
Cited alongside, same era.
Scene parsing by integrating function, geometry and appearance models
Y. Zhao and S.-C. Zhu · 2013
Cited alongside, same era.
Learning the visual interpretation of sentences
C. L. Zitnick, D. Parikh, and L. Vanderwende · 2013
Cited alongside, same era.
Learning spatial knowledge for text to 3d scene generation
A. X. Chang, M. Savva, and C. D. Manning · 2014
Cited alongside, same era.
Learning semantic relationships for better action retrieval in images
V. Ramanathan, C. Li, J. Deng, W. Han, Z. Li, K. Gu, Y. Song, S. Bengio, C. Rossenberg, and L. Fei-Fei · 2015
Later among the works it cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Later among the works it cites.
Describing common human visual actions in images
M. R. Ronchi and P. Perona · 2015
Later among the works it cites.
Scene understanding by reasoning stability and safety
B. Zheng, Y. Zhao, J. Yu, K. Ikeuchi, and S.-C. Zhu · 2015
Later among the works it cites.
Conditional random fields as recurrent neural networks
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. Torr · 2015
Later among the works it cites.
Spice: Semantic propositional image caption evaluation
P. Anderson, B. Fernando, M. Johnson, and S. Gould · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On the properties of neural machine translation: Encoder-decoder approaches
K. Cho, B. Van Merriënboer, D. Bahdanau, and Y. Bengio · 2014
Cited alongside, same era.
The role of context for object detection and semantic segmentation in the wild
R. Mottaghi, X. Chen, X. Liu, N.-G. Cho, S.-W. Lee, S. Fidler, R. Urtasun, and A. Yuille · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Inside-outside net: Detecting objects in context with skip pooling and recurrent neural networks
S. Bell, C. L. Zitnick, K. Bala, and R. Girshick · 2015
Cited alongside, same era.
Hico: A benchmark for recognizing human-object interactions in images
Y.-W. Chao, Z. Wang, Y. He, J. Wang, and J. Deng · 2015
Cited alongside, same era.
Hico: A benchmark for recognizing human-object interactions in images
Y.-W. Chao, Z. Wang, Y. He, J. Wang, and J. Deng · 2015
Cited alongside, same era.
Later among the works it cites.
Region-based convolutional networks for accurate object detection and segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, M. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Semantic object parsing with graph lstm
X. Liang, X. Shen, J. Feng, L. Lin, and S. Yan · 2016
Later among the works it cites.
On support relations and semantic scene graphs
W. Liao, M. Y. Yang, H. Ackermann, and B. Rosenhahn · 2016
Later among the works it cites.
Visual relationship detection with language priors
C. Lu, R. Krishna, M. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Graph-structured representations for visual question answering
D. Teney, L. Liu, and A. v. d. Hengel · 2016
Later among the works it cites.
Situation recognition: Visual semantic role labeling for image understanding
M. Yatskar, L. Zettlemoyer, and A. Farhadi · 2016
Later among the works it cites.