Fetching the paper…
Reading the bibliography…
Indoor scene understanding is central to applications such as robot navigation and human companion assistance.
Metropolis light transport
E. Veach and L. J. Guibas · 1997
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
Evaluation of image features using a photorealistic virtual world
B. Kaneva, A. Torralba, and W. T. Freeman · 2011
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus · 2012
Earlier work this paper cites.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Earlier work this paper cites.
Perceptual organization and recognition of indoor scenes from rgb-d images
S. Gupta, P. Arbelaez, and J. Malik · 2013
Earlier work this paper cites.
A benchmark for RGB-D visual odometry, 3D reconstruction and SLAM
A. Handa, T. Whelan, J. McDonald, and A. Davison · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Understanding deep features with computer-generated imagery
M. Aubry and B. C. Russell · 2015
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
A. X. Chang, T. Funkhouser, L. Guibas, P. Hanrahan, Q. Huang, Z. Li, S. Savarese, M. Savva, S. Song, H. Su, et al · 2015
Cited alongside, same era.
Fast edge detection using structured forests
P. Dollár and C. L. Zitnick · 2015
Cited alongside, same era.
Flownet: Learning optical flow with convolutional networks
A. Dosovitskiy, P. Fischery, E. Ilg, C. Hazirbas, V. Golkov, P. van der Smagt, D. Cremers, T. Brox, et al · 2015
Cited alongside, same era.
Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture
D. Eigen and R. Fergus · 2015
Cited alongside, same era.
Aligning 3d models to rgb-d images of cluttered scenes
S. Gupta, P. Arbeláez, R. Girshick, and J. Malik · 2015
Cited alongside, same era.
Holistically-nested edge detection
S. Xie and Z. Tu · 2015
Later among the works it cites.
Marr revisited: 2D-3D alignment via surface normal prediction
A. Bansal, B. C. Russell, and A. Gupta · 2016
Closest in time.
How useful is photo-realistic rendering for visual learning?
Y. Movshovitz-Attias, T. Kanade, and Y. Sheikh · 2016
Closest in time.
Playing for data: Ground truth from computer games
S. R. Richter, V. Vineet, S. Roth, and V. Koltun · 2016
Closest in time.
The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes
G. Ros, L. Sellart, J. Materzynska, D. Vazquez, and A. M. Lopez · 2016
Closest in time.
Semantic Scene Completion from a Single Depth Image
S. Shuran, Y. Fisher, Z. Andy, X. C. Angel, S. Manolis, and F. Thomas · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Handa, V. Patraucean, V. Badrinarayanan, S. Stent, and R. Cipolla · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
SUN RGB-D: A RGB-D scene understanding benchmark suite
S. Song, S. Lichtenberg, and J. Xiao · 2015
Cited alongside, same era.
Render for cnn: Viewpoint estimation in images using cnns trained with rendered 3d model views
H. Su, C. R. Qi, Y. Li, and L. J. Guibas · 2015
Cited alongside, same era.
http://www.mitsuba-renderer.org/
Mitsuba physically based renderer
Cited in the paper.
Discriminatively trained dense surface normal estimation
B. Z. L’ubor Ladickỳ and M. Pollefeys
Cited in the paper.
Closest in time.
Objectnet3d: A large scale database for 3d object recognition
Y. Xiang, W. Kim, W. Chen, J. Ji, C. Choy, H. Su, R. Mottaghi, L. Guibas, and S. Savarese · 2016
Closest in time.
Multi-scale context aggregation by dilated convolutions
F. Yu and V. Koltun · 2016
Closest in time.
Deepcontext: Context-encoding neural pathways for 3d holistic scene understanding
Y. Zhang, M. Bai, P. Kohli, S. Izadi, and J. Xiao · 2016
Closest in time.