Fetching the paper…
Reading the bibliography…
Visual scene understanding is an important capability that enables robots to purposefully act in their environment.
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle, “Greedy layer-wise training of deep networks,” in Advances in Neural Information Processing Systems (NIPS)
2007
Earlier work this paper cites.
P. Krähenbühl and V. Koltun, “Efficient inference in fully connected CRFs with Gaussian edge potentials,” in Advances in Neural Information Processing Systems (NIPS)
2011
Earlier work this paper cites.
P. K. Nathan Silberman, Derek Hoiem and R. Fergus, “Indoor segmentation and support inference from RGBD images,” in Europ. Conf. on Computer Vision (ECCV)
2012
Earlier work this paper cites.
L. Bottou, “Stochastic gradient descent tricks,” in Neural Networks: Tricks of the Trade
2012
Earlier work this paper cites.
R. F. Salas-Moreno, R. A. Newcombe, H. Strasdat, P. H. Kelly, and A. J. Davison, “SLAM++: Simultaneous localisation and mapping at the level of objects,” IEEE Int. Conf. on Computer Vision and Pattern Recognition (CVPR)
2013
Earlier work this paper cites.
C. Kerl, J. Sturm, and D. Cremers, “Dense visual SLAM for RGB-D cameras,” in IEEE/RSJ Intelligent Robots and Systems (IROS)
2013
Earlier work this paper cites.
S. Gupta, P. Arbelaez, and J. Malik, “Perceptual organization and recognition of indoor scenes from RGB-D images,” in IEEE Conf. on Computer Vision and Pattern Recognition (CVPR)
2013
Earlier work this paper cites.
S. Gupta, R. Girshick, P. Arbeláez, and J. Malik, “Learning rich features from RGB-D images for object detection and segmentation,” in Europ. Conf. on Computer Vision (ECCV)
2014
Earlier work this paper cites.
R. Girshick, J. Donahue, T. Darrell, and J. Malik, “Rich feature hierarchies for accurate object detection and semantic segmentation,” in IEEE Computer Vision and Pattern Recognition (CVPR)
2014
Earlier work this paper cites.
A. Hermans, G. Floros, and B. Leibe, “Dense 3D semantic mapping of indoor scenes from rgb-d images,” in IEEE Int. Conf. on Robotics and Automation (ICRA)
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition (CVPR)
2015
Cited alongside, same era.
H. Noh, S. Hong, and B. Han, “Learning deconvolution network for semantic segmentation,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition (CVPR)
2015
Cited alongside, same era.
D. Eigen and R. Fergus, “Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture,” in IEEE Int. Conf. on Computer Vision (ICCV)
2015
Cited alongside, same era.
J. Stückler, B. Waldvogel, H. Schulz, and S. Behnke, “Dense real-time mapping of object-class semantics from RGB-D video,” J. of Real-Time Image Processing
2015
Cited alongside, same era.
H. Su, S. Maji, E. Kalogerakis, and E. G. Learned-Miller, “Multi-view convolutional neural networks for 3d shape recognition,” in IEEE Int. Conf. on Computer Vision (ICCV)
2016
Later among the works it cites.
Z. Li, Y. Gan, X. Liang, Y. Yu, H. Cheng, and L. Lin, “LSTM-CF: Unifying context modeling and fusion with LSTMs for RGB-D scene labeling,” in Europ. Conf. on Computer Vision (ECCV)
2016
Later among the works it cites.
2016
Later among the works it cites.
I. Armeni, O. Sener, A. R. Zamir, H. Jiang, I. Brilakis, M. Fischer, and S. Savarese, “3D semantic parsing of large-scale indoor spaces,” in IEEE Computer Vision and Pattern Recognition (CVPR)
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
C. Lee, S. Xie, P. W. Gallagher, Z. Zhang, and Z. Tu, “Deeply-supervised nets,” in Proc. of the 18th Int. Conf. on Artificial Intelligence and Statistics (AISTATS)
2015
Cited alongside, same era.
A. Dosovitskiy, P. Fischer, E. Ilg, P. Hausser, C. Hazirbas, V. Golkov, P. van der Smagt, D. Cremers, and T. Brox, “Flownet: Learning optical flow with convolutional networks,” in The IEEE Int. Conf. on Computer Vision (ICCV)
2015
Cited alongside, same era.
M. Jaderberg, K. Simonyan, A. Zisserman, and k. kavukcuoglu, “Spatial transformer networks,” in Advances in Neural Information Processing Systems (NIPS)
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in IEEE Int. Conf. on Computer Vision (ICCV)
2015
Cited alongside, same era.
C. Hazirbas, L. Ma, C. Domokos, and D. Cremers, “Fusenet: incorporating depth into semantic segmentation via fusion-based cnn architecture,” in Asian Conf. on Computer Vision (ACCV)
2016
Cited alongside, same era.
F. Yu and V. Koltun, “Multi-scale context aggregation by dilated convolutions,” in Int. Conf. on Learning Representations (ICLR)
2016
Cited alongside, same era.
2016
Later among the works it cites.
2016
Later among the works it cites.
T. Whelan, R. F. Salas-Moreno, B. Glocker, A. J. Davison, and S. Leutenegger, “ElasticFusion: Real-time dense SLAM and light source estimation,” Intl. J. of Robotics Research (IJRR)
2016
Later among the works it cites.
A. Kundu, V. Vineet, and V. Koltun, “Feature space optimization for semantic video segmentation,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition (CVPR)
2016
Later among the works it cites.
A. Handa, V. Patraucean, V. Badrinarayanan, S. Stent, and R. Cipolla, “Scenenet: Understanding real world indoor scenes with synthetic data,” in IEEE Comp. Vision and Pattern Recognition (CVPR)
2016
Later among the works it cites.
G. Lin, A. Milan, C. Shen, and I. Reid, “RefineNet: Multi-path refinement networks for high-resolution semantic segmentation,” in IEEE Computer Vision and Pattern Recognition (CVPR)
2017
Closest in time.
Y. He, W. Chiu, M. Keuper, and M. Fritz, “STD2P: Rgbd semantic segmentation using spatio-temporal data driven pooling,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition (CVPR)
2017
Closest in time.