Fetching the paper…
Reading the bibliography…
Learning to reliably perceive and understand the scene is an integral enabler for robots to operate in the real-world.
He K, Zhang X, Ren S, Sun J (2015c) Spatial pyramid pooling in deep convolutional networks for visual recognition. IEEE transactions on pattern analysis and machine intelligence 37(9):1904–1916
1916
Earlier work this paper cites.
LeCun Y, Denker JS, Solla SA (1990) Optimal brain damage. In: Advances in neural information processing systems, pp 598–605
1990
Earlier work this paper cites.
Huete A, Justice C, Van Leeuwen W (1999) Modis vegetation index (mod13). Algorithm theoretical basis document 3:213
1999
Earlier work this paper cites.
Running SW, Nemani R, Glassy JM, Thornton PE (1999) Modis daily photosynthesis (psn) and annual net primary production (npp) product (mod17) algorithm theoretical basis document. University of Montana, SCF At-Launch Algorithm ATBD Documents
1999
Earlier work this paper cites.
Fei-Fei L, Koch C, Iyer A, Perona P (2004) What do we see when we glance at a scene? Journal of Vision 4(8):863–863
2004
Earlier work this paper cites.
Brostow GJ, Shotton J, Fauqueur J, Cipolla R (2008) Segmentation and recognition using structure from motion point clouds. In: Forsyth D, Torr P, Zisserman A (eds) Proceedings of the European Conference on Computer Vision
2008
Earlier work this paper cites.
Shotton J, Johnson M, Cipolla R (2008) Semantic texton forests for image categorization and segmentation. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2008
Earlier work this paper cites.
Deng J, Dong W, Socher R, Li LJ, Li K, Fei-Fei L (2009) Imagenet: A large-scale hierarchical image database. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2009
Earlier work this paper cites.
Fulkerson B, Vedaldi A, Soatto S (2009) Class segmentation and object localization with superpixel neighborhoods. In: Proceedings of the International Conference on Computer Vision
2009
Earlier work this paper cites.
Grangier D, Bottou L, , Collobert R (2009) Deep convolutional networks for scene parsing. In: ICML Workshop on Deep Learning
2009
Earlier work this paper cites.
Kohli P, Torr PH, et al (2009) Robust higher order potentials for enforcing label consistency. International Journal of Computer Vision 82(3):302–324
2009
Earlier work this paper cites.
Plath N, Toussaint M, Nakajima S (2009) Multi-class image segmentation using conditional random fields and global classification. In: Proceedings of the International Conference on Machine Learning
2009
Earlier work this paper cites.
Sturgess P, Alahari K, Ladicky L, Torr PHS (2009) Combining Appearance and Structure from Motion Features for Road Scene Understanding. In: Proceedings of the British Machine Vision Conference
2009
Earlier work this paper cites.
Zhang C, Wang L, Yang R (2010) Semantic segmentation of urban scenes using dense depth maps. In: Daniilidis K, Maragos P, Paragios N (eds) Proceedings of the European Conference on Computer Vision
2010
Earlier work this paper cites.
Farabet C, Couprie C, Najman L, LeCun Y (2012) Scene parsing with multiscale feature learning, purity trees, and optimal covers. In: Proceedings of the International Conference on Machine Learning
2012
Earlier work this paper cites.
Krizhevsky A, Sutskever I, Hinton GE (2012) Imagenet classification with deep convolutional neural networks. In: Advances in neural information processing systems, pp 1097–1105
2012
Earlier work this paper cites.
Munoz D, Bagnell JA, Hebert M (2012) Co-inference for multi-modal scene analysis. In: Proceedings of the European Conference on Computer Vision
2012
Earlier work this paper cites.
Ren X, Bo L, Fox D (2012) Rgb-(d) scene labeling: Features and algorithms. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2012
Earlier work this paper cites.
Silberman N, Hoiem D, Kohli P, Fergus R (2012) Indoor segmentation and support inference from rgbd images. In: Proceedings of the European Conference on Computer Vision
2012
Earlier work this paper cites.
Couprie C, Farabet C, Najman L, LeCun Y (2013) Indoor semantic segmentation using depth information. arXiv preprint arXiv:13013572
2013
Earlier work this paper cites.
Janoch A, Karayev S, Jia Y, Barron JT, Fritz M, Saenko K, Darrell T (2013) A category-level 3d object dataset: Putting the kinect to work. In: Proceedings of the IEEE International Conference on Consumer Depth Cameras for Computer Vision, pp 141–165
2013
Earlier work this paper cites.
Lin M, Chen Q, Yan S (2013) Network in network. arXiv preprint arXiv:13124400
2013
Earlier work this paper cites.
Xiao J, Owens A, Torralba A (2013) Sun3d: A database of big spaces reconstructed using sfm and object labels. In: Proceedings of the International Conference on Computer Vision
2013
Earlier work this paper cites.
Gupta S, Girshick R, Arbeláez P, Malik J (2014) Learning rich features from rgb-d images for object detection and segmentation. In: Proceedings of the European Conference on Computer Vision
2014
Earlier work this paper cites.
Hermans A, Floros G, Leibe B (2014) Dense 3d semantic mapping of indoor scenes from rgb-d images. In: Proceedings of the IEEE International Conference on Robotics and Automation
2014
Earlier work this paper cites.
Pinheiro PO, Collobert R (2014) Recurrent convolutional neural networks for scene labeling. In: Proceedings of the International Conference on Machine Learning
2014
Earlier work this paper cites.
Simonyan K, Zisserman A (2014) Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:14091556
2014
Earlier work this paper cites.
Zhou B, Khosla A, Lapedriza A, Oliva A, Torralba A (2014) Object detectors emerge in deep scene cnns. arXiv preprint arXiv:14126856
2014
Cited alongside, same era.
Abadi M, Agarwal A, Barham P, Brevdo E, Chen Z, Citro C, Corrado GS, Davis A, Dean J, Devin M, Ghemawat S, Goodfellow I, Harp A, Irving G, Isard M, Jia Y, Jozefowicz R, Kaiser L, Kudlur M, Levenberg J, Mané D, Monga R, Moore S, Murray D, Olah C, Schuster M, Shlens J, Steiner B, Sutskever I, Talwar K, Tucker P, Vanhoucke V, Vasudevan V, Viégas F, Vinyals O, Warden P, Wattenberg M, Wicke M, Yu Y, Zheng X (2015) TensorFlow: Large-scale machine learning on heterogeneous systems. Software available from tensorflow.org
2015
Cited alongside, same era.
Badrinarayanan V, Kendall A, Cipolla R (2015) Segnet: A deep convolutional encoder-decoder architecture for image segmentation. arXiv preprint arXiv: 151100561
2015
Cited alongside, same era.
Eitel A, Springenberg JT, Spinello L, Riedmiller MA, Burgard W (2015) Multimodal deep learning for robust rgb-d object recognition. In: Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems
Chen LC, Papandreou G, Schroff F, Adam H (2017) Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:170605587
2017
Later among the works it cites.
Dai A, Chang AX, Savva M, Halber M, Funkhouser T, Nießner M (2017) Scannet: Richly-annotated 3d reconstructions of indoor scenes. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2017
Later among the works it cites.
Hu J, Shen L, Sun G (2017) Squeeze-and-excitation networks. arXiv preprint arXiv:170901507
2017
Later among the works it cites.
Kim DK, Maturana D, Uenoyama M, Scherer S (2017) Season-invariant semantic segmentation with a deep multimodal network. In: Field and Service Robotics
2017
Later among the works it cites.
Li H, Kadav A, Durdanovic I, Samet H, Graf HP (2017) Pruning filters for efficient convnets. Proceedings of the International Conference on Learning Representations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
Everingham M, Eslami SA, Van Gool L, Williams CK, Winn J, Zisserman A (2015) The pascal visual object classes challenge: A retrospective. International journal of computer vision 111(1):98–136
2015
Cited alongside, same era.
Lee CY, Xie S, Gallagher P, Zhang Z, Tu Z (2015) Deeply-supervised nets. In: Artificial Intelligence and Statistics, pp 562–570
2015
Cited alongside, same era.
Liang-Chieh C, Papandreou G, Kokkinos I, Murphy K, Yuille A (2015) Semantic image segmentation with deep convolutional nets and fully connected crfs. In: International Conference on Learning Representations
2015
Cited alongside, same era.
Liu W, Rabinovich A, Berg AC (2015) Parsenet: Looking wider to see better. arXiv preprint arXiv: 150604579
2015
Cited alongside, same era.
Long J, Shelhamer E, Darrell T (2015) Fully convolutional networks for semantic segmentation. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2015
Cited alongside, same era.
Noh H, Hong S, Han B (2015) Learning deconvolution network for semantic segmentation. In: Proceedings of the International Conference on Computer Vision, pp 1520–1528
2015
Cited alongside, same era.
Ronneberger O, Fischer P, Brox T (2015) U-net: Convolutional networks for biomedical image segmentation. In: International Conference on Medical image computing and computer-assisted intervention, pp 234–241
2015
Cited alongside, same era.
Song S, Lichtenberg SP, Xiao J (2015) Sun rgb-d: A rgb-d scene understanding benchmark suite. In: Proceedings of the Conference on Computer Vision and Pattern Recognition, vol 5, p 6
2015
Cited alongside, same era.
2017
Later among the works it cites.
Lin G, Milan A, Shen C, Reid ID (2017) Refinenet: Multi-path refinement networks for high-resolution semantic segmentation. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2017
Later among the works it cites.
Liu Z, Li J, Shen Z, Huang G, Yan S, Zhang C (2017) Learning efficient convolutional networks through network slimming. In: Proceedings of the International Conference on Computer Vision
2017
Later among the works it cites.
Molchanov P, Tyree S, Karras T, Aila T, Kautz J (2017) Pruning convolutional neural networks for resource efficient inference. Proceedings of the International Conference on Learning Representation
2017
Later among the works it cites.
Noh H, Araujo A, Sim J, Weyand T, Han B (2017) Largescale image retrieval with attentive deep local features. In: Proceedings of the IEEE International Conference on Computer Vision, pp 3456–3465
2017
Later among the works it cites.
Qi X, Liao R, Jia J, Fidler S, Urtasun R (2017) 3d graph neural networks for rgbd semantic segmentation. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp 5199–5208
2017
Later among the works it cites.
Schneider L, Jasch M, Fröhlich B, Weber T, Franke U, Pollefeys M, Rätsch M (2017) Multimodal neural networks: Rgb-d for semantic segmentation and object detection. In: Sharma P, Bianchi FM (eds) Image Analysis, Cham, pp 98–109
2017
Later among the works it cites.
Valada A, Vertens J, Dhall A, Burgard W (2017) Adapnet: Adaptive semantic segmentation in adverse environmental conditions. In: Proceedings of the IEEE International Conference on Robotics and Automation
2017
Later among the works it cites.
Xiang Y, Fox D (2017) Da-rnn: Semantic mapping with data associated recurrent neural networks. arXiv preprint arXiv:170303098
2017
Later among the works it cites.
Xie S, Girshick R, Dollár P, Tu Z, He K (2017) Aggregated residual transformations for deep neural networks. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2017
Later among the works it cites.
Zhao H, Shi J, Qi X, Wang X, Jia J (2017) Pyramid scene parsing network. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2017
Later among the works it cites.
Audebert N, Le Saux B, Lefèvre S (2018) Beyond rgb: Very high resolution urban remote sensing with multimodal deep networks. ISPRS Journal of Photogrammetry and Remote Sensing 140:20–32
2018
Closest in time.
Bulò SR, Porzi L, Kontschieder P (2018) In-place activated batchnorm for memory-optimized training of dnns. In: Proceedings of the Conference on Computer Vision and Pattern Recognition
2018
Closest in time.
Dai A, Nießner M (2018) 3dmv: Joint 3d-multi-view prediction for 3d semantic scene segmentation. arXiv preprint arXiv:180310409
2018
Closest in time.
Hu J, Shen L, Sun G (2018) Squeeze-and-excitation networks. In: Proceedings of the Conference on Computer Vision and Pattern Recognition, pp 7132–7141
2018
Closest in time.
Ku J, Harakeh A, Waslander SL (2018) In defense of classical image processing: Fast depth completion on the cpu. arXiv preprint arXiv:180200036
2018
Closest in time.
Romera E, Alvarez JM, Bergasa LM, Arroyo R (2018) Erfnet: Efficient residual factorized convnet for real-time semantic segmentation. IEEE Transactions on Intelligent Transportation Systems 19(1):263–272
2018
Closest in time.
Sandler M, Howard A, Zhu M, Zhmoginov A, Chen LC (2018) Mobilenetv2: Inverted residuals and linear bottlenecks. In: Proceedings of the Conference on Computer Vision and Pattern Recognition, pp 4510–4520
2018
Closest in time.
Yang M, Yu K, Zhang C, Li Z, Yang K (2018) Denseaspp for semantic segmentation in street scenes. In: Proceedings of the Conference on Computer Vision and Pattern Recognition, pp 3684–3692
2018
Closest in time.
2018
Closest in time.
Boniardi F, Valada A, Mohan R, Caselitz T, Burgard W (2019) Robot localization in floor plans using a room layout edge extraction network. arXiv preprint arXiv:190301804
2019
Closest in time.
Wen W, Wu C, Wang Y, Chen Y, Li H (2016) Learning structured sparsity in deep neural networks. In: Advances in Neural Information Processing Systems, pp 2074–2082
2082
Closest in time.