Fetching the paper…
Reading the bibliography…
The 3D scene understanding is mainly considered as a crucial requirement in computer vision and robotics applications.
J. Freixenet, X. Muñoz, D. Raba, J. Martí, and X. Cufí, “Yet another survey on image segmentation: Region and boundary information integration,” in European Conference on Computer Vision . Springer, 2002, pp. 408–422
2002
Earlier work this paper cites.
J. Shotton, J. Winn, C. Rother, and A. Criminisi, “Textonboost: Joint appearance, shape and context modeling for multi-class object recognition and segmentation,” in European conference on computer vision . Springer, 2006, pp. 1–15
2006
Earlier work this paper cites.
N. Silberman and R. Fergus, “Indoor scene segmentation using a structured light sensor,” in Computer Vision Workshops (ICCV Workshops), 2011 IEEE International Conference on . IEEE, 2011, pp. 601–608
2011
Earlier work this paper cites.
P. Arbelaez, M. Maire, C. Fowlkes, and J. Malik, “Contour detection and hierarchical image segmentation,” IEEE transactions on pattern analysis and machine intelligence , vol. 33, no. 5, pp. 898–916, 2011
2011
Earlier work this paper cites.
X. Ren, L. Bo, and D. Fox, “Rgb-(d) scene labeling: Features and algorithms,” in Computer Vision and Pattern Recognition (CVPR), 2012 IEEE Conference on . IEEE, 2012, pp. 2759–2766
2012
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from rgbd images,” in European Conference on Computer Vision . Springer, 2012, pp. 746–760
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
J. Han, L. Shao, D. Xu, and J. Shotton, “Enhanced computer vision with microsoft kinect sensor: A review,” IEEE transactions on cybernetics , vol. 43, no. 5, pp. 1318–1334, 2013
2013
Earlier work this paper cites.
C. Cadena and J. Košecka, “Semantic parsing for priming object detection in rgb-d scenes,” in 3rd Workshop on Semantic Perception, Mapping and Exploration . Citeseer, 2013
2013
Earlier work this paper cites.
S. Gupta, P. Arbelaez, and J. Malik, “Perceptual organization and recognition of indoor scenes from rgb-d images,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2013, pp. 564–571
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
C. Farabet, C. Couprie, L. Najman, and Y. LeCun, “Learning hierarchical features for scene labeling,” IEEE transactions on pattern analysis and machine intelligence , vol. 35, no. 8, pp. 1915–1929, 2013
2013
Earlier work this paper cites.
G. Csurka, D. Larlus, F. Perronnin, and F. Meylan, “What is a good evaluation measure for semantic segmentation?.” in BMVC , vol. 27. Citeseer, 2013, p. 2013
2013
Earlier work this paper cites.
S. Gupta, R. Girshick, P. Arbeláez, and J. Malik, “Learning rich features from rgb-d images for object detection and segmentation,” in European Conference on Computer Vision . Springer, 2014, pp. 345–360
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
A. C. Müller and S. Behnke, “Learning depth-sensitive conditional random fields for semantic segmentation of rgb-d images,” in Robotics and Automation (ICRA), 2014 IEEE International Conference on . Citeseer, 2014, pp. 6232–6237
2014
Earlier work this paper cites.
R. Girshick, J. Donahue, T. Darrell, and J. Malik, “Rich feature hierarchies for accurate object detection and semantic segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2014, pp. 580–587
2014
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in Advances in neural information processing systems , 2015, pp. 91–99
2015
Earlier work this paper cites.
L. Wang, Y. Qiao, and X. Tang, “Action recognition with trajectory-pooled deep-convolutional descriptors,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 4305–4314
2015
Earlier work this paper cites.
F. Liu, C. Shen, and G. Lin, “Deep convolutional neural fields for depth estimation from a single image,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 5162–5170
2015
Earlier work this paper cites.
D. Eigen and R. Fergus, “Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2650–2658
2015
Earlier work this paper cites.
A. Kendall, M. Grimes, and R. Cipolla, “Posenet: A convolutional network for real-time 6-dof camera relocalization,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2938–2946
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 3431–3440
2015
Cited alongside, same era.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “Semantic image segmentation with deep convolutional nets and fully connected crfs,” International Conference on Learning Representations , 2015
2015
Cited alongside, same era.
H. Noh, S. Hong, and B. Han, “Learning deconvolution network for semantic segmentation,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 1520–1528
2015
Cited alongside, same era.
F. Fooladgar and S. Kasaei, “Semantic segmentation of rgb-d images using 3d and local neighbouring features,” in Digital Image Computing: Techniques and Applications (DICTA), 2015 International Conference on . IEEE, 2015, pp. 1–7
2015
Cited alongside, same era.
R. Quan, J. Han, D. Zhang, F. Nie, X. Qian, and X. Li, “Unsupervised salient object detection via inferring from imperfect saliency models,” IEEE Transactions on Multimedia , vol. 20, no. 5, pp. 1101–1112, 2017
2017
Later among the works it cites.
K. Fu, I. Y.-H. Gu, and J. Yang, “Saliency detection by fully learning a continuous conditional random field,” IEEE Transactions on Multimedia , vol. 19, no. 7, pp. 1531–1544, 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
X. Qi, R. Liao, J. Jia, S. Fidler, and R. Urtasun, “3d graph neural networks for rgbd semantic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 5199–5208
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 1–9
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
S. Song, S. P. Lichtenberg, and J. Xiao, “Sun rgb-d: A rgb-d scene understanding benchmark suite,” in 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2015, pp. 567–576
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
B. Shuai, Z. Zuo, G. Wang, and B. Wang, “Scene parsing with integration of parametric and non-parametric models,” IEEE Transactions on Image Processing , vol. 25, no. 5, pp. 2379–2391, 2016
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Identity mappings in deep residual networks,” in European conference on computer vision . Springer, 2016, pp. 630–645
2016
Cited alongside, same era.
C. Hazirbas, L. Ma, C. Domokos, and D. Cremers, “Fusenet: Incorporating depth into semantic segmentation via fusion-based cnn architecture,” in Asian Conference on Computer Vision . Springer, 2016, pp. 213–228
2016
Cited alongside, same era.
Y. Cheng, R. Cai, Z. Li, X. Zhao, and K. Huang, “Localitysensitive deconvolution networks with gated fusion for rgb-d indoor semantic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 3, 2017
2017
Later among the works it cites.
H. Shi, H. Li, F. Meng, Q. Wu, L. Xu, and K. N. Ngan, “Hierarchical parsing net: Semantic scene parsing from global scene to objects,” IEEE Transactions on Multimedia , vol. 20, no. 10, pp. 2670–2682, 2018
2018
Later among the works it cites.
Y. Li, Y. Guo, J. Guo, Z. Ma, X. Kong, and Q. Liu, “Joint crf and locality-consistent dictionary learning for semantic segmentation,” IEEE Transactions on Multimedia , vol. 21, no. 4, pp. 875–886, 2018
2018
Later among the works it cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 4, pp. 834–848, 2018
2018
Later among the works it cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 7132–7141
2018
Later among the works it cites.
K. Fu, Q. Zhao, and I. Y.-H. Gu, “Refinet: A deep segmentation assisted refinement network for salient object detection,” IEEE Transactions on Multimedia , vol. 21, no. 2, pp. 457–469, 2018
2018
Later among the works it cites.
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam, “Encoder-decoder with atrous separable convolution for semantic image segmentation,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 801–818
2018
Later among the works it cites.
H. Liu, W. Wu, X. Wang, and Y. Qian, “Rgb-d joint modelling with scene geometric information for indoor semantic segmentation,” Multimedia Tools and Applications , pp. 1–14, 2018
2018
Later among the works it cites.
W. Wang and U. Neumann, “Depth-aware cnn for rgb-d segmentation,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 135–150
2018
Later among the works it cites.
B. Kang, Y. Lee, and T. Q. Nguyen, “Depth-adaptive deep neural network for semantic segmentation,” IEEE Transactions on Multimedia , vol. 20, no. 9, pp. 2478–2490, 2018
2018
Later among the works it cites.
S. Woo, J. Park, J.-Y. Lee, and I. So Kweon, “Cbam: Convolutional block attention module,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 3–19
2018
Later among the works it cites.
G. Lin, C. Shen, A. Van Den Hengel, and I. Reid, “Exploring context with deep structured models for semantic segmentation,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 6, pp. 1352–1366, 2018
2018
Later among the works it cites.
K. Tateno, N. Navab, and F. Tombari, “Distortion-aware convolutional filters for dense prediction in panoramic images,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 707–722
2018
Later among the works it cites.
2018
Later among the works it cites.
Q. Wang, C. Yuan, and Y. Liu, “Learning deep conditional neural network for image segmentation,” IEEE Transactions on Multimedia , 2019
2019
Closest in time.
L. Khelifi and M. Mignotte, “Mc-ssm: Nonparametric semantic image segmentation with the icm algorithm,” IEEE TRANSACTIONS ON MULTIMEDIA , vol. 21, no. 8, pp. 1946–1959, 2019
2019
Closest in time.
F. Fooladgar and S. Kasaei, “3m2rnet: Multi-modal multi-resolution refinement network for semantic segmentation,” in Science and Information Conference . Springer, 2019, pp. 544–557
2019
Closest in time.