Fetching the paper…
Reading the bibliography…
RGB-D semantic segmentation methods conventionally use two independent encoders to extract features from the RGB and depth data.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
J. Shotton, J. Winn, C. Rother, and A. Criminisi, “Textonboost for image understanding: Multi-class object recognition and segmentation by jointly modeling texture, layout, and context,” Int. J. Comput. Vis , vol. 81, no. 1, pp. 2–23, 2009
2009
Earlier work this paper cites.
P. K. Atrey, M. A. Hossain, A. El Saddik, and M. S. Kankanhalli, “Multimodal fusion for multimedia analysis: a survey,” Multimed. Syst. , vol. 16, no. 6, pp. 345–379, 2010
2010
Earlier work this paper cites.
J. Ngiam, A. Khosla, M. Kim, J. Nam, H. Lee, and A. Y. Ng, “Multimodal deep learning,” in Int. Conf. Mach. Learn. , 2011, pp. 689–696
2011
Earlier work this paper cites.
C. Farabet, C. Couprie, L. Najman, and Y. LeCun, “Learning hierarchical features for scene labeling,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 35, no. 8, pp. 1915–1929, 2012
2012
Earlier work this paper cites.
C. Couprie, C. Farabet, L. Najman, and Y. Lecun, “Indoor semantic segmentation using depth information,” in Int. Conf. Learn. Represent. , 2013
2013
Earlier work this paper cites.
A. Hermans, G. Floros, and B. Leibe, “Dense 3d semantic mapping of indoor scenes from rgb-d images,” in IEEE Int. Conf. Robot. Autom. IEEE, 2014, pp. 2631–2638
2014
Earlier work this paper cites.
A. C. Müller and S. Behnke, “Learning depth-sensitive conditional random fields for semantic segmentation of rgb-d images,” in IEEE Int. Conf. Robot. Autom. IEEE, 2014, pp. 6232–6237
2014
Earlier work this paper cites.
S. Gupta, R. Girshick, P. Arbeláez, and J. Malik, “Learning rich features from rgb-d images for object detection and segmentation,” in Eur. Conf. Comput. Vis. Springer, 2014, pp. 345–360
2014
Earlier work this paper cites.
R. K. Srivastava, K. Greff, and J. Schmidhuber, “Training very deep networks,” in Adv. Neural Inf. Process. Syst. , 2015, pp. 2377–2385
2015
Earlier work this paper cites.
M. Everingham, S. A. Eslami, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes challenge: A retrospective,” Int. J. Comput. Vis , vol. 111, no. 1, pp. 98–136, 2015
2015
Earlier work this paper cites.
C. Hazirbas, L. Ma, C. Domokos, and D. Cremers, “Fusenet: Incorporating depth into semantic segmentation via fusion-based cnn architecture,” in Asian Conf. Comput. Vis. Springer, 2016, pp. 213–228
2016
Earlier work this paper cites.
A. Valada, G. L. Oliveira, T. Brox, and W. Burgard, “Deep multispectral semantic scene understanding of forested environments using multimodal fusion,” in Int. Symp. Exp. Robot. Springer, 2016, pp. 465–477
2016
Earlier work this paper cites.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2016, pp. 3213–3223
2016
Cited alongside, same era.
J. Wang, Z. Wang, D. Tao, S. See, and G. Wang, “Learning common and specific features for rgb-d semantic segmentation with deconvolutional networks,” in Eur. Conf. Comput. Vis. Springer, 2016, pp. 664–679
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2016, pp. 770–778
2016
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Later among the works it cites.
A. Milioto, P. Lottes, and C. Stachniss, “Real-time semantic segmentation of crop and weed for precision agriculture robots leveraging background knowledge in cnns,” in IEEE Int. Conf. Robot. Autom. IEEE, 2018, pp. 2229–2235
2018
Later among the works it cites.
E. Stenborg, C. Toft, and L. Hammarstrand, “Long-term visual localization using semantically segmented images,” in IEEE Int. Conf. Robot. Autom. IEEE, 2018, pp. 6484–6490
2018
Later among the works it cites.
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam, “Encoder-decoder with atrous separable convolution for semantic image segmentation,” in Eur. Conf. Comput. Vis. , 2018, pp. 801–818
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Shelhamer, J. Long, and T. Darrell, “Fully convolutional networks for semantic segmentation,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 39, no. 4, pp. 640–651, April 2017
2017
Cited alongside, same era.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” in IEEE Conf. Comput. Vis. Pattern Recog. , July 2017, pp. 6230–6239
2017
Cited alongside, same era.
E. Romera, J. M. Alvarez, L. M. Bergasa, and R. Arroyo, “ERFNet: Efficient Residual Factorized ConvNet for Real-Time Semantic Segmentation,” IEEE Trans. Intell. Transp. Syst. , 2017
2017
Cited alongside, same era.
L. Deng, M. Yang, Y. Qian, C. Wang, and B. Wang, “CNN based semantic segmentation for urban traffic scenes using fisheye camera,” in IEEE Intell. Veh. Symp , 2017, pp. 231–236
2017
Cited alongside, same era.
Y. Cheng, R. Cai, Z. Li, X. Zhao, and K. Huang, “Locality-sensitive deconvolution networks with gated fusion for rgb-d indoor semantic segmentation,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2017, pp. 3029–3037
2017
Cited alongside, same era.
S.-J. Park, K.-S. Hong, and S. Lee, “Rdfnet: Rgb-d multi-level residual feature fusion for indoor semantic segmentation,” in IEEE Int. Conf. Comput. Vis. , 2017, pp. 4980–4989
2017
Cited alongside, same era.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner, “Scannet: Richly-annotated 3d reconstructions of indoor scenes,” in IEEE Conf. Comput. Vis. Pattern Recog. , 2017, pp. 5828–5839
2017
Cited alongside, same era.
D. Ramachandram and G. W. Taylor, “Deep multimodal learning: A survey on recent advances and trends,” IEEE Signal Process. Mag. , vol. 34, no. 6, pp. 96–108, 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
J. Ku, A. Harakeh, and S. L. Waslander, “In defense of classical image processing: Fast depth completion on the cpu,” in Conf. Comput. Robot Vis. IEEE, 2018, pp. 16–22
2018
Later among the works it cites.
A. Dai and M. Nießner, “3dmv: Joint 3d-multi-view prediction for 3d semantic scene segmentation,” in Eur. Conf. Comput. Vis. , 2018, pp. 452–468
2018
Later among the works it cites.
Y. Chen, M. Yang, C. Wang, and B. Wang, “3d semantic modelling with label correction for extensive outdoor scene,” in IEEE Intell. Veh. Symp . IEEE, 2019, pp. 1262–1267
2019
Closest in time.
L. Deng, M. Yang, B. Hu, T. Li, H. Li, and C. Wang, “Semantic segmentation-based lane-level localization using around view monitoring system,” IEEE Sensors J. , 2019
2019
Closest in time.
A. Valada, R. Mohan, and W. Burgard, “Self-supervised model adaptation for multimodal semantic segmentation,” Int. J. Comput. Vis , jul 2019
2019
Closest in time.
L. Deng, M. Yang, H. Li, T. Li, B. Hu, and C. Wang, “Restricted deformable convolution based road scene semantic segmentation using surround view cameras,” IEEE Trans. Intell. Transp. Syst. , 2019
2019
Closest in time.
K. Yang, X. Hu, L. M. Bergasa, E. Romera, X. Huang, D. Sun, and K. Wang, “Can we pass beyond the field of view? panoramic annular semantic segmentation for real-world surrounding perception,” in IEEE Intell. Veh. Symp . IEEE, 2019, pp. 446–453
2019
Closest in time.