Fetching the paper…
Reading the bibliography…
Scene understanding based on image segmentation is a crucial component of autonomous vehicles.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from RGBD images,” in ECCV , 2012
2012
Earlier work this paper cites.
S. Gupta, R. Girshick, P. Arbeláez, and J. Malik, “Learning rich features from RGB-D images for object detection and segmentation,” in ECCV , 2014
2014
Earlier work this paper cites.
S. Song, S. P. Lichtenberg, and J. Xiao, “SUN RGB-D: A RGB-D scene understanding benchmark suite,” in CVPR , 2015
2015
Earlier work this paper cites.
O. Russakovsky et al. , “ImageNet large scale visual recognition challenge,” IJCV , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in ICLR , 2015
2015
Earlier work this paper cites.
M. Cordts et al. , “The cityscapes dataset for semantic urban scene understanding,” in CVPR , 2016
2016
Earlier work this paper cites.
C. Hazirbas, L. Ma, C. Domokos, and D. Cremers, “FuseNet: Incorporating depth into semantic segmentation via fusion-based CNN architecture,” in ACCV , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” in CVPR , 2017
2017
Earlier work this paper cites.
Q. Ha, K. Watanabe, T. Karasawa, Y. Ushiku, and T. Harada, “MFNet: Towards real-time semantic segmentation for autonomous vehicles with multi-spectral scenes,” in IROS , 2017
2017
Earlier work this paper cites.
A. Vaswani et al. , “Attention is all you need,” in NeurIPS , 2017
2017
Earlier work this paper cites.
L. Chen et al. , “SCA-CNN: Spatial and channel-wise attention in convolutional networks for image captioning,” in CVPR , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner, “ScanNet: Richly-annotated 3D reconstructions of indoor scenes,” in CVPR , 2017
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “CARLA: An open urban driving simulator,” in CoRL , 2017
2017
Earlier work this paper cites.
X. Qi, R. Liao, J. Jia, S. Fidler, and R. Urtasun, “3D graph neural networks for RGBD semantic segmentation,” in ICCV , 2017
2017
Earlier work this paper cites.
Y. Cheng, R. Cai, Z. Li, X. Zhao, and K. Huang, “Locality-sensitive deconvolution networks with gated fusion for RGB-D indoor semantic segmentation,” in CVPR , 2017
2017
Earlier work this paper cites.
D. Lin, G. Chen, D. Cohen-Or, P.-A. Heng, and H. Huang, “Cascaded feature network for semantic segmentation of RGB-D images,” in ICCV , 2017
2017
Earlier work this paper cites.
S.-J. Park, K.-S. Hong, and S. Lee, “RDFNet: RGB-D multi-level residual feature fusion for indoor semantic segmentation,” in ICCV , 2017
2017
Earlier work this paper cites.
T. Pohlen, A. Hermans, M. Mathias, and B. Leibe, “Full-resolution residual networks for semantic segmentation in street scenes,” in CVPR , 2017
2017
Earlier work this paper cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs,” TPAMI , vol. 40, no. 4, pp. 834–848, 2018
2018
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in CVPR , 2018
2018
Earlier work this paper cites.
W. Wang and U. Neumann, “Depth-aware CNN for RGB-D segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
S. Kong and C. C. Fowlkes, “Recurrent scene parsing with perspective understanding in the loop,” in CVPR , 2018
2018
Earlier work this paper cites.
A. Dai and M. Nießner, “3DMV: Joint 3D-multi-view prediction for 3D semantic scene segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
D. Xu, W. Ouyang, X. Wang, and N. Sebe, “PAD-net: Multi-tasks guided prediction-and-distillation network for simultaneous depth estimation and scene parsing,” in CVPR , 2018
2018
Earlier work this paper cites.
E. Romera, J. M. Alvarez, L. M. Bergasa, and R. Arroyo, “ERFNet: Efficient residual factorized ConvNet for real-time semantic segmentation,” T-ITS , vol. 19, no. 1, pp. 263–272, 2018
2018
Earlier work this paper cites.
C. Yu, J. Wang, C. Peng, C. Gao, G. Yu, and N. Sang, “Learning a discriminative feature network for semantic segmentation,” in CVPR , 2018
2018
Earlier work this paper cites.
C. Yu, J. Wang, C. Peng, C. Gao, G. Yu, and N. Sang, “BiSeNet: Bilateral segmentation network for real-time semantic segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam, “Encoder-decoder with atrous separable convolution for semantic image segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
T. Xiao, Y. Liu, B. Zhou, Y. Jiang, and J. Sun, “Unified perceptual parsing for scene understanding,” in ECCV , 2018
2018
Earlier work this paper cites.
J. Fu et al. , “Dual attention network for scene segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
X. Hu, K. Yang, L. Fei, and K. Wang, “ACNet: Attention based network to exploit complementary features for RGBD semantic segmentation,” in ICIP , 2019
2019
Earlier work this paper cites.
D. Sun, X. Huang, and K. Yang, “A multimodal vision sensor for autonomous driving,” in SPIE , 2019
2019
Earlier work this paper cites.
Z. Huang, X. Wang, L. Huang, C. Huang, Y. Wei, and W. Liu, “CCNet: Criss-cross attention for semantic segmentation,” in ICCV , 2019
2019
Earlier work this paper cites.
Y. Sun, W. Zuo, and M. Liu, “RTFNet: RGB-thermal fusion network for semantic segmentation of urban scenes,” RA-L , vol. 4, no. 3, pp. 2576–2583, 2019
2019
Cited alongside, same era.
Z. Zhang, Z. Cui, C. Xu, Y. Yan, N. Sebe, and J. Yang, “Pattern-affinitive propagation across depth, surface normal and semantic segmentation,” in CVPR , 2019
2019
Cited alongside, same era.
P. Zhang, W. Liu, Y. Lei, and H. Lu, “Hyperfusion-net: Hyper-densely reflective feature fusion for salient object detection,” PR , vol. 93, pp. 521–533, 2019
2019
Cited alongside, same era.
I. Alonso and A. C. Murillo, “EV-SegNet: Semantic segmentation for event-based cameras,” in CVPRW , 2019
2019
Cited alongside, same era.
W. Wang et al. , “Pyramid vision transformer: A versatile backbone for dense prediction without convolutions,” in ICCV , 2021
2021
Later among the works it cites.
Y. Yuan et al. , “HRFormer: High-resolution transformer for dense prediction,” in NeurIPS , 2021
2021
Later among the works it cites.
Y. Sun, W. Zuo, P. Yun, H. Wang, and M. Liu, “FuseSeg: Semantic segmentation of urban scenes based on RGB and thermal data fusion,” T-ASE , vol. 18, no. 3, pp. 1000–1011, 2021
2021
Later among the works it cites.
W. Zhou, J. Liu, J. Lei, L. Yu, and J.-N. Hwang, “GMNet: Graded-feature multilabel-learning network for RGB-thermal urban scene semantic segmentation,” TIP , vol. 30, pp. 7790–7802, 2021
2021
Later among the works it cites.
Z. Shen, M. Zhang, H. Zhao, S. Yi, and H. Li, “Efficient attention: Attention with linear complexities,” in WACV , 2021
2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
A. Valada, R. Mohan, and W. Burgard, “Self-supervised model adaptation for multimodal semantic segmentation,” IJCV , vol. 128, no. 5, pp. 1239–1285, 2019
2019
Cited alongside, same era.
M. Orsic, I. Kreso, P. Bevandic, and S. Segvic, “In defense of pre-trained ImageNet architectures for real-time semantic segmentation of road-driving images,” in CVPR , 2019
2019
Cited alongside, same era.
T. Takikawa, D. Acuna, V. Jampani, and S. Fidler, “Gated-SCNN: Gated shape CNNs for semantic segmentation,” in ICCV , 2019
2019
Cited alongside, same era.
F. Zhang et al. , “ACFNet: Attentional class feature network for semantic segmentation,” in ICCV , 2019
2019
Cited alongside, same era.
R. P. K. Poudel, S. Liwicki, and R. Cipolla, “Fast-SCNN: Fast semantic segmentation network,” in BMVC , 2019
2019
Cited alongside, same era.
W. Zhou, J. S. Berrio, S. Worrall, and E. Nebot, “Automated evaluation of semantic segmentation robustness for autonomous driving,” T-ITS , vol. 21, no. 5, pp. 1951–1963, 2020
2020
Cited alongside, same era.
L. Sun, K. Yang, X. Hu, W. Hu, and K. Wang, “Real-time fusion network for RGB-D semantic segmentation incorporating unexpected obstacle detection for road-driving images,” RA-L , vol. 5, no. 4, pp. 5558–5565, 2020
2020
Cited alongside, same era.
Later among the works it cites.
2021
Later among the works it cites.
R. Yan, K. Yang, and K. Wang, “NLFNet: Non-local fusion towards generalized multimodal semantic segmentation across RGB-depth, polarization, and thermal images,” in ROBIO , 2021
2021
Later among the works it cites.
G. Zhang, J.-H. Xue, P. Xie, S. Yang, and G. Wang, “Non-local aggregation for RGB-D semantic segmentation,” SPL , vol. 28, pp. 658–662, 2021
2021
Later among the works it cites.
Y. Yue, W. Zhou, J. Lei, and L. Yu, “Two-stage cascaded decoder for semantic segmentation of RGB-D images,” SPL , vol. 28, pp. 1115–1119, 2021
2021
Later among the works it cites.
D. Seichter, M. Köhler, B. Lewandowski, T. Wengefeld, and H.-M. Gross, “Efficient RGB-D semantic segmentation for indoor scene analysis,” in ICRA , 2021
2021
Later among the works it cites.
W. Hu, H. Zhao, L. Jiang, J. Jia, and T.-T. Wong, “Bidirectional projection network for cross dimension scene understanding,” in CVPR , 2021
2021
Later among the works it cites.
J. Wang et al. , “Deep high-resolution representation learning for visual recognition,” TPAMI , vol. 43, no. 10, pp. 3349–3364, 2021
2021
Later among the works it cites.
J. Xu, K. Lu, and H. Wang, “Attention fusion network for multi-spectral semantic segmentation,” PRL , vol. 146, pp. 179–184, 2021
2021
Later among the works it cites.
T. Wu, S. Tang, R. Zhang, and Y. Zhang, “CGNet: A light-weight context guided network for semantic segmentation,” TIP , vol. 30, pp. 1169–1179, 2021
2021
Later among the works it cites.
J. Zhang, K. Yang, A. Constantinescu, K. Peng, K. Müller, and R. Stiefelhagen, “Trans4Trans: Efficient transformer for transparent object segmentation to help visually impaired people navigate in the real world,” in ICCVW , 2021
2021
Later among the works it cites.
A. Prakash, K. Chitta, and A. Geiger, “Multi-modal fusion transformer for end-to-end autonomous driving,” in CVPR , 2021
2021
Later among the works it cites.
K. Yang, X. Hu, Y. Fang, K. Wang, and R. Stiefelhagen, “Omnisupervised omnidirectional semantic segmentation,” T-ITS , vol. 23, no. 2, pp. 1184–1199, 2022
2022
Closest in time.
J. Zhang, K. Yang, A. Constantinescu, K. Peng, K. Müller, and R. Stiefelhagen, “Trans4Trans: Efficient transformer for transparent object and semantic scene segmentation in real-world navigation assistance,” T-ITS , vol. 23, no. 10, pp. 19 173–19 186, 2022
2022
Closest in time.
R. Girdhar, M. Singh, N. Ravi, L. van der Maaten, A. Joulin, and I. Misra, “Omnivore: A single model for many visual modalities,” in CVPR , 2022
2022
Closest in time.
T. Zhou, W. Wang, E. Konukoglu, and L. Van Gool, “Rethinking semantic segmentation: A prototype view,” in CVPR , 2022
2022
Closest in time.
Y. Zhang, B. Pang, and C. Lu, “Semantic segmentation by early region proxy,” in CVPR , 2022
2022
Closest in time.
Y. Qian, L. Deng, T. Li, C. Wang, and M. Yang, “Gated-residual block for semantic segmentation using RGB-D data,” T-ITS , vol. 23, no. 8, pp. 11 836–11 844, 2022
2022
Closest in time.
H. Zhou, L. Qi, H. Huang, X. Yang, Z. Wan, and X. Wen, “CANet: Co-attention network for RGB-D semantic segmentation,” PR , vol. 124, p. 108468, 2022
2022
Closest in time.
J. Zhang, K. Yang, and R. Stiefelhagen, “Exploring event-driven dynamic context for accident scene segmentation,” T-ITS , vol. 23, no. 3, pp. 2606–2622, 2022
2022
Closest in time.
R. Bachmann, D. Mizrahi, A. Atanov, and A. Zamir, “MultiMAE: Multi-modal multi-task masked autoencoders,” in ECCV , 2022
2022
Closest in time.
Z. Sun, N. Messikommer, D. Gehrig, and D. Scaramuzza, “ESS: Learning event-based semantic segmentation from still images,” in ECCV , 2022
2022
Closest in time.
W. Shi et al. , “RGB-D semantic segmentation and label-oriented voxelgrid fusion for accurate 3D semantic mapping,” TCSVT , vol. 32, no. 1, pp. 183–197, 2022
2022
Closest in time.
Y. Wang, X. Chen, L. Cao, W. Huang, F. Sun, and Y. Wang, “Multimodal token fusion for vision transformers,” in CVPR , 2022
2022
Closest in time.
Y. Liao, J. Xie, and A. Geiger, “KITTI-360: A novel dataset and benchmarks for urban scene understanding in 2D and 3D,” TPAMI , vol. 45, no. 3, pp. 3292–3310, 2023
2023
Closest in time.
F. Lin, Z. Liang, J. He, M. Zheng, S. Tian, and K. Chen, “StructToken : Rethinking semantic segmentation with structural prior,” TCSVT , 2023
2023
Closest in time.
Y. Pang, X. Zhao, L. Zhang, and H. Lu, “CAVER: Cross-modal view-mixed transformer for bi-modal salient object detection,” TIP , 2023
2023
Closest in time.
X. Zhang, S. Zhang, Z. Cui, Z. Li, J. Xie, and J. Yang, “Tube-embedded transformer for pixel prediction,” TMM , vol. 25, pp. 2503–2514, 2023
2023
Closest in time.
Y. Cai, W. Zhou, L. Zhang, L. Yu, and T. Luo, “DHFNet: Dual-decoding hierarchical fusion network for RGB-thermal semantic segmentation,” The Visual Computer , pp. 1–11, 2023
2023
Closest in time.
T. Broedermann, C. Sakaridis, D. Dai, and L. Van Gool, “HRFuser: A multi-resolution sensor fusion architecture for 2D object detection,” in ITSC , 2023
2023
Closest in time.