Fetching the paper…
Reading the bibliography…
Semantic segmentation benchmarks in the realm of autonomous driving are dominated by large pre-trained transformers, yet their widespread adoption is impeded by substantial computational costs and prolonged training durations.
L. Van der Maaten and G. Hinton, “Visualizing data using t-SNE,” JMLR , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: A large-scale hierarchical image database,” in CVPR , 2009
2009
Earlier work this paper cites.
M. Everingham, L. V. Gool, C. K. I. Williams, J. M. Winn, and A. Zisserman, “The pascal visual object classes (VOC) challenge,” IJCV , vol. 88, pp. 303–338, 2010
2010
Earlier work this paper cites.
B. Hariharan, P. Arbeláez, L. Bourdev, S. Maji, and J. Malik, “Semantic contours from inverse detectors,” in ICCV , 2011
2011
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from RGBD images,” in ECCV , 2012
2012
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in CVPR , 2015
2015
Earlier work this paper cites.
M. Cordts et al. , “The cityscapes dataset for semantic urban scene understanding,” in CVPR , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Vaswani et al. , “Attention is all you need,” in NeurIPS , 2017
2017
Earlier work this paper cites.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” in CVPR , 2017
2017
Earlier work this paper cites.
V. Badrinarayanan, A. Kendall, and R. Cipolla, “SegNet: A deep convolutional encoder-decoder architecture for image segmentation,” TPAMI , vol. 39, no. 12, pp. 2481–2495, 2017
2017
Earlier work this paper cites.
E. Romera, J. M. Alvarez, L. M. Bergasa, and R. Arroyo, “ERFNet: Efficient residual factorized ConvNet for real-time semantic segmentation,” T-ITS , vol. 19, no. 1, pp. 263–272, 2018
2018
Earlier work this paper cites.
S. Mehta, M. Rastegari, A. Caspi, L. Shapiro, and H. Hajishirzi, “ESPNet: Efficient spatial pyramid of dilated convolutions for semantic segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam, “Encoder-decoder with atrous separable convolution for semantic image segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in CVPR , 2018
2018
Earlier work this paper cites.
H. Zhao, X. Qi, X. Shen, J. Shi, and J. Jia, “ICNet for real-time semantic segmentation on high-resolution images,” in ECCV , 2018
2018
Earlier work this paper cites.
C. Yu, J. Wang, C. Peng, C. Gao, G. Yu, and N. Sang, “BiSeNet: Bilateral segmentation network for real-time semantic segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “MobileNetV2: Inverted residuals and linear bottlenecks,” in CVPR , 2018
2018
Earlier work this paper cites.
N. Ma, X. Zhang, H.-T. Zheng, and J. Sun, “ShuffleNet V2: Practical guidelines for efficient CNN architecture design,” in ECCV , 2018
2018
Earlier work this paper cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in CVPR , 2018
2018
Earlier work this paper cites.
H. Zhang et al. , “Context encoding for semantic segmentation,” in CVPR , 2018
2018
Earlier work this paper cites.
R. P. Poudel, U. Bonde, S. Liwicki, and C. Zach, “ContextNet: Exploring context and detail for semantic segmentation in real-time,” in BMVC , 2018
2018
Earlier work this paper cites.
J. Xie, B. Shuai, J.-F. Hu, J. Lin, and W.-S. Zheng, “Improving fast segmentation with teacher-student learning,” in BMVC , 2018
2018
Earlier work this paper cites.
Y. Liu, K. Chen, C. Liu, Z. Qin, Z. Luo, and J. Wang, “Structured knowledge distillation for semantic segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
J. Fu et al. , “Dual attention network for scene segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
M. Orsic, I. Kreso, P. Bevandic, and S. Segvic, “In defense of pre-trained ImageNet architectures for real-time semantic segmentation of road-driving images,” in CVPR , 2019
2019
Earlier work this paper cites.
H. Li, P. Xiong, H. Fan, and J. Sun, “DFANet: Deep feature aggregation for real-time semantic segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
T. He, C. Shen, Z. Tian, D. Gong, C. Sun, and Y. Yan, “Knowledge adaptation for efficient semantic segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
X. Li, W. Wang, X. Hu, and J. Yang, “Selective kernel networks,” in CVPR , 2019
2019
Earlier work this paper cites.
I. Loshchilov and F. Hutter, “Decoupled weight decay regularization,” in ICLR , 2019
2019
Cited alongside, same era.
P. Mishra and K. Sarawadekar, “Polynomial learning rate policy with warm restart for deep neural network,” in TENCON , 2019
2019
Cited alongside, same era.
Z. Huang, X. Wang, L. Huang, C. Huang, Y. Wei, and W. Liu, “CCNet: Criss-cross attention for semantic segmentation,” in ICCV , 2019
2019
Cited alongside, same era.
R. P. K. Poudel, S. Liwicki, and R. Cipolla, “Fast-SCNN: Fast semantic segmentation network,” in BMVC , 2019
2019
Cited alongside, same era.
M. Tan and Q. Le, “EfficientNet: Rethinking model scaling for convolutional neural networks,” in ICML , 2019
2019
Cited alongside, same era.
L. Liu et al. , “Exploring inter-channel correlation for diversity-preserved knowledge distillation,” in ICCV , 2021
2021
Later among the works it cites.
C. Dong, G. Wang, H. Xu, J. Peng, X. Ren, and X. Liang, “EfficientBERT: Progressively searching multilayer perceptron via warm-up knowledge distillation,” in EMNLP , 2021
2021
Later among the works it cites.
B. Cheng, A. Schwing, and A. Kirillov, “Per-pixel classification is not all you need for semantic segmentation,” in NeurIPS , 2021
2021
Later among the works it cites.
T. Wu, S. Tang, R. Zhang, J. Cao, and Y. Zhang, “CGNet: A light-weight context guided network for semantic segmentation,” TIP , vol. 30, pp. 1169–1179, 2021
2021
Later among the works it cites.
K. Muhammad et al. , “Vision-based semantic segmentation in scene understanding for autonomous driving: Recent achievements, challenges, and outlooks,” T-ITS , vol. 23, no. 12, pp. 22 694–22 715, 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Wang, W. Zhou, T. Jiang, X. Bai, and Y. Xu, “Intra-class feature variation distillation for semantic segmentation,” in ECCV , 2020
2020
Cited alongside, same era.
S. I. Mirzadeh, M. Farajtabar, A. Li, N. Levine, A. Matsukawa, and H. Ghasemzadeh, “Improved knowledge distillation via teacher assistant,” in AAAI , 2020
2020
Cited alongside, same era.
X. Jiao et al. , “TinyBERT: Distilling BERT for natural language understanding,” in EMNLP , 2020
2020
Cited alongside, same era.
W. Wang, F. Wei, L. Dong, H. Bao, N. Yang, and M. Zhou, “MiniLM: Deep self-attention distillation for task-agnostic compression of pre-trained transformers,” in NeurIPS , 2020
2020
Cited alongside, same era.
S. Choi, J. T. Kim, and J. Choo, “Cars can’t fly up in the sky: Improving urban-scene segmentation via height-driven attention networks,” in CVPR , 2020
2020
Cited alongside, same era.
Y. Yuan, X. Chen, and J. Wang, “Object-contextual representations for semantic segmentation,” in ECCV , 2020
2020
Cited alongside, same era.
C. Shu, Y. Liu, J. Gao, Z. Yan, and C. Shen, “Channel-wise knowledge distillation for dense prediction,” in ICCV , 2021
2021
Cited alongside, same era.
2022
Closest in time.
J. Zhang, K. Yang, A. Constantinescu, K. Peng, K. Müller, and R. Stiefelhagen, “Trans4Trans: Efficient transformer for transparent object and semantic scene segmentation in real-world navigation assistance,” T-ITS , vol. 23, no. 10, pp. 19 173–19 186, 2022
2022
Closest in time.
J. Zhang, K. Yang, and R. Stiefelhagen, “Exploring event-driven dynamic context for accident scene segmentation,” T-ITS , vol. 23, no. 3, pp. 2606–2622, 2022
2022
Closest in time.
D. Ji, H. Wang, M. Tao, J. Huang, X.-S. Hua, and H. Lu, “Structural and statistical texture knowledge distillation for semantic segmentation,” in CVPR , 2022
2022
Closest in time.
S. An, Q. Liao, Z. Lu, and J.-H. Xue, “Efficient semantic segmentation via self-attention and self-distillation,” T-ITS , vol. 23, no. 9, pp. 15 256–15 266, 2022
2022
Closest in time.
W. Wang et al. , “PVT v2: Improved baselines with pyramid vision transformer,” CVM , vol. 8, no. 3, pp. 415–424, 2022
2022
Closest in time.
T. Huang, S. You, F. Wang, C. Qian, and C. Xu, “Knowledge distillation from a stronger teacher,” in NeurIPS , 2022
2022
Closest in time.
C. Yang et al. , “Lite vision transformer with enhanced self-attention,” in CVPR , 2022
2022
Closest in time.
W. Yu et al. , “MetaFormer is actually what you need for vision,” in CVPR , 2022
2022
Closest in time.
W. Wang et al. , “CrossFormer: A versatile vision transformer hinging on cross-scale attention,” in ICLR , 2022
2022
Closest in time.
K. Wu et al. , “TinyViT: Fast pretraining distillation for small vision transformers,” in ECCV , 2022
2022
Closest in time.
S. Mehta and M. Rastegari, “MobileViT: Light-weight, general-purpose, and mobile-friendly vision transformer,” in ICLR , 2022
2022
Closest in time.
Y. Chen et al. , “Mobile-Former: Bridging MobileNet and transformer,” in CVPR , 2022
2022
Closest in time.
L. Wang and K.-J. Yoon, “Knowledge distillation and student-teacher learning for visual intelligence: A review and new outlooks,” TPAMI , vol. 44, no. 6, pp. 3048–3068, 2022
2022
Closest in time.
C. Yang, H. Zhou, Z. An, X. Jiang, Y. Xu, and Q. Zhang, “Cross-image relational knowledge distillation for semantic segmentation,” in CVPR , 2022
2022
Closest in time.
Z. Zhang, C. Zhou, and Z. Tu, “Distilling inter-class distance for semantic segmentation,” in IJCAI , 2022
2022
Closest in time.
Z. Tian et al. , “Adaptive perspective distillation for semantic segmentation,” TPAMI , 2022
2022
Closest in time.
S. Lin et al. , “Knowledge distillation via the target-aware transformer,” in CVPR , 2022
2022
Closest in time.
D. Zhou et al. , “Understanding the robustness in vision transformers,” in ICML , 2022
2022
Closest in time.
Y. Lee, J. Kim, J. Willette, and S. J. Hwang, “MPViT: Multi-path vision transformer for dense prediction,” in CVPR , 2022
2022
Closest in time.
X. Dong et al. , “CSWin transformer: A general vision transformer backbone with cross-shaped windows,” in CVPR , 2022
2022
Closest in time.
J. Zhang, H. Liu, K. Yang, X. Hu, R. Liu, and R. Stiefelhagen, “CMX: cross-modal fusion for RGB-X semantic segmentation with transformers,” T-ITS , vol. 24, no. 12, pp. 14 679–14 694, 2023
2023
Closest in time.
Y. Zheng, F. Zhou, S. Liang, W. Song, and X. Bai, “Semantic segmentation in thermal videos: A new benchmark and multi-granularity contrastive learning-based framework,” T-ITS , 2023
2023
Closest in time.
K. Li et al. , “UniFormer: Unifying convolution and self-attention for visual recognition,” TPAMI , vol. 45, no. 10, pp. 12 581–12 600, 2023
2023
Closest in time.
Z. Chen et al. , “Vision transformer adapter for dense predictions,” in ICLR , 2023
2023
Closest in time.