Fetching the paper…
Reading the bibliography…
In this paper, we address panoramic semantic segmentation which is under-explored due to two critical challenges: (1) image distortions and object deformations on panoramas; (2) lack of semantic annotations in the 360{\deg} imagery.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. C. Courville, and Y. Bengio, “Generative adversarial nets,” in NeurIPS , 2014
2014
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in CVPR , 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in ICLR , 2015
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: convolutional networks for biomedical image segmentation,” in MICCAI , 2015
2015
Earlier work this paper cites.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in CVPR , 2016
2016
Earlier work this paper cites.
G. Ros, L. Sellart, J. Materzynska, D. Vazquez, and A. M. Lopez, “The SYNTHIA dataset: A large collection of synthetic images for semantic segmentation of urban scenes,” in CVPR , 2016
2016
Earlier work this paper cites.
S. R. Richter, V. Vineet, S. Roth, and V. Koltun, “Playing for data: Ground truth from computer games,” in ECCV , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
V. Badrinarayanan, A. Kendall, and R. Cipolla, “SegNet: A deep convolutional encoder-decoder architecture for image segmentation,” TPAMI , vol. 39, no. 12, pp. 2481–2495, 2017
2017
Earlier work this paper cites.
G. Lin, A. Milan, C. Shen, and I. Reid, “RefineNet: Multi-path refinement networks for high-resolution semantic segmentation,” in CVPR , 2017
2017
Earlier work this paper cites.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” in CVPR , 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017
2017
Earlier work this paper cites.
Y. Zhang, P. David, and B. Gong, “Curriculum domain adaptation for semantic segmentation of urban scenes,” in ICCV , 2017
2017
Earlier work this paper cites.
J. Dai, H. Qi, Y. Xiong, Y. Li, G. Zhang, H. Hu, and Y. Wei, “Deformable convolutional networks,” in ICCV , 2017
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “CARLA: An open urban driving simulator,” in CoRL , 2017
2017
Earlier work this paper cites.
G. P. de La Garanderie, A. A. Abarghouei, and T. P. Breckon, “Eliminating the blind spot: Adapting 3D object detection and monocular depth estimation to 360 ∘ panoramic imagery,” in ECCV , 2018
2018
Earlier work this paper cites.
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam, “Encoder-decoder with atrous separable convolution for semantic image segmentation,” in ECCV , 2018
2018
Earlier work this paper cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs,” TPAMI , vol. 40, no. 4, pp. 834–848, 2018
2018
Earlier work this paper cites.
H. Zhang, K. J. Dana, J. Shi, Z. Zhang, X. Wang, A. Tyagi, and A. Agrawal, “Context encoding for semantic segmentation,” in CVPR , 2018
2018
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in CVPR , 2018
2018
Earlier work this paper cites.
K. Tateno, N. Navab, and F. Tombari, “Distortion-aware convolutional filters for dense prediction in panoramic images,” in ECCV , 2018
2018
Earlier work this paper cites.
J. Hoffman, E. Tzeng, T. Park, J. Zhu, P. Isola, K. Saenko, A. A. Efros, and T. Darrell, “CyCADA: Cycle-consistent adversarial domain adaptation,” in ICML , 2018
2018
Earlier work this paper cites.
Y.-H. Tsai, W.-C. Hung, S. Schulter, K. Sohn, M.-H. Yang, and M. Chandraker, “Learning to adapt structured output space for semantic segmentation,” in CVPR , 2018
2018
Earlier work this paper cites.
W.-S. Lai, Y. Huang, N. Joshi, C. Buehler, M.-H. Yang, and S. B. Kang, “Semantic-driven generation of hyperlapse from 360 degree video,” TVCG , vol. 24, no. 9, pp. 2610–2621, 2018
2018
Earlier work this paper cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in CVPR , 2018
2018
Earlier work this paper cites.
E. Romera, J. M. Alvarez, L. M. Bergasa, and R. Arroyo, “ERFNet: Efficient residual factorized ConvNet for real-time semantic segmentation,” T-ITS , vol. 19, no. 1, pp. 263–272, 2018
2018
Earlier work this paper cites.
M. Xu, Y. Song, J. Wang, M. Qiao, L. Huo, and Z. Wang, “Predicting head movement in panoramic video: A deep reinforcement learning approach,” TPAMI , vol. 41, no. 11, pp. 2693–2708, 2019
2019
Earlier work this paper cites.
Y. Luo, L. Zheng, T. Guan, J. Yu, and Y. Yang, “Taking a closer look at domain shift: Category-level adversaries for semantics consistent domain adaptation,” in CVPR , 2019
2019
Earlier work this paper cites.
Y. Zou, Z. Yu, X. Liu, B. V. K. V. Kumar, and J. Wang, “Confidence regularized self-training,” in ICCV , 2019
2019
Earlier work this paper cites.
J. Fu, J. Liu, H. Tian, Y. Li, Y. Bao, Z. Fang, and H. Lu, “Dual attention network for scene segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
Z. Huang, X. Wang, L. Huang, C. Huang, Y. Wei, and W. Liu, “CCNet: Criss-cross attention for semantic segmentation,” in ICCV , 2019
2019
Earlier work this paper cites.
S. K. Yogamani, C. Witt, H. Rashed, S. Nayak, S. Mansoor, P. Varley, X. Perrotton, D. O’Dea, P. Pérez, C. Hughes, J. Horgan, G. Sistu, S. Chennupati, M. Uricár, S. Milz, M. Simon, and K. Amende, “WoodScape: A multi-task, multi-camera fisheye dataset for autonomous driving,” in ICCV , 2019
2019
Earlier work this paper cites.
Y. Xu, K. Wang, K. Yang, D. Sun, and J. Fu, “Semantic segmentation of panoramic images using a synthetic dataset,” in SPIE , 2019
2019
Earlier work this paper cites.
C. M. Jiang, J. Huang, K. Kashinath, Prabhat, P. Marcus, and M. Nießner, “Spherical CNNs on unstructured grids,” in ICLR , 2019
2019
Earlier work this paper cites.
Y. Lee, J. Jeong, J. Yun, W. Cho, and K.-J. Yoon, “SpherePHD: Applying CNNs on a spherical PolyHeDron representation of 360° images,” in CVPR , 2019
2019
Earlier work this paper cites.
W.-L. Chang, H.-P. Wang, W.-H. Peng, and W.-C. Chiu, “All about structure: Adapting structural information across domains for boosting semantic segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
Q. Lian, L. Duan, F. Lv, and B. Gong, “Constructing self-motivated pyramid curriculums for cross-domain semantic segmentation: A non-adversarial approach,” in ICCV , 2019
2019
Earlier work this paper cites.
Y. Li, L. Yuan, and N. Vasconcelos, “Bidirectional learning for domain adaptation of semantic segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
M. Chen, H. Xue, and D. Cai, “Domain adaptation for semantic segmentation with maximum squares loss,” in ICCV , 2019
2019
Earlier work this paper cites.
R. P. K. Poudel, S. Liwicki, and R. Cipolla, “Fast-SCNN: Fast semantic segmentation network,” in BMVC , 2019
2019
Earlier work this paper cites.
M. Orsic, I. Kreso, P. Bevandic, and S. Segvic, “In defense of pre-trained ImageNet architectures for real-time semantic segmentation of road-driving images,” in CVPR , 2019
2019
Earlier work this paper cites.
A. Kirillov, R. Girshick, K. He, and P. Dollár, “Panoptic feature pyramid networks,” in CVPR , 2019
2019
Earlier work this paper cites.
T. Kalluri, G. Varma, M. Chandraker, and C. V. Jawahar, “Universal semi-supervised semantic segmentation,” in ICCV , 2019
2019
Earlier work this paper cites.
L. Porzi, S. R. Bulò, A. Colovic, and P. Kontschieder, “Seamless scene segmentation,” in CVPR , 2019
2019
Earlier work this paper cites.
T. Cohen, M. Weiler, B. Kicanaoglu, and M. Welling, “Gauge equivariant convolutional networks and the icosahedral CNN,” in ICML , 2019
2019
Cited alongside, same era.
C. Zhang, S. Liwicki, W. Smith, and R. Cipolla, “Orientation-aware semantic segmentation on icosahedron spheres,” in ICCV , 2019
2019
Cited alongside, same era.
K. Yang, X. Hu, L. M. Bergasa, E. Romera, and K. Wang, “PASS: Panoramic annular semantic segmentation,” T-ITS , vol. 21, no. 10, pp. 4171–4185, 2020
2020
Cited alongside, same era.
J. Zheng, J. Zhang, J. Li, R. Tang, S. Gao, and Z. Zhou, “Structured3D: A large photo-realistic dataset for structured 3D modeling,” in ECCV , 2020
2020
Cited alongside, same era.
Q. Hou, L. Zhang, M.-M. Cheng, and J. Feng, “Strip pooling: Rethinking spatial pooling for scene parsing,” in CVPR , 2020
2020
Cited alongside, same era.
Q. Gu, Q. Zhou, M. Xu, Z. Feng, G. Cheng, X. Lu, J. Shi, and L. Ma, “PIT: Position-invariant transform for cross-FoV domain adaptation,” in ICCV , 2021
2021
Later among the works it cites.
X. Yue, Z. Zheng, S. Zhang, Y. Gao, T. Darrell, K. Keutzer, and A. L. Sangiovanni-Vincentelli, “Prototypical cross-domain self-supervised learning for few-shot unsupervised domain adaptation,” in CVPR , 2021
2021
Later among the works it cites.
P. Zhang, B. Zhang, T. Zhang, D. Chen, Y. Wang, and F. Wen, “Prototypical pseudo label denoising and target structure learning for domain adaptive semantic segmentation,” in CVPR , 2021
2021
Later among the works it cites.
P. Hu, F. Perazzi, F. C. Heilbron, O. Wang, Z. Lin, K. Saenko, and S. Sclaroff, “Real-time semantic segmentation with fast attention,” RA-L , vol. 6, no. 1, pp. 263–270, 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Yu, J. Wang, C. Gao, G. Yu, C. Shen, and N. Sang, “Context prior for scene segmentation,” in CVPR , 2020
2020
Cited alongside, same era.
L. Deng, M. Yang, H. Li, T. Li, B. Hu, and C. Wang, “Restricted deformable convolution-based road scene semantic segmentation using surround view cameras,” T-ITS , vol. 21, no. 10, pp. 4350–4362, 2020
2020
Cited alongside, same era.
J. Huang, S. Lu, D. Guan, and X. Zhang, “Contextual-relation consistent domain adaptation for semantic segmentation,” in ECCV , 2020
2020
Cited alongside, same era.
Y. Yang and S. Soatto, “FDA: Fourier domain adaptation for semantic segmentation,” in CVPR , 2020
2020
Cited alongside, same era.
Z. Wang, M. Yu, Y. Wei, R. Feris, J. Xiong, W. Hwu, T. S. Huang, and H. Shi, “Differential treatment for stuff and things: A simple unsupervised domain adaptation method for semantic segmentation,” in CVPR , 2020
2020
Cited alongside, same era.
F. Pan, I. Shin, F. Rameau, S. Lee, and I. S. Kweon, “Unsupervised intra-domain adaptation for semantic segmentation through self-supervision,” in CVPR , 2020
2020
Cited alongside, same era.
T. Chen, S. Kornblith, K. Swersky, M. Norouzi, and G. Hinton, “Big self-supervised models are strong semi-supervised learners,” in NeurIPS , 2020
2020
Cited alongside, same era.
J. Zhang, K. Yang, A. Constantinescu, K. Peng, K. Müller, and R. Stiefelhagen, “Trans4Trans: Efficient transformer for transparent object segmentation to help visually impaired people navigate in the real world,” in ICCVW , 2021
2021
Later among the works it cites.
S. Gao, K. Yang, H. Shi, K. Wang, and J. Bai, “Review on panoramic imaging and its applications in scene understanding,” TIM , vol. 71, pp. 1–34, 2022
2022
Closest in time.
Y. Xu, Z. Zhang, and S. Gao, “Spherical DNNs and their applications in 360 ∘ images and videos,” TPAMI , vol. 44, no. 10, pp. 7235–7252, 2022
2022
Closest in time.
2022
Closest in time.
J. Zhang, K. Yang, C. Ma, S. Reiß, K. Peng, and R. Stiefelhagen, “Bending reality: Distortion-aware transformers for adapting to panoramic semantic segmentation,” in CVPR , 2022
2022
Closest in time.
J. Zhang, C. Ma, K. Yang, A. Roitberg, K. Peng, and R. Stiefelhagen, “Transfer beyond the field of view: Dense panoramic semantic segmentation via unsupervised domain adaptation,” T-ITS , vol. 23, no. 7, pp. 9478–9491, 2022
2022
Closest in time.
W. Yu, M. Luo, P. Zhou, C. Si, Y. Zhou, X. Wang, J. Feng, and S. Yan, “MetaFormer is actually what you need for vision,” in CVPR , 2022
2022
Closest in time.
S. Chen, E. Xie, C. Ge, D. Liang, and P. Luo, “CycleMLP: A MLP-like architecture for dense prediction,” in ICLR , 2022
2022
Closest in time.
D. Lian, Z. Yu, X. Sun, and S. Gao, “AS-MLP: An axial shifted MLP architecture for vision,” in ICLR , 2022
2022
Closest in time.
D. Zhou, Z. Yu, E. Xie, C. Xiao, A. Anandkumar, J. Feng, and J. M. Alvarez, “Understanding the robustness in vision transformers,” in ICML , 2022
2022
Closest in time.
L. Hoyer, D. Dai, and L. Van Gool, “DAFormer: Improving network architectures and training strategies for domain-adaptive semantic segmentation,” in CVPR , 2022
2022
Closest in time.
Y. Liu, Y. Chen, P. Lasang, and Q. Sun, “Covariance attention for semantic segmentation,” TPAMI , vol. 44, no. 4, pp. 1805–1818, 2022
2022
Closest in time.
Z. Li, Y. Sun, L. Zhang, and J. Tang, “CTNet: Context-based tandem network for semantic segmentation,” TPAMI , vol. 44, no. 12, pp. 9904–9917, 2022
2022
Closest in time.
X. Dong, J. Bao, D. Chen, W. Zhang, N. Yu, L. Yuan, D. Chen, and B. Guo, “CSWin transformer: A general vision transformer backbone with cross-shaped windows,” in CVPR , 2022
2022
Closest in time.
Y.-H. Wu, Y. Liu, X. Zhan, and M.-M. Cheng, “P2T: Pyramid pooling transformer for scene understanding,” TPAMI , 2022
2022
Closest in time.
J. Gu, H. Kwon, D. Wang, W. Ye, M. Li, Y.-H. Chen, L. Lai, V. Chandra, and D. Z. Pan, “Multi-scale high-resolution vision transformer for semantic segmentation,” in CVPR , 2022
2022
Closest in time.
A. Petrovai and S. Nedevschi, “Semantic cameras for 360-degree environment perception in automated urban driving,” T-ITS , vol. 23, no. 10, pp. 17 271–17 283, 2022
2022
Closest in time.
S. Orhan and Y. Bastanlar, “Semantic segmentation of outdoor panoramic images,” SIVP , vol. 16, no. 3, pp. 643–650, 2022
2022
Closest in time.
X. Hu, Y. An, C. Shao, and H. Hu, “Distortion convolution module for semantic segmentation of panoramic images based on the image-forming principle,” TIM , vol. 71, pp. 1–12, 2022
2022
Closest in time.
J. Mei, A. Z. Zhu, X. Yan, H. Yan, S. Qiao, Y. Zhu, L.-C. Chen, H. Kretzschmar, and D. Anguelov, “Waymo open dataset: Panoramic video panoptic segmentation,” in ECCV , 2022
2022
Closest in time.
2022
Closest in time.
Z. Xia, X. Pan, S. Song, L. E. Li, and G. Huang, “Vision transformer with deformable attention,” in CVPR , 2022
2022
Closest in time.
H. Yin, A. Vahdat, J. M. Alvarez, A. Mallya, J. Kautz, and P. Molchanov, “A-ViT: Adaptive tokens for efficient vision transformer,” in CVPR , 2022
2022
Closest in time.
K. Liu, T. Wu, C. Liu, and G. Guo, “Dynamic group transformer: A general vision transformer backbone with dynamic group attention,” in IJCAI , 2022
2022
Closest in time.
R. Li, S. Li, C. He, Y. Zhang, X. Jia, and L. Zhang, “Class-balanced pixel-level self-labeling for domain adaptive semantic segmentation,” in CVPR , 2022
2022
Closest in time.
X. Huo, L. Xie, H. Hu, W. Zhou, H. Li, and Q. Tian, “Domain-agnostic prior for transfer semantic segmentation,” in CVPR , 2022
2022
Closest in time.
X. Lai, Z. Tian, X. Xu, Y. Chen, S. Liu, H. Zhao, L. Wang, and J. Jia, “DecoupleNet: Decoupled network for domain adaptive semantic segmentation,” in ECCV , 2022
2022
Closest in time.
Z. Jiang, Y. Li, C. Yang, P. Gao, Y. Wang, Y. Tai, and C. Wang, “Prototypical contrast adaptation for domain adaptive semantic segmentation,” in ECCV , 2022
2022
Closest in time.
A. R. Sekkat, Y. Dupuis, V. R. Kumar, H. Rashed, S. K. Yogamani, P. Vasseur, and P. Honeine, “SynWoodScape: Synthetic surround-view fisheye camera dataset for autonomous driving,” RA-L , vol. 7, no. 3, pp. 8502–8509, 2022
2022
Closest in time.
H. Zhang, C. Wu, Z. Zhang, Y. Zhu, Z. Zhang, H. Lin, Y. Sun, T. He, J. Mueller, R. Manmatha, M. Li, and A. J. Smola, “ResNeSt: Split-attention networks,” in CVPRW , 2022
2022
Closest in time.
B. Cheng, I. Misra, A. G. Schwing, A. Kirillov, and R. Girdhar, “Masked-attention mask transformer for universal image segmentation,” in CVPR , 2022
2022
Closest in time.
L. Hoyer, D. Dai, and L. Van Gool, “HRDA: Context-aware high-resolution domain-adaptive semantic segmentation,” in ECCV , 2022
2022
Closest in time.
W. Wang, E. Xie, X. Li, D. Fan, K. Song, D. Liang, T. Lu, P. Luo, and L. Shao, “PVT v2: Improved baselines with pyramid vision transformer,” CVM , vol. 8, no. 3, pp. 415–424, 2022
2022
Closest in time.
2023
Closest in time.
Y. Li, T. Yao, Y. Pan, and T. Mei, “Contextual transformer networks for visual recognition,” TPAMI , vol. 45, no. 2, pp. 1489–1500, 2023
2023
Closest in time.
Q. Hou, Z. Jiang, L. Yuan, M.-M. Cheng, S. Yan, and J. Feng, “Vision permutator: A permutable MLP-like architecture for visual recognition,” TPAMI , vol. 45, no. 1, pp. 1328–1334, 2023
2023
Closest in time.
A. Jaus, K. Yang, and R. Stiefelhagen, “Panoramic panoptic segmentation: Insights into surrounding parsing for mobile agents via unsupervised contrastive learning,” T-ITS , vol. 24, no. 4, pp. 4438–4453, 2023
2023
Closest in time.
B. Xie, S. Li, M. Li, C. H. Liu, G. Huang, and G. Wang, “SePiCo: Semantic-guided pixel contrast for domain adaptive semantic segmentation,” TPAMI , vol. 45, no. 7, pp. 9004–9021, 2023
2023
Closest in time.
Y. Liao, J. Xie, and A. Geiger, “KITTI-360: A novel dataset and benchmarks for urban scene understanding in 2D and 3D,” TPAMI , vol. 45, no. 3, pp. 3292–3310, 2023
2023
Closest in time.
P. Testolina, F. Barbato, U. Michieli, M. Giordani, P. Zanuttigh, and M. Zorzi, “SELMA: Semantic large-scale multimodal acquisitions in variable weather, daytime and viewpoints,” T-ITS , 2023
2023
Closest in time.
L. Hoyer, D. Dai, H. Wang, and L. Van Gool, “MIC: Masked image consistency for context-enhanced domain adaptation,” in CVPR , 2023
2023
Closest in time.