Fetching the paper…
Reading the bibliography…
A semantic map of the road scene, covering fundamental road elements, is an essential ingredient in autonomous driving systems.
H. A. Mallot, H. H. Bülthoff, J. Little, and S. Bohrer, “Inverse perspective mapping simplifies optical flow computation and obstacle detection,” Biol. Cybern. , vol. 64, no. 3, pp. 177–185, 1991
1991
Earlier work this paper cites.
S. Sengupta, P. Sturgess, L. Ladickỳ, and P. H. Torr, “Automatic dense visual semantic mapping from street-level imagery,” in Proc. IROS , 2012, pp. 857–862
2012
Earlier work this paper cites.
A. Bar-Hillel, R. Lerner, D. Levi, and G. Raz, “Recent progress in road and lane detection: a survey,” Mach. Vis. Appl. , vol. 25, no. 3, pp. 727–745, 2014
2014
Earlier work this paper cites.
Y. Zhang, T. Xiang, T. M. Hospedales, and H. Lu, “Deep mutual learning,” in Proc. CVPR , 2018, pp. 4320–4328
2018
Earlier work this paper cites.
S. Ammar Abbas and A. Zisserman, “A geometric approach to obtain a bird’s eye view from an image,” in Proc. CVPR , 2019, pp. 4095–4104
2019
Earlier work this paper cites.
C. Lu, M. J. G. van de Molengraft, and G. Dubbelman, “Monocular semantic occupancy grid mapping with convolutional variational encoder–decoder networks,” IEEE Robot. Autom. Lett. , vol. 4, no. 2, pp. 445–452, 2019
2019
Earlier work this paper cites.
M. Tan and Q. Le, “EfficientNet: Rethinking model scaling for convolutional neural networks,” in Proc. ICML , 2019, pp. 6105–6114
2019
Earlier work this paper cites.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3D,” in Proc. ECCV , 2020, pp. 194–210
2020
Earlier work this paper cites.
H. Caesar et al. , “nuScenes: A multimodal dataset for autonomous driving,” in Proc. CVPR , 2020, pp. 11 618–11 628
2020
Earlier work this paper cites.
L. Reiher et al. , “Cam2BEV,” 2020. [Online]. Available: https://github.com/ika-rwth-aachen/Cam2BEV
2020
Earlier work this paper cites.
L. Reiher, B. Lampe, and L. Eckstein, “A Sim2Real deep learning approach for the transformation of images from multiple vehicle-mounted cameras to a semantically segmented image in bird’s eye view,” in Proc. ITSC , 2020, pp. 1–7
2020
Earlier work this paper cites.
T. Roddick and R. Cipolla, “Predicting semantic map representations from images using pyramid occupancy networks,” in Proc. CVPR , 2020, pp. 11 135–11 144
2020
Earlier work this paper cites.
B. Pan, J. Sun, H. Y. T. Leung, A. Andonian, and B. Zhou, “Cross-view semantic segmentation for sensing surroundings,” IEEE Robot. Autom. Lett. , vol. 5, no. 3, pp. 4867–4873, 2020
2020
Earlier work this paper cites.
W. Yang et al. , “Projecting your view attentively: Monocular road scene layout estimation via cross-view transformation,” in Proc. CVPR , 2021, pp. 15 531–15 540
2021
Cited alongside, same era.
A. Hu et al. , “FIERY: Future instance prediction in bird’s-eye view from surround monocular cameras,” in Proc. ICCV , 2021, pp. 15 253–15 262
2021
Cited alongside, same era.
J. Kim, M. Hyun, I. Chung, and N. Kwak, “Feature fusion for online mutual knowledge distillation,” in Proc. ICPR , 2021, pp. 4619–4625
2021
Cited alongside, same era.
Z. Li et al. , “BEVFormer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers,” in Proc. ECCV , 2022, pp. 1–18
2022
Cited alongside, same era.
N. Gosala and A. Valada, “Bird’s-eye-view panoptic segmentation using monocular frontal view images,” IEEE Robot. Autom. Lett. , vol. 7, no. 2, pp. 1968–1975, 2022
2022
Later among the works it cites.
S. Gao, Q. Wang, and Y. Sun, “S2G2: Semi-supervised semantic bird-eye-view grid-map generation using a monocular camera for autonomous driving,” IEEE Robot. Autom. Lett. , vol. 7, no. 4, pp. 11 974–11 981, 2022
2022
Later among the works it cites.
B. Zhou and P. Krähenbühl, “Cross-view transformers for real-time map-view semantic segmentation,” in Proc. CVPR , 2022, pp. 13 750–13 759
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
K. Peng et al. , “MASS: Multi-attentional semantic segmentation of LiDAR data for dense top-view understanding,” IEEE Trans. Intell. Transp. Syst. , vol. 23, no. 9, pp. 15 824–15 840, 2022
2022
Cited alongside, same era.
Y. B. Can, A. Liniger, O. Unal, D. Paudel, and L. Van Gool, “Understanding bird’s-eye view of road semantics using an onboard camera,” IEEE Robot. Autom. Lett. , vol. 7, no. 2, pp. 3302–3309, 2022
2022
Cited alongside, same era.
Q. Li, Y. Wang, Y. Wang, and H. Zhao, “HDMapNet: An online HD map construction and evaluation framework,” in Proc. ICRA , 2022, pp. 4628–4634
2022
Cited alongside, same era.
Y. Ma et al. , “Vision-centric BEV perception: A survey,” arXiv preprint arXiv:2208.02797 , 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Y. B. Can, A. Liniger, O. Unal, D. Paudel, and L. Van Gool, “Understanding bird’s-eye view of road semantics using an onboard camera,” IEEE Robot. Autom. Lett. , vol. 7, no. 2, pp. 3302–3309, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Later among the works it cites.
S. Gong et al. , “GitNet: Geometric prior-based transformation for birds-eye-view segmentation,” in Proc. ECCV , 2022, pp. 396–411
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Hong, H. Dai, and Y. Ding, “Cross-modality knowledge distillation network for monocular 3D object detection,” in Proc. ECCV , 2022, pp. 87–104
2022
Later among the works it cites.
Y. Liu, T. Yuan, Y. Wang, Y. Wang, and H. Zhao, “Vectormapnet: End-to-end vectorized hd map learning,” Proc. ICML , 2023
2023
Closest in time.
2023
Closest in time.
Y. Li et al. , “BEVDepth: Acquisition of reliable depth for multi-view 3D object detection,” in Proc. AAAI , 2023
2023
Closest in time.
B. Liao, S. Chen, X. Wang, T. Cheng, Q. Zhang, W. Liu, and C. Huang, “Maptr: Structured modeling and learning for online vectorized hd map construction,” Proc. ICLR , 2023
2023
Closest in time.
L. Peng, Z. Chen, Z. Fu, P. Liang, and E. Cheng, “BEVSegFormer: Bird’s eye view semantic segmentation from arbitrary camera rigs,” in Proc. WACV , 2023, pp. 5935–5943
2023
Closest in time.