Fetching the paper…
Reading the bibliography…
Building 3D perception systems for autonomous vehicles that do not rely on high-density LiDAR is a critical research problem because of the expense of LiDAR systems compared to cameras and other sensors.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in ECCV , D. Fleet, T. Pajdla, B. Schiele, and T. Tuytelaars, Eds., 2014
2014
Earlier work this paper cites.
A. Bewley, Z. Ge, L. Ott, F. Ramos, and B. Upcroft, “Simple online and realtime tracking,” in ICIP . IEEE, 2016, pp. 3464–3468
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
M. Parker, “Chapter 20 – automotive radar,” in Digital Signal Processing 101 , 2nd ed., M. Parker, Ed. Newnes, 2017, pp. 253–276
2017
Earlier work this paper cites.
J. Lombacher, K. Laudt, M. Hahn, J. Dickmann, and C. Wöhler, “Semantic radar grids,” in Intelligent vehicles symposium (IV) . IEEE, 2017, pp. 1170–1175
2017
Earlier work this paper cites.
V. L. Dmitry Ulyanov, Andrea Vedaldi, “Improved texture networks: Maximizing quality and diversity in feed-forward stylization and texture synthesis,” in CVPR , 2017
2017
Earlier work this paper cites.
I. Loshchilov and F. Hutter, “Decoupled weight decay regularization,” arXiv:1711.05101 , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Schulter, M. Zhai, N. Jacobs, and M. Chandraker, “Learning to look around objects for top-view representations of outdoor scenes,” in ECCV , 2018, pp. 787–802
2018
Earlier work this paper cites.
R. Cheng, Z. Wang, and K. Fragkiadaki, “Geometry-aware recurrent neural networks for active visual recognition,” in NeurIPS , 2018
2018
Earlier work this paper cites.
O. Schumann, M. Hahn, J. Dickmann, and C. Wöhler, “Semantic segmentation on radar point clouds,” in International Conference on Information Fusion (FUSION) . IEEE, 2018, pp. 2179–2186
2018
Earlier work this paper cites.
A. Kendall, Y. Gal, and R. Cipolla, “Multi-task learning using uncertainty to weigh losses for scene geometry and semantics,” in CVPR , 2018, pp. 7482–7491
2018
Earlier work this paper cites.
V. Sitzmann, J. Thies, F. Heide, M. Nießner, G. Wetzstein, and M. Zollhofer, “DeepVoxels: Learning persistent 3D feature embeddings,” in CVPR , 2019, pp. 2437–2446
2019
Earlier work this paper cites.
H.-Y. F. Tung, R. Cheng, and K. Fragkiadaki, “Learning spatial common sense with geometry-aware recurrent networks,” CVPR , 2019
2019
Earlier work this paper cites.
T.-Y. Lim, A. Ansari, B. Major, D. Fontijne, M. Hamilton, R. Gowaikar, and S. Subramanian, “Radar and camera early fusion for vehicle detection in advanced driver assistance systems,” in Machine Learning for Autonomous Driving Workshop at NeurIPS , vol. 2, 2019, p. 7
2019
Earlier work this paper cites.
L. Sless, B. El Shlomo, G. Cohen, and S. Oron, “Road scene understanding by occupancy grid learning from sparse radar clusters using semantic segmentation,” in CVPR Workshops , 2019, pp. 0–0
2019
Cited alongside, same era.
M. Meyer and G. Kuschk, “Deep learning based 3D object detection for automotive radar and camera,” in European Radar Conference (EuRAD) . IEEE, 2019, pp. 133–136
2019
Cited alongside, same era.
2019
Cited alongside, same era.
L. N. Smith and N. Topin, “Super-convergence: Very fast training of neural networks using large learning rates,” in Artificial Intelligence and Machine Learning for Multi-Domain Operations Applications , vol. 11006. International Society for Optics and Photonics, 2019, p. 1100612
2019
Cited alongside, same era.
A. Saha, O. Mendez, C. Russell, and R. Bowden, “Enabling spatio-temporal aggregation in birds-eye-view vehicle estimation,” in ICRA . IEEE, 2021, pp. 5133–5139
2021
Later among the works it cites.
C. Reading, A. Harakeh, J. Chae, and S. L. Waslander, “Categorical depth distribution network for monocular 3D object detection,” in CVPR , 2021
2021
Later among the works it cites.
A. Hu, Z. Murez, N. Mohan, S. Dudas, J. Hawke, V. Badrinarayanan, R. Cipolla, and A. Kendall, “FIERY: Future instance prediction in bird’s-eye view from surround monocular cameras,” in CVPR , 2021, pp. 15 273–15 282
2021
Later among the works it cites.
H. Wang, P. Cai, Y. Sun, L. Wang, and M. Liu, “Learning interpretable end-to-end vision-based motion planning for autonomous driving with optical flow distillation,” in 2021 ICRA (ICRA) . IEEE, 2021, pp. 13 731–13 737
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
B. Pan, J. Sun, H. Y. T. Leung, A. Andonian, and B. Zhou, “Cross-view semantic segmentation for sensing surroundings,” IEEE Robotics and Automation Letters , vol. 5, no. 3, pp. 4867–4873, 2020
2020
Cited alongside, same era.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3D,” in ECCV . Springer, 2020, pp. 194–210
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. W. Harley, S. K. Lakshmikanth, F. Li, X. Zhou, H.-Y. F. Tung, and K. Fragkiadaki, “Learning from unlabelled videos using contrastive predictive neural 3D mapping,” in ICLR , 2020
2020
Cited alongside, same era.
B. Liu, B. Zhuang, S. Schulter, P. Ji, and M. Chandraker, “Understanding road layout from videos as a whole,” in CVPR , 2020, pp. 4414–4423
2020
Cited alongside, same era.
T. Roddick and R. Cipolla, “Predicting semantic map representations from images using pyramid occupancy networks,” in CVPR , 2020, pp. 11 138–11 147
2020
Cited alongside, same era.
B. Yang, R. Guo, M. Liang, S. Casas, and R. Urtasun, “Radarnet: Exploiting radar for robust perception of dynamic objects,” in ECCV . Springer, 2020, pp. 496–512
2020
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
W. Yang, Q. Li, W. Liu, Y. Yu, Y. Ma, S. He, and J. Pan, “Projecting your view attentively: Monocular road scene layout estimation via cross-view transformation,” in CVPR , 2021, pp. 15 536–15 545
2021
Later among the works it cites.
D. Park, R. Ambrus, V. Guizilini, J. Li, and A. Gaidon, “Is pseudo-lidar needed for monocular 3D object detection?” in CVPR , 2021, pp. 3142–3152
2021
Later among the works it cites.
2022
Closest in time.
Z. Li, W. Wang, H. Li, E. Xie, C. Sima, T. Lu, Y. Qiao, and J. Dai, “BEVFormer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers,” ECCV , 2022
2022
Closest in time.
Y. B. Can, A. Liniger, O. Unal, D. Paudel, and L. Van Gool, “Understanding bird’s-eye view of road semantics using an onboard camera,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 3302–3309, 2022
2022
Closest in time.
N. Gosala and A. Valada, “Bird’s-eye-view panoptic segmentation using monocular frontal view images,” IEEE Robotics and Automation Letters , 2022
2022
Closest in time.
A. Saha, O. Mendez, C. Russell, and R. Bowden, “Translating images into maps,” in ICRA , 2022
2022
Closest in time.
B. Zhou and P. Krähenbühl, “Cross-view transformers for real-time map-view semantic segmentation,” in CVPR , 2022, pp. 13 760–13 769
2022
Closest in time.