Fetching the paper…
Reading the bibliography…
Holistic understanding and reasoning in 3D scenes are crucial for the success of autonomous driving systems.
NAS-FPN: Learning Scalable Feature Pyramid Architecture for Object Detection
Ghiasi, G.; Lin, T.-Y.; Pang, R.; and Le, Q. V. 2019 · 1904
Earlier work this paper cites.
EfficientDet: Scalable and Efficient Object Detection
Tan, M.; Pang, R.; and Le, Q. V. 2020 · 1911
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2010
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Deformable Convolutional Networks
Dai, J.; Qi, H.; Xiong, Y.; Li, Y.; Zhang, G.; Hu, H.; and Wei, Y. 2017 · 2017
Earlier work this paper cites.
Feature Pyramid Networks for Object Detection
Lin, T.-Y.; Dollár, P.; Girshick, R.; He, K.; Hariharan, B.; and Belongie, S. 2017 · 2017
Earlier work this paper cites.
The lovász-softmax loss: A tractable surrogate for the optimization of the intersection-over-union measure in neural networks
Berman, M.; Triki, A. R.; and Blaschko, M. B. 2018 · 2018
Earlier work this paper cites.
PointRCNN: 3D object proposal generation and detection from point cloud
Shi, S.; Wang, X.; and Li, H. 2018 · 2018
Earlier work this paper cites.
PointPillars: Fast Encoders for Object Detection from Point Clouds
Lang, A. H.; Vora, S.; Caesar, H.; Zhou, L.; Yang, J.; and Beijbom, O. 2019 · 2019
Earlier work this paper cites.
Occupancy Networks: Learning 3D Reconstruction in Function Space
Mescheder, L.; Oechsle, M.; Niemeyer, M.; Nowozin, S.; and Geiger, A. 2019 · 2019
Earlier work this paper cites.
Disentangling monocular 3D object detection
Simonelli, A.; Bulò, S. R. R.; Porzi, L.; López-Antequera, M.; and Kontschieder, P. 2019 · 2019
Cited alongside, same era.
GridMask Data Augmentation
Chen, P.; Liu, S.; Zhao, H.; Wang, X.; and Jia, J. 2020 · 2020
Cited alongside, same era.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3D
Philion, J.; and Fidler, S. 2020 · 2020
Cited alongside, same era.
LMSCNet: Lightweight Multiscale 3D Semantic Completion
Roldão, L.; de Charette, R.; and Verroust-Blondet, A. 2020 · 2020
Cited alongside, same era.
Pseudo-LiDAR from Visual Depth Estimation: Bridging the Gap in 3D Object Detection for Autonomous Driving
Wang, Y.; Chao, W.-L.; Garg, D.; Hariharan, B.; Campbell, M.; and Weinberger, K. Q. 2020 · 2020
Cited alongside, same era.
Panoptic nuScenes: A Large-Scale Benchmark for LiDAR Panoptic Segmentation and Tracking
MonoScene: Monocular 3D Semantic Scene Completion
Cao, A.-Q.; and de Charette, R. 2022 · 2022
Later among the works it cites.
BEVDet4D: Exploit Temporal Cues in Multi-camera 3D Object Detection
Huang, J.; and Huang, G. 2022 · 2022
Later among the works it cites.
UniFusion: Unified multi-view fusion transformer for spatial-temporal representation in bird’s-eye-view
Qin, Z.; Chen, J.; Chen, C.; Chen, X.; and Li, X. 2022 · 2022
Later among the works it cites.
BEVerse: Unified Perception and Prediction in Birds-Eye-View for Vision-Centric Autonomous Driving
Zhang, Y.; Zhu, Z.; Zheng, W.; Huang, J.; Huang, G.; Zhou, J.; and Lu, J. 2022 · 2022
Later among the works it cites.
Tri-Perspective View for Vision-Based 3D Semantic Occupancy Prediction
Huang, Y.; Zheng, W.; Zhang, Y.; Zhou, J.; and Lu, J. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fong, W. K.; Mohan, R.; Hurtado, J. V.; Zhou, L.; Caesar, H.; Beijbom, O.; and Valada, A. 2021 · 2021
Cited alongside, same era.
BEVDet: High-performance multi-camera 3D object detection in Bird-Eye-View
Huang, J.; Huang, G.; Zhu, Z.; Ye, Y.; and Du, D. 2021 · 2021
Cited alongside, same era.
HDMapNet: An Online HD Map Construction and Evaluation Framework
Li, Q.; Wang, Y.; Wang, Y.; and Zhao, H. 2021 · 2021
Cited alongside, same era.
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
Liu, Z.; Lin, Y.; Cao, Y.; Hu, H.; Wei, Y.; Zhang, Z.; Lin, S.; and Guo, B. 2021 · 2021
Cited alongside, same era.
ImVoxelNet: Image to Voxels Projection for Monocular and Multi-View General-Purpose 3D Object Detection
Rukhovich, D.; Vorontsova, A.; and Konushin, A. 2021 · 2021
Cited alongside, same era.
Cylindrical and asymmetrical 3D convolution networks for LiDAR-based perception
Zhu, X.; Zhou, H.; Wang, T.; Hong, F.; Li, W.; Ma, Y.; Li, H.; Yang, R.; and Lin, D. 2021 · 2021
Cited alongside, same era.
BEVDepth: Acquisition of reliable depth for multi-view 3D object detection
Li, Y.; Ge, Z.; Yu, G.; Yang, J.; Wang, Z.; Shi, Y.; Sun, J.; and Li, Z. 2022a
Cited in the paper.
Occ-BEV: Multi-Camera Unified Pre-training via 3D Scene Reconstruction
Min, C.; Xu, X.; Li, F.; Si, S.; Xue, H.; Jiang, W.; Zhang, Z.; Li, J.; Zhao, D.; Xiao, L.; Xu, J.; Nie, Y.; and Dai, B. 2023 · 2023
Later among the works it cites.
Scene as Occupancy
Sima, C.; Tong, W.; Wang, T.; Chen, L.; Wu, S.; Deng, H.; Gu, Y.; Lu, L.; Luo, P.; Lin, D.; and Li, H. 2023 · 2023
Later among the works it cites.
Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving
Tian, X.; Jiang, T.; Yun, L.; Wang, Y.; Wang, Y.; and Zhao, H. 2023 · 2023
Later among the works it cites.
SurroundOcc: Multi-Camera 3D Occupancy Prediction for Autonomous Driving
Wei, Y.; Zhao, L.; Zheng, W.; Zhu, Z.; Zhou, J.; and Lu, J. 2023 · 2023
Later among the works it cites.
OccFormer: Dual-path transformer for vision-based 3D semantic occupancy prediction
Zhang, Y.; Zhu, Z.; and Du, D. 2023 · 2023
Later among the works it cites.