Fetching the paper…
Reading the bibliography…
Recent camera-based 3D object detection is limited by the precision of transforming from image to 3D feature spaces, as well as the accuracy of object localization within the 3D space.
Eigen, David, Christian Puhrsch, and Rob Fergus. ”Depth map prediction from a single image using a multi-scale deep network.” Advances in neural information processing systems 27 (2014)
2014
Earlier work this paper cites.
He, Kaiming, et al. ”Deep residual learning for image recognition.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2016
2016
Earlier work this paper cites.
Tarvainen, Antti, and Harri Valpola. ”Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results.” Advances in neural information processing systems 30 (2017)
2017
Earlier work this paper cites.
Fu, Huan, et al. ”Deep ordinal regression network for monocular depth estimation.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2018
2018
Earlier work this paper cites.
Li, Buyu, et al. ”Gs3d: An efficient 3d object detection framework for autonomous driving.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2019
2019
Earlier work this paper cites.
Li, Peiliang, Xiaozhi Chen, and Shaojie Shen. ”Stereo r-cnn based 3d object detection for autonomous driving.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Zhou, Ling, et al. ”Pattern-structure diffusion for multi-task learning.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2020
2020
Earlier work this paper cites.
Caesar, Holger, et al. ”nuscenes: A multimodal dataset for autonomous driving.” Proceedings of the IEEE/CVF conference on computer vision and pattern recognition. 2020
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
Reading, Cody, et al. ”Categorical depth distribution network for monocular 3d object detection.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2021
2021
Earlier work this paper cites.
Shi, Xuepeng, et al. ”Geometry-based distance decomposition for monocular 3d object detection.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2021
2021
Cited alongside, same era.
Bhat, Shariq Farooq, Ibraheem Alhashim, and Peter Wonka. ”Adabins: Depth estimation using adaptive bins.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2021
2021
Cited alongside, same era.
Ranftl, René, Alexey Bochkovskiy, and Vladlen Koltun. ”Vision transformers for dense prediction.” Proceedings of the IEEE/CVF international conference on computer vision. 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Wei, Zizhuang, et al. ”Bidirectional hybrid LSTM based recurrent neural network for multi-view stereo.” IEEE Transactions on Visualization and Computer Graphics (2022)
2022
Later among the works it cites.
Li, Feng, et al. ”Dn-detr: Accelerate detr training by introducing query denoising.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Wei, Zizhuang, et al. ”Aa-rmvsnet: Adaptive aggregation recurrent multi-view stereo network.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2021
2021
Cited alongside, same era.
Liu, Ze, et al. ”Swin transformer: Hierarchical vision transformer using shifted windows.” Proceedings of the IEEE/CVF international conference on computer vision. 2021
2021
Cited alongside, same era.
Wang, Tai, et al. ”Fcos3d: Fully convolutional one-stage monocular 3d object detection.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2021
2021
Cited alongside, same era.
Wang, Yue, et al. ”Detr3d: 3d object detection from multi-view images via 3d-to-2d queries.” Conference on Robot Learning. PMLR, 2022
2022
Cited alongside, same era.
Liu, Yingfei, et al. ”Petr: Position embedding transformation for multi-view 3d object detection.” European Conference on Computer Vision. Cham: Springer Nature Switzerland, 2022
2022
Cited alongside, same era.
Li, Zhiqi, et al. ”Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers.” European conference on computer vision. Cham: Springer Nature Switzerland, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Later among the works it cites.
Li, Yinhao, et al. ”Bevdepth: Acquisition of reliable depth for multi-view 3d object detection.” Proceedings of the AAAI Conference on Artificial Intelligence. Vol. 37. No. 2. 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Yang, Chenyu, et al. ”BEVFormer v2: Adapting Modern Image Backbones to Bird’s-Eye-View Recognition via Perspective Supervision.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Jiang, Yanqin, et al. ”Polarformer: Multi-camera 3d object detection with polar transformer.” Proceedings of the AAAI Conference on Artificial Intelligence. Vol. 37. No. 1. 2023
2023
Later among the works it cites.
Yang, Jie, et al. ”Neural interactive keypoint detection.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2023
2023
Later among the works it cites.