Fetching the paper…
Reading the bibliography…
LiDAR-camera fusion can enhance the performance of 3D object detection by utilizing complementary information between depth-aware LiDAR points and semantically rich images.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in Computer Vision and Pattern Recognition (CVPR), 2012 IEEE Conference on . IEEE, 2012, pp. 3354–3361. [Online]. Available: https://ieeexplore.ieee.org/abstract/document/6248074
2012
Earlier work this paper cites.
A. Mousavian, D. Anguelov, J. Flynn, and J. Kosecka, “3D Bounding Box Estimation Using Deep Learning and Geometry,” in 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , Jul 2017. [Online]. Available: http://dx.doi.org/10.1109/cvpr.2017.597
2017
Earlier work this paper cites.
R. Q. Charles, S. Hao, M. Kaichun, and J. G. Leonidas, “PointNet: Deep Learning on Point Sets for 3D Classification and Segmentation,” pp. 77–85, 2017. [Online]. Available: https://doi.org/10.1109/CVPR.2017.16
2017
Earlier work this paper cites.
C. R. Qi, L. Yi, H. Su, and L. J. Guibas, “PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space.” in NIPS , I. Guyon, U. von Luxburg, S. Bengio, H. M. Wallach, R. Fergus, S. V. N. Vishwanathan, and R. Garnett, Eds., 2017, pp. 5099–5108. [Online]. Available: http://dblp.uni-trier.de/db/conf/nips/nips2017.html#QiYSG17
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
X. Chen, H. Ma, J. Wan, B. Li, and T. Xia, “Multi-view 3D Object Detection Network for Autonomous Driving.” in CVPR . IEEE Computer Society, 2017, pp. 6526–6534. [Online]. Available: http://dblp.uni-trier.de/db/conf/cvpr/cvpr2017.html#ChenMWLX17
2017
Earlier work this paper cites.
C. R. Qi, W. Liu, C. Wu, H. Su, and L. J. Guibas, “Frustum pointnets for 3d object detection from rgb-d data,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 918–927
2018
Earlier work this paper cites.
Y. Zhou and O. Tuzel, “VoxelNet: End-to-End Learning for Point Cloud Based 3D Object Detection.” in CVPR . IEEE Computer Society, 2018, pp. 4490–4499. [Online]. Available: http://dblp.uni-trier.de/db/conf/cvpr/cvpr2018.html#ZhouT18
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Y. Yan, Y. Mao, and B. Li, “SECOND: Sparsely Embedded Convolutional Detection.” Sensors , vol. 18, no. 10, p. 3337, 2018. [Online]. Available: http://dblp.uni-trier.de/db/journals/sensors/sensors18.html#YanML18
2018
Earlier work this paper cites.
B. Graham, M. Engelcke, and L. van der Maaten, “3D Semantic Segmentation with Submanifold Sparse Convolutional Networks,” CVPR , 2018
2018
Earlier work this paper cites.
Y. Wang, W.-L. Chao, D. Garg, B. Hariharan, M. Campbell, and K. Q. Weinberger, “Pseudo-lidar from visual depth estimation: Bridging the gap in 3d object detection for autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 8445–8453
2019
Earlier work this paper cites.
S. Shi, X. Wang, and H. Li, “PointRCNN: 3D Object Proposal Generation and Detection From Point Cloud.” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 770–779. [Online]. Available: http://dblp.uni-trier.de/db/conf/cvpr/cvpr2019.html#ShiWL19
2019
Earlier work this paper cites.
Z. Wang and K. Jia, “Frustum convnet: Sliding frustums to aggregate local point-wise features for amodal 3d object detection,” in 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2019, pp. 1742–1749
2019
Earlier work this paper cites.
V. A. Sindagi, Y. Zhou, and O. Tuzel, “Mvx-net: Multimodal voxelnet for 3d object detection,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 7276–7282
2019
Earlier work this paper cites.
M. Liang, B. Yang, Y. Chen, R. Hu, and R. Urtasun, “Multi-task multi-sensor fusion for 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 7345–7353
2019
Earlier work this paper cites.
T. Huang, Z. Liu, X. Chen, and X. Bai, “Epnet: Enhancing point features with image semantics for 3d object detection,” in European Conference on Computer Vision . Springer, 2020, pp. 35–52
2020
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 11 621–11 631
2020
Earlier work this paper cites.
Z. Liu, Z. Wu, and R. Tóth, “Smoke: Single-stage monocular 3d object detection via keypoint estimation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops , 2020, pp. 996–997
2020
Earlier work this paper cites.
P. Li, H. Zhao, P. Liu, and F. Cao, “Rtm3d: Real-time monocular 3d detection from object keypoints for autonomous driving,” in European Conference on Computer Vision . Springer, 2020, pp. 644–660
2020
Earlier work this paper cites.
Y. Chen, L. Tai, K. Sun, and M. Li, “MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships,” in 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , Jun 2020. [Online]. Available: http://dx.doi.org/10.1109/cvpr42600.2020.01211
2020
Earlier work this paper cites.
Z. Yang, Y. Sun, S. Liu, and J. Jia, “3dssd: Point-based 3d single stage object detector,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 11 040–11 048
2020
Earlier work this paper cites.
S. Vora, A. H. Lang, B. Helou, and O. Beijbom, “Pointpainting: Sequential fusion for 3d object detection,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 4604–4612
2020
Earlier work this paper cites.
L. Xie, C. Xiang, Z. Yu, G. Xu, Z. Yang, D. Cai, and X. He, “PI-RCNN: An efficient multi-sensor 3D object detector with point-based attentive cont-conv fusion module,” in Proceedings of the AAAI conference on artificial intelligence , vol. 34, no. 07, 2020, pp. 12 460–12 467
2020
Earlier work this paper cites.
J. H. Yoo, Y. Kim, J. Kim, and J. W. Choi, “3d-cvf: Generating joint camera and lidar features using cross-view spatial feature fusion for 3d object detection,” in European Conference on Computer Vision . Springer, 2020, pp. 720–736
2020
Earlier work this paper cites.
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang, and H. Li, “Pv-rcnn: Point-voxel feature set abstraction for 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 10 529–10 538
2020
Earlier work this paper cites.
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou, and X. Bai, “Tanet: Robust 3d object detection from point clouds with triple attention,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 07, 2020, pp. 11 677–11 684
2020
Cited alongside, same era.
J. Wang, S. Lan, M. Gao, and L. S. Davis, “Infofocus: 3d object detection for autonomous driving with dynamic information modeling,” in European Conference on Computer Vision . Springer, 2020, pp. 405–420
2020
Cited alongside, same era.
T. Wang, X. Zhu, J. Pang, and D. Lin, “Fcos3d: Fully convolutional one-stage monocular 3d object detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 913–922
2021
Cited alongside, same era.
J. Deng, S. Shi, P. Li, W. Zhou, Y. Zhang, and H. Li, “Voxel r-cnn: Towards high performance voxel-based 3d object detection,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 35, no. 2, 2021, pp. 1201–1209
2021
X. Bai, Z. Hu, X. Zhu, Q. Huang, Y. Chen, H. Fu, and C.-L. Tai, “Transfusion: Robust lidar-camera fusion for 3d object detection with transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 1090–1099
2022
Later among the works it cites.
S. Pang, D. Morris, and H. Radha, “Fast-CLOCs: Fast camera-LiDAR object candidates fusion for 3D object detection,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2022, pp. 187–196
2022
Later among the works it cites.
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He, and D. Cai, “Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph,” in European Conference on Computer Vision . Springer, 2022, pp. 662–679
2022
Later among the works it cites.
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang, and X. Bai, “EPNet++: Cascade bi-directional fusion for multi-modal 3D object detection,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2022
2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
J. Deng, W. Zhou, Y. Zhang, and H. Li, “From multi-view to hollow-3D: Hallucinated hollow-3D R-CNN for 3D object detection,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 31, no. 12, pp. 4722–4734, 2021
2021
Cited alongside, same era.
H. Sheng, S. Cai, Y. Liu, B. Deng, J. Huang, X.-S. Hua, and M.-J. Zhao, “Improving 3d object detection with channel-wise transformer,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 2743–2752
2021
Cited alongside, same era.
T. Yin, X. Zhou, and P. Krähenbühl, “Multimodal virtual point 3d detection,” Advances in Neural Information Processing Systems , vol. 34, pp. 16 494–16 507, 2021
2021
Cited alongside, same era.
C. Wang, C. Ma, M. Zhu, and X. Yang, “Pointaugmenting: Cross-modal augmentation for 3d object detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 11 794–11 803
2021
Cited alongside, same era.
T. Yin, X. Zhou, and P. Krahenbuhl, “Center-based 3d object detection and tracking,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 11 784–11 793
2021
Cited alongside, same era.
P. Wang, L. Shi, B. Chen, Z. Hu, J. Qiao, and Q. Dong, “Pursuing 3-D scene structures with optical satellite images from affine reconstruction to Euclidean reconstruction,” IEEE Transactions on Geoscience and Remote Sensing , vol. 60, pp. 1–14, 2022
2022
Cited alongside, same era.
X. Wu, L. Peng, H. Yang, L. Xie, C. Huang, C. Deng, H. Liu, and D. Cai, “Sparse fuse dense: Towards high quality 3d detection with depth completion,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 5418–5427
2022
Cited alongside, same era.
Y. Li, X. Qi, Y. Chen, L. Wang, Z. Li, J. Sun, and J. Jia, “Voxel Field Fusion for 3D Object Detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 1120–1129
2022
Cited alongside, same era.
Later among the works it cites.
W. Zheng, M. Hong, L. Jiang, and C.-W. Fu, “Boosting 3d object detection by simulating multimodality on point clouds,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 13 638–13 647
2022
Later among the works it cites.
Y. Hu, Z. Ding, R. Ge, W. Shao, L. Huang, K. Li, and Q. Liu, “Afdetv2: Rethinking the necessity of the second stage for object detection from point clouds,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 36, no. 1, 2022, pp. 969–979
2022
Later among the works it cites.
S. Deng, Z. Liang, L. Sun, and K. Jia, “Vista: Boosting 3d object detection via dual cross-view spatial attention,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 8448–8457
2022
Later among the works it cites.
2022
Later among the works it cites.
O. Team et al. , “Openpcdet: An open-source toolbox for 3d object detection from point clouds,” 2020
2022
Later among the works it cites.
Z. Song, H. Wei, L. Bai, L. Yang, and C. Jia, “GraphAlign: Enhancing Accurate Feature Alignment by Graph matching for Multi-Modal 3D Object Detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2023, pp. 3358–3369
2023
Later among the works it cites.
Z. Song, C. Jia, L. Yang, H. Wei, and L. Liu, “GraphAlign++: An Accurate Feature Alignment by Graph Matching for Multi-Modal 3D Object Detection,” IEEE Transactions on Circuits and Systems for Video Technology , pp. 1–1, 2023
2023
Later among the works it cites.
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, J. Zhou, and J. Lu, “Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 21 729–21 740
2023
Later among the works it cites.
L. Yang, X. Zhang, J. Li, L. Wang, M. Zhu, and L. Zhu, “Lite-fpn for keypoint-based monocular 3d object detection,” Knowledge-Based Systems , vol. 271, p. 110517, 2023
2023
Later among the works it cites.
L. Yang, X. Zhang, J. Li, L. Wang, M. Zhu, C. Zhang, and H. Liu, “Mix-teaching: A simple, unified and effective semi-supervised learning framework for monocular 3d object detection,” IEEE Transactions on Circuits and Systems for Video Technology , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
L. Piccinelli, C. Sakaridis, and F. Yu, “iDisc: Internal Discretization for Monocular Depth Estimation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 21 477–21 487
2023
Later among the works it cites.
Y. Li, Z. Ge, G. Yu, J. Yang, Z. Wang, Y. Shi, J. Sun, and Z. Li, “Bevdepth: Acquisition of reliable depth for multi-view 3d object detection,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 37, no. 2, 2023, pp. 1477–1485
2023
Later among the works it cites.
T. Xie, L. Wang, K. Wang, R. Li, X. Zhang, H. Zhang, L. Yang, H. Liu, and J. Li, “FARP-Net: Local-Global Feature Aggregation and Relation-Aware Proposals for 3D Object Detection,” IEEE Transactions on Multimedia , pp. 1–15, 2023
2023
Later among the works it cites.
Z. Song, H. Wei, C. Jia, Y. Xia, X. Li, and C. Zhang, “VP-Net: Voxels as Points for 3D Object Detection,” IEEE Transactions on Geoscience and Remote Sensing , vol. , pp. 1–1, 2023
2023
Later among the works it cites.
Q. Xia, Y. Chen, G. Cai, G. Chen, D. Xie, J. Su, and Z. Wang, “3-D HANet: A Flexible 3-D Heatmap Auxiliary Network for Object Detection,” IEEE Transactions on Geoscience and Remote Sensing , vol. 61, pp. 1–13, 2023
2023
Later among the works it cites.
L. Wang, X. Zhang, Z. Song, J. Bi, G. Zhang, H. Wei, L. Tang, L. Yang, J. Li, C. Jia et al. , “Multi-modal 3D Object Detection in Autonomous Driving: A Survey and Taxonomy,” IEEE Transactions on Intelligent Vehicles , 2023
2023
Later among the works it cites.
L. Wang, Z. Song, X. Zhang, C. Wang, G. Zhang, L. Zhu, J. Li, and H. Liu, “SAT-GCN: Self-attention graph convolutional network-based 3D object detection for autonomous driving,” Knowledge-Based Systems , vol. 259, p. 110080, 2023
2023
Later among the works it cites.
Y. Jiao, Z. Jie, S. Chen, J. Chen, L. Ma, and Y.-G. Jiang, “MSMDFusion: Fusing LiDAR and Camera at Multiple Scales With Multi-Depth Seeds for 3D Object Detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2023, pp. 21 643–21 652
2023
Later among the works it cites.
H. Wu, C. Wen, S. Shi, X. Li, and C. Wang, “Virtual Sparse Convolution for Multimodal 3D Object Detection,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2023, pp. 21 653–21 662
2023
Later among the works it cites.
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu, and X. Li, “Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection,” IEEE Transactions on Circuits and Systems for Video Technology , 2023
2023
Later among the works it cites.
X. Tian, M. Yang, Q. Yu, J. Yong, and D. Xu, “MedoidsFormer: A Strong 3D Object Detection Backbone by Exploiting Interaction with Adjacent Medoid Tokens,” IEEE Transactions on Circuits and Systems for Video Technology , 2023
2023
Later among the works it cites.
Y. Chen, J. Liu, X. Zhang, X. Qi, and J. Jia, “Voxelnext: Fully sparse voxelnet for 3d object detection and tracking,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 21 674–21 683
2023
Later among the works it cites.