Fetching the paper…
Reading the bibliography…
3D occupancy prediction holds significant promise in the fields of robot perception and autonomous driving, which quantifies 3D scenes into grid cells with semantic labels.
L. Rudin and S. Osher, “Total variation based image restoration with free local constraints,” in Proceedings of 1st International Conference on Image Processing , Dec 2002. [Online]. Available: http://dx.doi.org/10.1109/icip.1994.413269
1994
Earlier work this paper cites.
N. Max, “Optical models for direct volume rendering,” IEEE Transactions on Visualization and Computer Graphics , p. 99–108, Jun 1995. [Online]. Available: http://dx.doi.org/10.1109/2945.468400
1995
Earlier work this paper cites.
D. Eigen, C. Puhrsch, and R. Fergus, “Depth map prediction from a single image using a multi-scale deep network,” Advances in neural information processing systems , vol. 27, 2014
2014
Earlier work this paper cites.
A. Garcia-Garcia, F. Gomez-Donoso, J. Garcia-Rodriguez, S. Orts-Escolano, M. Cazorla, and J. Azorin-Lopez, “Pointnet: A 3d convolutional neural network for real-time object class recognition,” in 2016 International Joint Conference on Neural Networks (IJCNN) , Jul 2016. [Online]. Available: http://dx.doi.org/10.1109/ijcnn.2016.7727386
2016
Earlier work this paper cites.
J. L. Schonberger and J.-M. Frahm, “Structure-from-motion revisited,” in 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , Jun 2016. [Online]. Available: http://dx.doi.org/10.1109/cvpr.2016.445
2016
Earlier work this paper cites.
C. R. Qi, L. Yi, H. Su, and L. J. Guibas, “Pointnet++: Deep hierarchical feature learning on point sets in a metric space,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
Y. Yan, Y. Mao, and B. Li, “Second: Sparsely embedded convolutional detection,” Sensors , p. 3337, Oct 2018. [Online]. Available: http://dx.doi.org/10.3390/s18103337
2018
Earlier work this paper cites.
E. Arnold, O. Y. Al-Jarrah, M. Dianati, S. Fallah, D. Oxtoby, and A. Mouzakitis, “A survey on 3d object detection methods for autonomous driving applications,” IEEE Transactions on Intelligent Transportation Systems , vol. 20, no. 10, pp. 3782–3795, 2019
2019
Earlier work this paper cites.
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall, “Semantickitti: A dataset for semantic scene understanding of lidar sequences,” in 2019 IEEE/CVF International Conference on Computer Vision (ICCV) , Oct 2019. [Online]. Available: http://dx.doi.org/10.1109/iccv.2019.00939
2019
Earlier work this paper cites.
A. H. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang, and O. Beijbom, “Pointpillars: Fast encoders for object detection from point clouds,” in 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , Jun 2019. [Online]. Available: http://dx.doi.org/10.1109/cvpr.2019.01298
2019
Earlier work this paper cites.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d,” in European Conference on Computer Vision . Springer, 2020, pp. 194–210
2020
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 11 621–11 631
2020
Earlier work this paper cites.
L. Roldao, R. de Charette, and A. Verroust-Blondet, “Lmscnet: Lightweight multiscale 3d semantic completion.” in 2020 International Conference on 3D Vision (3DV) , Nov 2020. [Online]. Available: http://dx.doi.org/10.1109/3dv50981.2020.00021
2020
Earlier work this paper cites.
J. Li, K. Han, P. Wang, Y. Liu, and X. Yuan, “Anisotropic convolutional networks for 3d semantic scene completion,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 3351–3359
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
“Tesla ai day,” https://www.youtube.com/watch?v=j0z4FweCy4M , 2021
2021
Earlier work this paper cites.
T. Yin, X. Zhou, and P. Krahenbuhl, “Center-based 3d object detection and tracking.” in 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , Jun 2021. [Online]. Available: http://dx.doi.org/10.1109/cvpr46437.2021.01161
2021
Earlier work this paper cites.
T. Wang, X. Zhu, J. Pang, and D. Lin, “Fcos3d: Fully convolutional one-stage monocular 3d object detection,” in 2021 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) , Oct 2021. [Online]. Available: http://dx.doi.org/10.1109/iccvw54120.2021.00107
2021
Cited alongside, same era.
2021
Cited alongside, same era.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” Communications of the ACM , vol. 65, no. 1, pp. 99–106, 2021
2021
Cited alongside, same era.
J. T. Barron, B. Mildenhall, M. Tancik, P. Hedman, R. Martin-Brualla, and P. P. Srinivasan, “Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 5855–5864
A. Chen, Z. Xu, A. Geiger, J. Yu, and H. Su, “Tensorf: Tensorial radiance fields,” in European Conference on Computer Vision . Springer, 2022, pp. 333–350
2022
Later among the works it cites.
Y. Hu, J. Yang, L. Chen, K. Li, C. Sima, X. Zhu, S. Chai, S. Du, T. Lin, W. Wang et al. , “Planning-oriented autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 17 853–17 862
2023
Closest in time.
W. Tong, C. Sima, T. Wang, L. Chen, S. Wu, H. Deng, Y. Gu, L. Lu, P. Luo, D. Lin et al. , “Scene as occupancy,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 8406–8415
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows.” in 2021 IEEE/CVF International Conference on Computer Vision (ICCV) , Oct 2021. [Online]. Available: http://dx.doi.org/10.1109/iccv48922.2021.00986
2021
Cited alongside, same era.
Y. Wang, V. C. Guizilini, T. Zhang, Y. Wang, H. Zhao, and J. Solomon, “Detr3d: 3d object detection from multi-view images via 3d-to-2d queries,” in Conference on Robot Learning . PMLR, 2022, pp. 180–191
2022
Cited alongside, same era.
Z. Li, W. Wang, H. Li, E. Xie, C. Sima, T. Lu, Y. Qiao, and J. Dai, “Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers,” in European conference on computer vision . Springer, 2022, pp. 1–18
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, J. Zhou, and J. Lu, “Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 21 729–21 740
2023
Closest in time.
J. Li, M. Lu, J. Liu, Y. Guo, Y. Du, L. Du, and S. Zhang, “Bev-lgkd: A unified lidar-guided knowledge distillation framework for multi-view bev 3d object detection,” IEEE Transactions on Intelligent Vehicles , 2023
2023
Closest in time.
Y. Liu, J. Yan, F. Jia, S. Li, A. Gao, T. Wang, and X. Zhang, “Petrv2: A unified framework for 3d perception from multi-camera images,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 3262–3272
2023
Closest in time.
X. Chi, J. Liu, M. Lu, R. Zhang, Z. Wang, Y. Guo, and S. Zhang, “Bev-san: Accurate bev 3d object detection via slice attention networks,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 17 461–17 470
2023
Closest in time.
Y. Li, Z. Yu, C. Choy, C. Xiao, J. M. Alvarez, S. Fidler, C. Feng, and A. Anandkumar, “Voxformer: Sparse voxel transformer for camera-based 3d semantic scene completion,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 9087–9098
2023
Closest in time.
Y. Huang, W. Zheng, Y. Zhang, J. Zhou, and J. Lu, “Tri-perspective view for vision-based 3d semantic occupancy prediction,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 9223–9232
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
F. Wimbauer, N. Yang, C. Rupprecht, and D. Cremers, “Behind the scenes: Density fields for single view reconstruction,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 9076–9086
2023
Closest in time.
2023
Closest in time.
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, Y. Rao, G. Huang, J. Lu, and J. Zhou, “Surrounddepth: Entangling surrounding views for self-supervised multi-camera depth estimation,” in Conference on Robot Learning . PMLR, 2023, pp. 539–549
2023
Closest in time.
X. Tian, T. Jiang, L. Yun, Y. Mao, H. Yang, Y. Wang, Y. Wang, and H. Zhao, “Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.