Fetching the paper…
Reading the bibliography…
3D occupancy prediction is an emerging task that aims to estimate the occupancy states and semantics of 3D scenes using multi-view images.
Cylinder3d: An effective 3d framework for driving-scene lidar semantic segmentation
Zhou, H.; Zhu, X.; Song, X.; Ma, Y.; Wang, Z.; Li, H.; and Lin, D. 2020 · 2008
Earlier work this paper cites.
SLIC superpixels compared to state-of-the-art superpixel methods
Achanta, R.; Shaji, A.; Smith, K.; Lucchi, A.; Fua, P.; and Süsstrunk, S. 2012 · 2012
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
Eigen, D.; Puhrsch, C.; and Fergus, R. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G.; Vinyals, O.; and Dean, J. 2015 · 2015
Earlier work this paper cites.
Deformable convolutional networks
Dai, J.; Qi, H.; Xiong, Y.; Li, Y.; Zhang, G.; Hu, H.; and Wei, Y. 2017 · 2017
Earlier work this paper cites.
Semantic scene completion from a single depth image
Song, S.; Yu, F.; Zeng, A.; Chang, A. X.; Savva, M.; and Funkhouser, T. 2017 · 2017
Earlier work this paper cites.
Second: Sparsely embedded convolutional detection
Yan, Y.; Mao, Y.; and Li, B. 2018 · 2018
Earlier work this paper cites.
Pointpillars: Fast encoders for object detection from point clouds
Lang, A. H.; Vora, S.; Caesar, H.; Zhou, L.; Yang, J.; and Beijbom, O. 2019 · 2019
Earlier work this paper cites.
Structured knowledge distillation for semantic segmentation
Liu, Y.; Chen, K.; Liu, C.; Qin, Z.; Luo, Z.; and Wang, J. 2019 · 2019
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
Caesar, H.; Bankiti, V.; Lang, A. H.; Vora, S.; Liong, V. E.; Xu, Q.; Krishnan, A.; Pan, Y.; Baldan, G.; and Beijbom, O. 2020 · 2020
Earlier work this paper cites.
Inter-region affinity distillation for road marking segmentation
Hou, Y.; Ma, Z.; Liu, C.; Hui, T.-W.; and Loy, C. C. 2020 · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Philion, J.; and Fidler, S. 2020 · 2020
Earlier work this paper cites.
From points to parts: 3d object detection from point cloud with part-aware and part-aggregation network
Shi, S.; Wang, Z.; Shi, J.; Wang, X.; and Li, H. 2020 · 2020
Earlier work this paper cites.
Improve object detection with feature-based knowledge distillation: Towards accurate and efficient detectors
Zhang, L.; and Ma, K. 2020 · 2020
Cited alongside, same era.
Polarnet: An improved grid representation for online lidar point clouds semantic segmentation
Zhang, Y.; Zhou, Z.; David, P.; Yue, X.; Xi, Z.; Gong, B.; and Foroosh, H. 2020 · 2020
Cited alongside, same era.
General instance distillation for object detection
Dai, X.; Jiang, Z.; Wu, Z.; Bao, Y.; Wang, Z.; Liu, S.; and Zhou, E. 2021 · 2021
Cited alongside, same era.
Distilling object detectors via decoupled features
Guo, J.; Han, K.; Wang, Y.; Wu, H.; Chen, X.; Xu, C.; and Xu, C. 2021 · 2021
Cited alongside, same era.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
Huang, J.; Huang, G.; Zhu, Z.; Ye, Y.; and Du, D. 2021 · 2021
Cited alongside, same era.
Point-to-voxel knowledge distillation for lidar semantic segmentation
Hou, Y.; Zhu, X.; Ma, Y.; Loy, C. C.; and Li, Y. 2022 · 2022
Later among the works it cites.
Learning ego 3d representation as ray tracing
Lu, J.; Zhou, Z.; Zhu, X.; Xu, H.; and Zhang, L. 2022 · 2022
Later among the works it cites.
X-trans2cap: Cross-modal knowledge transfer using transformer for 3d dense captioning
Yuan, Z.; Yan, X.; Liao, Y.; Guo, Y.; Li, G.; Cui, S.; and Li, Z. 2022 · 2022
Later among the works it cites.
A Simple Attempt for 3D Occupancy Estimation in Autonomous Driving
Gan, W.; Mo, N.; Xu, H.; and Yokoya, N. 2023 · 2023
Closest in time.
Tri-perspective view for vision-based 3d semantic occupancy prediction
Huang, Y.; Zheng, W.; Zhang, Y.; Zhou, J.; and Lu, J. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Swin transformer: Hierarchical vision transformer using shifted windows
Liu, Z.; Lin, Y.; Cao, Y.; Hu, H.; Wei, Y.; Zhang, Z.; Lin, S.; and Guo, B. 2021 · 2021
Cited alongside, same era.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2021 · 2021
Cited alongside, same era.
Fcos3d: Fully convolutional one-stage monocular 3d object detection
Wang, T.; Zhu, X.; Pang, J.; and Lin, D. 2021 · 2021
Cited alongside, same era.
Sparse single sweep lidar point cloud segmentation via learning contextual shape priors from scene completion
Yan, X.; Gao, J.; Li, J.; Zhang, R.; Li, Z.; Huang, R.; and Cui, S. 2021 · 2021
Cited alongside, same era.
Monoscene: Monocular 3d semantic scene completion
Cao, A.-Q.; and de Charette, R. 2022 · 2022
Cited alongside, same era.
Bevdistill: Cross-modal bev distillation for multi-view 3d object detection
Chen, Z.; Li, Z.; Zhang, S.; Fang, L.; Jiang, Q.; and Zhao, F. 2022 · 2022
Cited alongside, same era.
Monodistill: Learning spatial features for monocular 3d object detection
Chong, Z.; Ma, X.; Zhang, H.; Yue, Y.; Li, H.; Wang, Z.; and Ouyang, W. 2022 · 2022
Cited alongside, same era.
Jiang, Y.; Zhang, L.; Miao, Z.; Zhu, X.; Gao, J.; Hu, W.; and Jiang, Y.-G. 2023 · 2023
Closest in time.
Kirillov, A.; Mintun, E.; Ravi, N.; Mao, H.; Rolland, C.; Gustafson, L.; Xiao, T.; Whitehead, S.; Berg, A. C.; Lo, W.-Y.; et al. 2023 · 2023
Closest in time.
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Li, Y.; Ge, Z.; Yu, G.; Yang, J.; Wang, Z.; Shi, Y.; Sun, J.; and Li, Z. 2023 · 2023
Closest in time.
Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation
Liu, Z.; Tang, H.; Amini, A.; Yang, X.; Mao, H.; Rus, D. L.; and Han, S. 2023 · 2023
Closest in time.
Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving
Tian, X.; Jiang, T.; Yun, L.; Wang, Y.; Wang, Y.; and Zhao, H. 2023 · 2023
Closest in time.
Tong, W.; Sima, C.; Wang, T.; Wu, S.; Deng, H.; Chen, L.; Gu, Y.; Lu, L.; Luo, P.; Lin, D.; et al. 2023 · 2023
Closest in time.
Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving
Wei, Y.; Zhao, L.; Zheng, W.; Zhu, Z.; Zhou, J.; and Lu, J. 2023 · 2023
Closest in time.
CPU: Codebook Lookup Transformer with Knowledge Distillation for Point Cloud Upsampling
Zhao, W.; Zhang, H.; Zheng, C.; Yan, X.; Cui, S.; and Li, Z. 2023 · 2023
Closest in time.