Fetching the paper…
Reading the bibliography…
3D semantic occupancy prediction, which seeks to provide accurate and comprehensive representations of environment scenes, is important to autonomous driving systems.
Are we ready for autonomous driving? the kitti vision benchmark suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Focal loss for dense object detection
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár · 2017
Earlier work this paper cites.
Sgdr: Stochastic gradient descent with warm restarts
I. Loshchilov and F. Hutter · 2017
Earlier work this paper cites.
Generalised dice overlap as a deep learning loss function for highly unbalanced segmentations
C. H. Sudre, W. Li, T. Vercauteren, S. Ourselin, and M. Jorge Cardoso · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
3d semantic segmentation with submanifold sparse convolutional networks
B. Graham, M. Engelcke, and L. Van Der Maaten · 2018
Earlier work this paper cites.
Orthographic feature transform for monocular 3d object detection
T. Roddick, A. Kendall, and R. Cipolla · 2018
Earlier work this paper cites.
Pixor: Real-time 3d object detection from point clouds
B. Yang, W. Luo, and R. Urtasun · 2018
Earlier work this paper cites.
Semantickitti: A dataset for semantic scene understanding of lidar sequences
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall · 2019
Earlier work this paper cites.
Rgb and lidar fusion based 3d semantic segmentation for autonomous driving
K. El Madawi, H. Rashed, A. El Sallab, O. Nasr, H. Kamel, and S. Yogamani · 2019
Earlier work this paper cites.
Decoupled weight decay regularization
I. Loshchilov and F. Hutter · 2019
Earlier work this paper cites.
Sensor fusion for joint 3d object detection and semantic segmentation
G. P. Meyer, J. Charland, D. Hegde, A. Laddha, and C. Vallespi-Gonzalez · 2019
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Earlier work this paper cites.
Yolov4: Optimal speed and accuracy of object detection
A. Bochkovskiy, C.-Y. Wang, and H.-Y. M. Liao · 2020
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom · 2020
Earlier work this paper cites.
3d sketch-aware semantic scene completion via semi-supervised structure prior
X. Chen, K.-Y. Lin, C. Qian, G. Zeng, and H. Li · 2020
Earlier work this paper cites.
Salsanext: Fast, uncertainty-aware semantic segmentation of lidar point clouds
T. Cortinhal, G. Tzelepis, and E. Erdal Aksoy · 2020
Earlier work this paper cites.
xmuda: Cross-modal unsupervised domain adaptation for 3d semantic segmentation
M. Jaritz, T.-H. Vu, R. d. Charette, E. Wirbel, and P. Pérez · 2020
Earlier work this paper cites.
Beyond the nav-graph: Vision-and-language navigation in continuous environments
J. Krantz, E. Wijmans, A. Majumdar, D. Batra, and S. Lee · 2020
Earlier work this paper cites.
Anisotropic convolutional networks for 3d semantic scene completion
J. Li, K. Han, P. Wang, Y. Liu, and X. Yuan · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
J. Philion and S. Fidler · 2020
Earlier work this paper cites.
Lmscnet: Lightweight multiscale 3d semantic completion
L. Roldao, R. de Charette, and A. Verroust-Blondet · 2020
Earlier work this paper cites.
Scalability in perception for autonomous driving: Waymo open dataset
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, et al · 2020
Earlier work this paper cites.
Pointpainting: Sequential fusion for 3d object detection
S. Vora, A. H. Lang, B. Helou, and O. Beijbom · 2020
Cited alongside, same era.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
J. Huang, G. Huang, Z. Zhu, Y. Ye, and D. Du · 2021
Cited alongside, same era.
Sparse single sweep lidar point cloud segmentation via learning contextual shape priors from scene completion
X. Yan, J. Gao, J. Li, R. Zhang, Z. Li, R. Huang, and S. Cui · 2021
Cited alongside, same era.
Center-based 3d object detection and tracking
T. Yin, X. Zhou, and P. Krahenbuhl · 2021
Cited alongside, same era.
Panoptic-polarnet: Proposal-free lidar point cloud panoptic segmentation
Z. Zhou, Y. Zhang, and H. Foroosh · 2021
Cited alongside, same era.
Perception-aware multi-sensor fusion for 3d lidar semantic segmentation
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Y. Li, Z. Ge, G. Yu, J. Yang, Z. Wang, Y. Shi, J. Sun, and Z. Li · 2023
Later among the works it cites.
Voxformer: Sparse voxel transformer for camera-based 3d semantic scene completion
Y. Li, Z. Yu, C. Choy, C. Xiao, J. M. Alvarez, S. Fidler, C. Feng, and A. Anandkumar · 2023
Later among the works it cites.
Fb-bev: Bev representation from forward-backward view transformations
Z. Li, Z. Yu, W. Wang, A. Anandkumar, T. Lu, and J. M. Alvarez · 2023
Later among the works it cites.
Uniseg: A unified multi-modal lidar segmentation network and the openpcseg codebase
Y. Liu, R. Chen, X. Li, L. Kong, Y. Yang, Z. Xia, Y. Bai, X. Zhu, Y. Ma, Y. Li, et al · 2023
Later among the works it cites.
A large-scale outdoor multi-modal dataset and benchmark for novel view synthesis and implicit scene reconstruction
C. Lu, F. Yin, X. Chen, W. Liu, T. Chen, G. Yu, and J. Fan · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Zhuang, R. Li, K. Jia, Q. Wang, Y. Li, and M. Tan · 2021
Cited alongside, same era.
Monoscene: Monocular 3d semantic scene completion
A.-Q. Cao and R. de Charette · 2022
Cited alongside, same era.
Learning active camera for multi-object navigation
P. Chen, D. Ji, K. Lin, W. Hu, W. Huang, T. Li, M. Tan, and C. Gan · 2022
Cited alongside, same era.
Efficient and robust 2d-to-bev representation learning via geometry-guided kernel transformer
S. Chen, T. Cheng, X. Wang, W. Meng, Q. Zhang, and W. Liu · 2022
Cited alongside, same era.
Masked autoencoders are scalable vision learners
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick · 2022
Cited alongside, same era.
Multi-camera calibration free bev representation for 3d object detection
H. Jiang, W. Meng, H. Zhu, Q. Zhang, and J. Yin · 2022
Cited alongside, same era.
Deepfusion: Lidar-camera deep fusion for multi-modal 3d object detection
Y. Li, A. W. Yu, T. Meng, B. Caine, J. Ngiam, D. Peng, J. Shen, Y. Lu, D. Zhou, Q. V. Le, et al · 2022
Cited alongside, same era.
Scene as occupancy
W. Tong, C. Sima, T. Wang, L. Chen, S. Wu, H. Deng, Y. Gu, L. Lu, P. Luo, D. Lin, et al · 2023
Later among the works it cites.
Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception
X. Wang, Z. Zhu, W. Xu, Y. Zhang, Y. Wei, X. Chi, Y. Ye, D. Du, J. Lu, and X. Wang · 2023
Later among the works it cites.
Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, J. Zhou, and J. Lu · 2023
Later among the works it cites.
Occformer: Dual-path transformer for vision-based 3d semantic occupancy prediction
Y. Zhang, Z. Zhu, and D. Du · 2023
Later among the works it cites.
Lidar-camera panoptic segmentation via geometry-consistent and semantic-aware alignment
Z. Zhang, Z. Zhang, Q. Yu, R. Yi, Y. Xie, and L. Ma · 2023
Later among the works it cites.
Matrixvt: Efficient multi-camera to bev transformation for 3d perception
H. Zhou, Z. Ge, Z. Li, and X. Zhang · 2023
Later among the works it cites.
Pointocc: Cylindrical tri-perspective view for point-based 3d semantic occupancy prediction
S. Zuo, W. Zheng, Y. Huang, J. Zhou, and J. Lu · 2023
Later among the works it cites.
Symphonize 3d semantic scene completion with contextual instance queries
H. Jiang, T. Cheng, N. Gao, H. Zhang, T. Lin, W. Liu, and X. Wang · 2024
Closest in time.
Cotr: Compact occupancy transformer for vision-based 3d occupancy prediction
Q. Ma, X. Tan, Y. Qu, L. Ma, Z. Zhang, and Y. Xie · 2024
Closest in time.
Co-occ: Coupling explicit feature fusion with volume rendering regularization for multi-modal 3d semantic occupancy prediction
J. Pan, Z. Wang, and L. Wang · 2024
Closest in time.
Epmf: Efficient perception-aware multi-sensor fusion for 3d semantic segmentation
M. Tan, Z. Zhuang, S. Chen, R. Li, K. Jia, Q. Wang, and Y. Li · 2024
Closest in time.
Sparseocc: Rethinking sparse latent representation for vision-based semantic occupancy prediction
P. Tang, Z. Wang, G. Wang, J. Zheng, X. Ren, B. Feng, and C. Ma · 2024
Closest in time.
Occ3d: A large-scale 3d occupancy prediction benchmark for autonomous driving
X. Tian, T. Jiang, L. Yun, Y. Mao, H. Yang, Y. Wang, Y. Wang, and H. Zhao · 2024
Closest in time.
Occgen: Generative multi-modal 3d occupancy prediction for autonomous driving
G. Wang, Z. Wang, P. Tang, J. Zheng, X. Ren, B. Feng, and C. Ma · 2024
Closest in time.
Not all voxels are equal: Hardness-aware semantic scene completion with self-distillation
S. Wang, J. Yu, W. Li, W. Liu, X. Liu, J. Chen, and J. Zhu · 2024
Closest in time.
Panoocc: Unified occupancy representation for camera-based 3d panoptic segmentation
Y. Wang, Y. Chen, X. Liao, L. Fan, and Z. Zhang · 2024
Closest in time.
H2gformer: Horizontal-to-global voxel transformer for 3d semantic scene completion
Y. Wang and C. Tong · 2024
Closest in time.
A survey on occupancy perception for autonomous driving: The information fusion perspective
H. Xu, J. Chen, S. Meng, Y. Wang, and L.-P. Chau · 2025
Closest in time.