Fetching the paper…
Reading the bibliography…
Multi-camera perception tasks have gained significant attention in the field of autonomous driving.
Learning from unlabelled videos using contrastive predictive neural 3d mapping
Harley, A. W.; Lakshmikanth, S. K.; Li, F.; Zhou, X.; Tung, H.-Y. F.; and Fragkiadaki, K. 2019 · 1906
Earlier work this paper cites.
Optical models for direct volume rendering
Max, N. 1995 · 1995
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
A survey on transfer learning
Pan, S. J.; and Yang, Q. 2009 · 2009
Earlier work this paper cites.
Mayavi: 3D visualization of scientific data
Ramachandran, P.; and Varoquaux, G. 2011 · 2011
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Pyramid stereo matching network
Chang, J.-R.; and Chen, Y.-S. 2018 · 2018
Earlier work this paper cites.
Geometry-aware recurrent neural networks for active visual recognition
Cheng, R.; Wang, Z.; and Fragkiadaki, K. 2018 · 2018
Earlier work this paper cites.
Image inpainting for irregular holes using partial convolutions
Liu, G.; Reda, F. A.; Shih, K. J.; Wang, T.-C.; Tao, A.; and Catanzaro, B. 2018 · 2018
Earlier work this paper cites.
Rangenet++: Fast and accurate lidar semantic segmentation
Milioto, A.; Vizzo, I.; Behley, J.; and Stachniss, C. 2019 · 2019
Earlier work this paper cites.
Deepsdf: Learning continuous signed distance functions for shape representation
Park, J. J.; Florence, P.; Straub, J.; Newcombe, R.; and Lovegrove, S. 2019 · 2019
Earlier work this paper cites.
Deepvoxels: Learning persistent 3d feature embeddings
Sitzmann, V.; Thies, J.; Heide, F.; Nießner, M.; Wetzstein, G.; and Zollhofer, M. 2019 · 2019
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
Caesar, H.; Bankiti, V.; Lang, A. H.; Vora, S.; Liong, V. E.; Xu, Q.; Krishnan, A.; Pan, Y.; Baldan, G.; and Beijbom, O. 2020 · 2020
Earlier work this paper cites.
Deep local shapes: Learning local sdf priors for detailed 3d reconstruction
Chabra, R.; Lenssen, J. E.; Ilg, E.; Schmidt, T.; Straub, J.; Lovegrove, S.; and Newcombe, R. 2020 · 2020
Earlier work this paper cites.
3d sketch-aware semantic scene completion via semi-supervised structure prior
Chen, X.; Lin, K.-Y.; Qian, C.; Zeng, G.; and Li, H. 2020 · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Philion, J.; and Fidler, S. 2020 · 2020
Earlier work this paper cites.
Lmscnet: Lightweight multiscale 3d semantic completion
Roldao, L.; de Charette, R.; and Verroust-Blondet, A. 2020 · 2020
Cited alongside, same era.
Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields
Barron, J. T.; Mildenhall, B.; Tancik, M.; Hedman, P.; Martin-Brualla, R.; and Srinivasan, P. P. 2021 · 2021
Cited alongside, same era.
FIERY: future instance prediction in bird’s-eye view from surround monocular cameras
Hu, A.; Murez, Z.; Mohan, N.; Dudas, S.; Hawke, J.; Badrinarayanan, V.; Cipolla, R.; and Kendall, A. 2021 · 2021
Cited alongside, same era.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
Huang, J.; Huang, G.; Zhu, Z.; Ye, Y.; and Du, D. 2021 · 2021
Cited alongside, same era.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2021 · 2021
Cited alongside, same era.
Tesla AI Day
Tesla. 2022 · 2022
Later among the works it cites.
Detr3d: 3d object detection from multi-view images via 3d-to-2d queries
Wang, Y.; Guizilini, V. C.; Zhang, T.; Wang, Y.; Zhao, H.; and Solomon, J. 2022 · 2022
Later among the works it cites.
Mˆ 2bev: Multi-camera joint 3d detection and segmentation with unified birds-eye view representation
Xie, E.; Yu, Z.; Zhou, D.; Philion, J.; Anandkumar, A.; Fidler, S.; Luo, P.; and Alvarez, J. M. 2022 · 2022
Later among the works it cites.
Lidarmultinet: Towards a unified multi-task network for lidar perception
Ye, D.; Zhou, Z.; Chen, W.; Xie, Y.; Wang, Y.; Wang, P.; and Foroosh, H. 2022 · 2022
Later among the works it cites.
Beverse: Unified perception and prediction in birds-eye-view for vision-centric autonomous driving
Zhang, Y.; Zhu, Z.; Zheng, W.; Huang, J.; Huang, G.; Zhou, J.; and Lu, J. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Categorical depth distribution network for monocular 3d object detection
Reading, C.; Harakeh, A.; Chae, J.; and Waslander, S. L. 2021 · 2021
Cited alongside, same era.
NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view Reconstruction
Wang, P.; Liu, L.; Liu, Y.; Theobalt, C.; Komura, T.; and Wang, W. 2021 · 2021
Cited alongside, same era.
Volume rendering of neural implicit surfaces
Yariv, L.; Gu, J.; Kasten, Y.; and Lipman, Y. 2021 · 2021
Cited alongside, same era.
Drinet++: Efficient voxel-as-point point cloud segmentation
Ye, M.; Wan, R.; Xu, S.; Cao, T.; and Chen, Q. 2021 · 2021
Cited alongside, same era.
In-place scene labelling and understanding with implicit scene representation
Zhi, S.; Laidlow, T.; Leutenegger, S.; and Davison, A. J. 2021 · 2021
Cited alongside, same era.
Cylindrical and asymmetrical 3d convolution networks for lidar segmentation
Zhu, X.; Zhou, H.; Wang, T.; Hong, F.; Ma, Y.; Li, W.; Li, H.; and Lin, D. 2021 · 2021
Cited alongside, same era.
Monoscene: Monocular 3d semantic scene completion
Cao, A.-Q.; and de Charette, R. 2022 · 2022
Cited alongside, same era.
Later among the works it cites.
Cross-view transformers for real-time map-view semantic segmentation
Zhou, B.; and Krähenbühl, P. 2022 · 2022
Later among the works it cites.
A Simple Attempt for 3D Occupancy Estimation in Autonomous Driving
Gan, W.; Mo, N.; Xu, H.; and Yokoya, N. 2023 · 2023
Closest in time.
Tri-Perspective View for Vision-Based 3D Semantic Occupancy Prediction
Huang, Y.; Zheng, W.; Zhang, Y.; Zhou, J.; and Lu, J. 2023 · 2023
Closest in time.
LERF: Language Embedded Radiance Fields
Kerr, J.; Kim, C. M.; Goldberg, K.; Kanazawa, A.; and Tancik, M. 2023 · 2023
Closest in time.
FB-BEV: BEV Representation from Forward-Backward View Transformations
Li, Z.; Yu, Z.; Wang, W.; Anandkumar, A.; Lu, T.; and Alvarez, J. M. 2023 · 2023
Closest in time.
UniOcc: Unifying Vision-Centric 3D Occupancy Prediction with Geometric and Semantic Rendering
Pan, M.; Liu, L.; Liu, J.; Huang, P.; Wang, L.; Zhang, S.; Xu, S.; Lai, Z.; and Yang, K. 2023 · 2023
Closest in time.
Scene as Occupancy
Sima, C.; Tong, W.; Wang, T.; Chen, L.; Wu, S.; Deng, H.; Gu, Y.; Lu, L.; Luo, P.; Lin, D.; and Li, H. 2023 · 2023
Closest in time.
Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving
Tian, X.; Jiang, T.; Yun, L.; Wang, Y.; Wang, Y.; and Zhao, H. 2023 · 2023
Closest in time.
OpenOccupancy: A Large Scale Benchmark for Surrounding Semantic Occupancy Perception
Wang, X.; Zhu, Z.; Xu, W.; Zhang, Y.; Wei, Y.; Chi, X.; Ye, Y.; Du, D.; Lu, J.; and Wang, X. 2023 · 2023
Closest in time.
SurroundOcc: Multi-Camera 3D Occupancy Prediction for Autonomous Driving
Wei, Y.; Zhao, L.; Zheng, W.; Zhu, Z.; Zhou, J.; and Lu, J. 2023 · 2023
Closest in time.
OccFormer: Dual-path Transformer for Vision-based 3D Semantic Occupancy Prediction
Zhang, Y.; Zhu, Z.; and Du, D. 2023 · 2023
Closest in time.