Fetching the paper…
Reading the bibliography…
In recent years, autonomous driving has garnered escalating attention for its potential to relieve drivers' burdens and improve driving safety.
Multi-view 3d object detection network for autonomous driving
Chen X, Ma H, Wan J, Li B, Xia T · 1915
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
Silberman N, Hoiem D, Kohli P, Fergus R · 2012
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
Geiger A, Lenz P, Urtasun R · 2012
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger O, Fischer P, Brox T · 2015
Earlier work this paper cites.
Monocular 3d object detection for autonomous driving
Chen X, Kundu K, Zhang Z, Ma H, Fidler S, Urtasun R · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He K, Zhang X, Ren S, Sun J · 2016
Earlier work this paper cites.
Feature pyramid networks for object detection
Lin T Y, Dollár P, Girshick R, He K, Hariharan B, Belongie S · 2017
Earlier work this paper cites.
Second: Sparsely embedded convolutional detection
Yan Y, Mao Y, Li B · 2018
Earlier work this paper cites.
Pseudo-lidar from visual depth estimation: Bridging the gap in 3d object detection for autonomous driving
Wang Y, Chao W L, Garg D, Hariharan B, Campbell M, Weinberger K Q · 2019
Earlier work this paper cites.
Stereo r-cnn based 3d object detection for autonomous driving
Li P, Chen X, Shen S · 2019
Earlier work this paper cites.
Pointrcnn: 3d object proposal generation and detection from point cloud
Shi S, Wang X, Li H · 2019
Earlier work this paper cites.
Occupancy networks: Learning 3d reconstruction in function space
Mescheder L, Oechsle M, Niemeyer M, Nowozin S, Geiger A · 2019
Earlier work this paper cites.
Semantickitti: A dataset for semantic scene understanding of lidar sequences
Behley J, Garbade M, Milioto A, Quenzel J, Behnke S, Stachniss C, Gall J · 2019
Earlier work this paper cites.
Dsgn: Deep stereo geometry network for 3d object detection
Chen Y, Liu S, Shen X, Jia J · 2020
Earlier work this paper cites.
Pointpainting: Sequential fusion for 3d object detection
Vora S, Lang A H, Helou B, Beijbom O · 2020
Earlier work this paper cites.
Clocs: Camera-lidar object candidates fusion for 3d object detection
Pang S, Morris D, Radha H · 2020
Earlier work this paper cites.
Epnet: Enhancing point features with image semantics for 3d object detection
Huang T, Liu Z, Chen X, Bai X · 2020
Earlier work this paper cites.
Convolutional occupancy networks
Peng S, Niemeyer M, Mescheder L, Pollefeys M, Geiger A · 2020
Earlier work this paper cites.
nuscenes: A multimodal dataset for autonomous driving
Caesar H, Bankiti V, Lang A H, Vora S, Liong V E, Xu Q, Krishnan A, Pan Y, Baldan G, Beijbom O · 2020
Earlier work this paper cites.
Scalability in perception for autonomous driving: Waymo open dataset
Sun P, Kretzschmar H, Dotiwalla X, Chouard A, Patnaik V, Tsui P, Guo J, Zhou Y, Chai Y, Caine B, others · 2020
Earlier work this paper cites.
3d sketch-aware semantic scene completion via semi-supervised structure prior
Chen X, Lin K Y, Qian C, Zeng G, Li H · 2020
Earlier work this paper cites.
Anisotropic convolutional networks for 3d semantic scene completion
Li J, Han K, Wang P, Liu Y, Yuan X · 2020
Earlier work this paper cites.
Lmscnet: Lightweight multiscale 3d semantic completion
Roldao L, Charette d R, Verroust-Blondet A · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Philion J, Fidler S · 2020
Earlier work this paper cites.
Deformable detr: Deformable transformers for end-to-end object detection
Zhu X, Su W, Lu L, Li B, Wang X, Dai J · 2020
Earlier work this paper cites.
Center-based 3d object detection and tracking
Yin T, Zhou X, Krahenbuhl P · 2021
Earlier work this paper cites.
Voxel r-cnn: Towards high performance voxel-based 3d object detection
Deng J, Shi S, Li P, Zhou W, Zhang Y, Li H · 2021
Earlier work this paper cites.
Pc-rgnn: Point cloud completion and graph neural network for 3d object detection
Zhang Y, Huang D, Wang Y · 2021
Earlier work this paper cites.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
Huang J, Huang G, Zhu Z, Ye Y, Du D · 2021
Earlier work this paper cites.
Sparse single sweep lidar point cloud segmentation via learning contextual shape priors from scene completion
Yan X, Gao J, Li J, Zhang R, Li Z, Huang R, Cui S · 2021
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Mildenhall B, Srinivasan P P, Tancik M, Barron J T, Ramamoorthi R, Ng R · 2021
Earlier work this paper cites.
Cat-det: Contrastively augmented transformer for multi-modal 3d object detection
Zhang Y, Chen J, Huang D · 2022
Cited alongside, same era.
Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers
Li Z, Wang W, Li H, Xie E, Sima C, Lu T, Qiao Y, Dai J · 2022
Cited alongside, same era.
Vision-centric bev perception: A survey
Ma Y, Wang T, Bai X, Yang H, Hou Y, Wang Y, Qiao Y, Yang R, Manocha D, Zhu X · 2022
Cited alongside, same era.
Vdbfusion: Flexible and efficient tsdf integration of range sensor data
Vizzo I, Guadagnino T, Behley J, Stachniss C · 2022
Cited alongside, same era.
Monoscene: Monocular 3d semantic scene completion
Cao A Q, Charette d R · 2022
Cited alongside, same era.
Masked autoencoders are scalable vision learners
Ovo: Open-vocabulary occupancy
Tan Z, Dong Z, Zhang C, Zhang W, Ji H, Li H · 2023
Later among the works it cites.
Selfocc: Self-supervised vision-based 3d occupancy prediction
Huang Y, Zheng W, Zhang B, Zhou J, Lu J · 2023
Later among the works it cites.
Occnerf: Self-supervised multi-camera occupancy prediction with neural radiance fields
Zhang C, Yan J, Wei Y, Li J, Liu L, Tang Y, Duan Y, Lu J · 2023
Later among the works it cites.
Renderocc: Vision-centric 3d occupancy prediction with 2d rendering supervision
Pan M, Liu J, Zhang R, Huang P, Li X, Liu L, Zhang S · 2023
Later among the works it cites.
Uniocc: Unifying vision-centric 3d occupancy prediction with geometric and semantic rendering
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
He K, Chen X, Xie S, Li Y, Dollár P, Girshick R · 2022
Cited alongside, same era.
Masked-attention mask transformer for universal image segmentation
Cheng B, Misra I, Schwing A G, Kirillov A, Girdhar R · 2022
Cited alongside, same era.
3d object detection from images for autonomous driving: a survey
Ma X, Ouyang W, Simonelli A, Ricci E · 2023
Cited alongside, same era.
Octr: Octree-based transformer for 3d object detection
Zhou C, Zhang Y, Chen J, Huang D · 2023
Cited alongside, same era.
Multi-modal 3d object detection in autonomous driving: A survey and taxonomy
Wang L, Zhang X, Song Z, Bi J, Zhang G, Wei H, Tang L, Yang L, Li J, Jia C, others · 2023
Cited alongside, same era.
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Li Y, Ge Z, Yu G, Yang J, Wang Z, Shi Y, Sun J, Li Z · 2023
Cited alongside, same era.
Sa-bev: Generating semantic-aware bird’s-eye-view feature for multi-view 3d object detection
Zhang J, Zhang Y, Liu Q, Wang Y · 2023
Cited alongside, same era.
Pan M, Liu L, Liu J, Huang P, Wang L, Zhang S, Xu S, Lai Z, Yang K · 2023
Later among the works it cites.
Fb-occ: 3d occupancy prediction based on forward-backward view transformation
Li Z, Yu Z, Austin D, Fang M, Lan S, Kautz J, Alvarez J M · 2023
Later among the works it cites.
Pointocc: Cylindrical tri-perspective view for point-based 3d semantic occupancy prediction
Zuo S, Zheng W, Huang Y, Zhou J, Lu J · 2023
Later among the works it cites.
Occdepth: A depth-aware method for 3d semantic scene completion
Miao R, Liu W, Chen M, Gong Z, Xu W, Hu C, Zhou S · 2023
Later among the works it cites.
Cotr: Compact occupancy transformer for vision-based 3d occupancy prediction
Ma Q, Tan X, Qu Y, Ma L, Zhang Z, Xie Y · 2023
Later among the works it cites.
Learning occupancy for monocular 3d object detection
Peng L, Xu J, Cheng H, Yang Z, Wu X, Qian W, Wang W, Wu B, Cai D · 2023
Later among the works it cites.
A simple attempt for 3d occupancy estimation in autonomous driving
Gan W, Mo N, Xu H, Yokoya N · 2023
Later among the works it cites.
Regulating intermediate 3d features for vision-centric autonomous driving
Xu J, Peng L, Cheng H, Xia L, Zhou Q, Deng D, Qian W, Wang W, Cai D · 2023
Later among the works it cites.
Behind the scenes: Density fields for single view reconstruction
Wimbauer F, Yang N, Rupprecht C, Cremers D · 2023
Later among the works it cites.
Panacea: Panoramic and controllable video generation for autonomous driving
Wen Y, Zhao Y, Liu Y, Jia F, Wang Y, Luo C, Zhang C, Wang T, Sun X, Zhang X · 2023
Later among the works it cites.
Gaia-1: A generative world model for autonomous driving
Hu A, Russell L, Yeo H, Murez Z, Fedoseev G, Kendall A, Shotton J, Corrado G · 2023
Later among the works it cites.
Wovogen: World volume-aware diffusion for controllable multi-camera driving scene generation
Lu J, Huang Z, Zhang J, Yang Z, Zhang L · 2023
Later among the works it cites.
Occworld: Learning a 3d occupancy world model for autonomous driving
Zheng W, Chen W, Huang Y, Zhang B, Duan Y, Lu J · 2023
Later among the works it cites.
Uniworld: Autonomous driving pre-training via world models
Min C, Zhao D, Xiao L, Nie Y, Dai B · 2023
Later among the works it cites.
Cam4docc: Benchmark for camera-only 4d occupancy forecasting in autonomous driving applications
Ma J, Chen X, Huang J, Xu J, Luo Z, Xu J, Gu W, Ai R, Wang H · 2023
Later among the works it cites.
Occtransformer: Improving bevformer for 3d camera-only occupancy prediction
Liu J, Zhang S, Kong C, Zhang W, Wu Y, Ding Y, Xu B, Ming R, Wei D, Liu X · 2024
Closest in time.
Fastocc: Accelerating 3d occupancy prediction by fusing the 2d bird’s-eye view and perspective view
Hou J, Li X, Guan W, Zhang G, Feng D, Du Y, Xue X, Pu J · 2024
Closest in time.
Monoocc: Digging into monocular semantic occupancy prediction
Zheng Y, Li X, Li P, Zheng Y, Jin B, Zhong C, Long X, Zhao H, Zhang Q · 2024
Closest in time.
Radocc: Learning cross-modality occupancy knowledge through rendering assisted distillation
Zhang H, Yan X, Bai D, Gao J, Wang P, Liu B, Cui S, Li Z · 2024
Closest in time.
Boeder S, Gigengack F, Risse B · 2024
Closest in time.
Silva S, Wannigama S B, Ragel R, Jayatilaka G · 2024
Closest in time.
Inversematrixvt3d: An efficient projection matrix-based approach for 3d occupancy prediction
Ming Z, Berrio J S, Shan M, Worrall S · 2024
Closest in time.
Univision: A unified framework for vision-centric 3d perception
Hong Y, Liu Q, Cheng H, Ma D, Dai H, Wang Y, Cao G, Ding Y · 2024
Closest in time.
Song R, Liang C, Cao H, Yan Z, Zimmer W, Gross M, Festag A, Knoll A · 2024
Closest in time.
Pop-3d: Open-vocabulary 3d occupancy prediction from images
Vobecky A, Siméoni O, Hurych D, Gidaris S, Bursuc A, Pérez P, Sivic J · 2024
Closest in time.
Deep manta: A coarse-to-fine many-task network for joint 2d and 3d vehicle analysis from monocular image
Chabot F, Chaouch M, Rabarisoa J, Teuliere C, Chateau T · 2049
Closest in time.