Fetching the paper…
Reading the bibliography…
Learning powerful representations in bird's-eye-view (BEV) for perception tasks is trending and drawing extensive attention both from industry and academia.
H. A. Mallot, H. H. Bülthoff, J. Little, and S. Bohrer, “Inverse perspective mapping simplifies optical flow computation and obstacle detection,”
1991
Earlier work this paper cites.
H. Mallot, H. Bülthoff, J. Little, and S. Bohrer, “Inverse perspective mapping simplifies optical flow computation and obstacle detection,”
1991
Earlier work this paper cites.
A. M. Andrew, “Multiple view geometry in computer vision,”
2001
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in
2012
Earlier work this paper cites.
B. Graham, “Spatially-sparse convolutional neural networks,”
2014
Earlier work this paper cites.
S. Song, S. P. Lichtenberg, and J. Xiao, “SUN RGB-D: A rgb-d scene understanding benchmark suite,” in
2015
Earlier work this paper cites.
R. Girshick, “Fast R-CNN,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
X. Chen, K. Kundu, Z. Zhang, H. Ma, S. Fidler, and R. Urtasun, “Monocular 3d object detection for autonomous driving,” in
2016
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster R-CNN: Towards real-time object detection with region proposal networks,”
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-CNN,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “Carla: An open urban driving simulator,” in
2017
Earlier work this paper cites.
A. Dai, A. X. Chang, M. Savva, M. Halber, T. Funkhouser, and M. Nießner, “ScanNet: Richly-annotated 3d reconstructions of indoor scenes,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Earlier work this paper cites.
C. R. Qi, H. Su, K. Mo, and L. J. Guibas, “PointNet: Deep learning on point sets for 3d classification and segmentation,” in
2017
Earlier work this paper cites.
C. R. Qi, L. Yi, H. Su, and L. J. Guibas, “PointNet++: Deep hierarchical feature learning on point sets in a metric space,” in
2017
Earlier work this paper cites.
X. Chen, H. Ma, J. Wan, B. Li, and T. Xia, “Multi-view 3d object detection network for autonomous driving,” in
2017
Earlier work this paper cites.
T.-Y. Lin, P. Dollár, R. B. Girshick, K. He, B. Hariharan, and S. J. Belongie, “Feature pyramid networks for object detection,”
2017
Earlier work this paper cites.
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe, “Unsupervised learning of depth and ego-motion from video,” in
2017
Earlier work this paper cites.
A. Mousavian, D. Anguelov, J. Flynn, and J. Kosecka, “3d bounding box estimation using deep learning and geometry,” in
2017
Earlier work this paper cites.
Y. Zhou and O. Tuzel, “Voxelnet: End-to-end learning for point cloud based 3d object detection,” in
2018
Earlier work this paper cites.
Y. Yan, Y. Mao, and B. Li, “Second: Sparsely embedded convolutional detection,”
2018
Earlier work this paper cites.
B. Xu and Z. Chen, “Multi-level fusion based 3d object detection from monocular images,” in
2018
Earlier work this paper cites.
B. Yang, W. Luo, and R. Urtasun, “PIXOR: Real-time 3d object detection from point clouds,” in
2018
Earlier work this paper cites.
B. Yang, M. Liang, and R. Urtasun, “Hdnet: Exploiting hd maps for 3d object detection,” in
2018
Earlier work this paper cites.
J. Beltrán, C. Guindel, F. M. Moreno, D. Cruzado, F. García, and A. De La Escalera, “BirdNet: A 3d object detection framework from lidar information,” in
2018
Earlier work this paper cites.
Y. Zeng, Y. Hu, S. Liu, J. Ye, Y. Han, X. Li, and N. Sun, “Rt3d: Real-time 3-d vehicle detection in lidar point cloud for autonomous driving,”
2018
Earlier work this paper cites.
W. Ali, S. Abdelkarim, M. Zidan, M. Zahran, and A. El Sallab, “Yolo3d: End-to-end real-time 3d oriented object bounding box detection from lidar point cloud,” in
2018
Earlier work this paper cites.
M. Simony, S. Milzy, K. Amendey, and H.-M. Gross, “Complex-YOLO: An euler-region-proposal for real-time 3d object detection on point clouds,” in
2018
Earlier work this paper cites.
H. Law and J. Deng, “Cornernet: Detecting objects as paired keypoints,” in
2018
Earlier work this paper cites.
M. Berman, A. R. Triki, and M. B. Blaschko, “The lovász-softmax loss: A tractable surrogate for the optimization of the intersection-over-union measure in neural networks,” in
2018
Earlier work this paper cites.
A. Kundu, Y. Li, and J. M. Rehg, “3D-RCNN: Instance-level 3d object reconstruction via render-and-compare,” in
2018
Earlier work this paper cites.
Y. Xu, T. Fan, M. Xu, L. Zeng, and Y. Qiao, “SpiderCNN: Deep learning on point sets with parameterized convolutional filters,” in
2018
Earlier work this paper cites.
Y. Li, R. Bu, M. Sun, W. Wu, X. Di, and B. Chen, “PointCNN: Convolution on x-transformed points,” in
2018
Earlier work this paper cites.
M. Liang, B. Yang, S. Wang, and R. Urtasun, “Deep continuous fusion for multi-sensor 3d object detection,” in
2018
Earlier work this paper cites.
C. R. Qi, W. Liu, C. Wu, H. Su, and L. J. Guibas, “Frustum pointnets for 3d object detection from rgb-d data,” in
2018
Earlier work this paper cites.
J. Ku, M. Mozifian, J. Lee, A. Harakeh, and S. L. Waslander, “Joint 3d proposal generation and object detection from view aggregation,” in
2018
Earlier work this paper cites.
E. Arnold, O. Y. Al-Jarrah, M. Dianati, S. Fallah, D. Oxtoby, and A. Mouzakitis, “A survey on 3d object detection methods for autonomous driving applications,”
2019
Earlier work this paper cites.
M.-F. Chang, J. W. Lambert, P. Sangkloy, J. Singh, S. Bak, A. Hartnett, D. Wang, P. Carr, S. Lucey, D. Ramanan, and J. Hays, “Argoverse: 3d tracking and forecasting with rich maps,” in
2019
Earlier work this paper cites.
X. Huang, P. Wang, X. Cheng, D. Zhou, Q. Geng, and R. Yang, “The apolloscape open dataset for autonomous driving and its application,”
2019
Earlier work this paper cites.
A. Patil, S. Malla, H. Gang, and Y.-T. Chen, “The h3d dataset for full-surround 3d multi-object detection and tracking in crowded urban scenes,” in
2019
Earlier work this paper cites.
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall, “SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR Sequences,” in
2019
Earlier work this paper cites.
T. Roddick, A. Kendall, and R. Cipolla, “Orthographic feature transform for monocular 3d object detection,” in
2019
Earlier work this paper cites.
A. H. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang, and O. Beijbom, “Pointpillars: Fast encoders for object detection from point clouds,” in
2019
Earlier work this paper cites.
N. Garnett, R. Cohen, T. Pe’er, R. Lahav, and D. Levi, “3d-lanenet: end-to-end 3d multiple lane detection,” in
2019
Earlier work this paper cites.
K. He, R. Girshick, and P. Dollár, “Rethinking imagenet pre-training,” in
2019
Earlier work this paper cites.
Y. Wang, W.-L. Chao, D. Garg, B. Hariharan, M. Campbell, and K. Q. Weinberger, “Pseudo-lidar from visual depth estimation: Bridging the gap in 3d object detection for autonomous driving,” in
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Z. Tian, C. Shen, H. Chen, and T. He, “FCOS: Fully convolutional one-stage object detection,” in
2019
Earlier work this paper cites.
C. Choy, J. Gwak, and S. Savarese, “4d spatio-temporal convnets: Minkowski convolutional neural networks,” in
2019
Earlier work this paper cites.
C. R. Qi, O. Litany, K. He, and L. J. Guibas, “Deep hough voting for 3d object detection in point clouds,” in
2019
Earlier work this paper cites.
A. Gordon, H. Li, R. Jonschkowski, and A. Angelova, “Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras,” in
2019
Earlier work this paper cites.
J. Li, Y. Liu, X. Yuan, C. Zhao, R. Siegwart, I. Reid, and C. Cadena, “Depth based semantic scene completion with position importance aware loss,”
2019
Earlier work this paper cites.
X. Zhou, D. Wang, and P. Krähenbühl, “Objects as points,”
2019
Earlier work this paper cites.
G. Brazil and X. Liu, “M3D-RPN: Monocular 3d region proposal network for object detection,” in
2019
Earlier work this paper cites.
J. Ku, A. D. Pon, and S. L. Waslander, “Monocular 3d object detection leveraging accurate proposals and shape reconstruction,” in
2019
Earlier work this paper cites.
F. Manhardt, W. Kehl, and A. Gaidon, “ROI-10D: Monocular lifting of 2d detection to 6d pose and metric shape,” in
2019
Earlier work this paper cites.
S. Shi, X. Wang, and H. Li, “PointRCNN: 3d object proposal generation and detection from point cloud,” in
2019
Earlier work this paper cites.
Y. Wang, Y. Sun, Z. Liu, S. E. Sarma, M. M. Bronstein, and J. M. Solomon, “Dynamic graph cnn for learning on point clouds,”
2019
Earlier work this paper cites.
H. Thomas, C. R. Qi, J.-E. Deschaud, B. Marcotegui, F. Goulette, and L. J. Guibas, “KPConv: Flexible and deformable convolution for point clouds,” in
2019
Earlier work this paper cites.
V. A. Sindagi, Y. Zhou, and O. Tuzel, “MVX-Net: Multimodal voxelnet for 3d object detection,” in
2019
Earlier work this paper cites.
M. Liang, B. Yang, Y. Chen, R. Hu, and R. Urtasun, “Multi-task multi-sensor fusion for 3d object detection,” in
2019
Earlier work this paper cites.
X. Zhang, F. Wan, C. Liu, R. Ji, and Q. Ye, “Freeanchor: Learning to match anchors for visual object detection,” in
2019
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in
2020
Earlier work this paper cites.
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, V. Vasudevan, W. Han, J. Ngiam, H. Zhao, A. Timofeev, S. Ettinger, M. Krivokon, A. Gao, A. Joshi, Y. Zhang, J. Shlens, C. Zhifeng, and D. Anguelov, “Scalability in perception for autonomous driving: Waymo open dataset,” in
2020
Cited alongside, same era.
(2020) Drago Anguelov – Machine Learning for Autonomous Driving at Scale . [Online]. Available:
2020
Cited alongside, same era.
2020
Cited alongside, same era.
J. Houston, G. Zuidhof, L. Bergamini, Y. Ye, L. Chen, A. Jain, S. Omari, V. Iglovikov, and P. Ondruska, “One thousand and one hours: Self-driving motion prediction dataset,” in
2020
Cited alongside, same era.
Z. Liu, Z. Zhang, Y. Cao, H. Hu, and X. Tong, “Group-free 3d object detection via transformers,” in
2021
Later among the works it cites.
H. Zhao, L. Jiang, J. Jia, P. H. Torr, and V. Koltun, “Point transformer,” in
2021
Later among the works it cites.
X. Zhu, H. Zhou, T. Wang, F. Hong, Y. Ma, W. Li, H. Li, and D. Lin, “Cylindrical and asymmetrical 3d convolution networks for lidar segmentation,” in
2021
Later among the works it cites.
R. Cheng, R. Razani, E. Taghavi, E. Li, and B. Liu, “(AF)2-S3Net: Attentive feature fusion with adaptive feature selection for sparse semantic segmentation network,” in
2021
Later among the works it cites.
M. Gerdzhev, R. Razani, E. Taghavi, and L. Bingbing, “TORNADO-Net: multiview total variation semantic segmentation with diamond inception module,” in
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Q.-H. Pham, P. Sevestre, R. S. Pahwa, H. Zhan, C. H. Pang, Y. Chen, A. Mustafa, V. Chandrasekhar, and J. Lin, “A* 3d dataset: Towards autonomous driving in challenging environments,” in
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
L. Reiher, B. Lampe, and L. Eckstein, “A sim2real deep learning approach for the transformation of images from multiple vehicle-mounted cameras to a semantically segmented image in bird’s eye view,” in
2020
Cited alongside, same era.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d,” in
2020
Cited alongside, same era.
C. Yilun, S. Liu, X. Shen, and J. Jia, “Dsgn: Deep stereo geometry network for 3d object detection,”
2020
Cited alongside, same era.
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang, and H. Li, “PV-RCNN: Point-voxel feature set abstraction for 3d object detection,” in
2020
Cited alongside, same era.
S. Vora, A. H. Lang, B. Helou, and O. Beijbom, “PointPainting: Sequential fusion for 3d object detection,” in
2020
Cited alongside, same era.
M. Ye, S. Xu, T. Cao, and Q. Chen, “DRINet: A dual-representation iterative learning network for point cloud segmentation,” in
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Xu, R. Zhang, J. Dou, Y. Zhu, J. Sun, and S. Pu, “RPVNet: A deep and efficient range-point-voxel fusion network for lidar point cloud segmentation,” in
2021
Later among the works it cites.
K. Genova, X. Yin, A. Kundu, C. Pantofaru, F. Cole, A. Sud, B. Brewington, B. Shucker, and T. Funkhouser, “Learning 3d semantic segmentation with only 2d image supervision,” in
2021
Later among the works it cites.
R. Nabati and H. Qi, “CenterFusion: Center-based radar and camera fusion for 3d object detection,” in
2021
Later among the works it cites.
D. Meng, X. Chen, Z. Fan, G. Zeng, H. Li, Y. Yuan, L. Sun, and J. Wang, “Conditional detr for fast training convergence,” in
2021
Later among the works it cites.
R. Solovyev, W. Wang, and T. Gabruseva, “Weighted boxes fusion: Ensembling boxes from different object detection models,”
2021
Later among the works it cites.
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li, and Q. Jiang, “Monocular 3d object detection: An extrinsic parameter free approach,” in
2021
Later among the works it cites.
2021
Later among the works it cites.
E. Xie, Z. Yu, D. Zhou, J. Philion, A. Anandkumar, S. Fidler, P. Luo, and J. M. Alvarez, “M
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
H. Chen, P. Wang, F. Wang, W. Tian, L. Xiong, and H. Li, “EPro-PnP: Generalized end-to-end probabilistic perspective-n-points for monocular object pose estimation,” in
2022
Closest in time.
T. Wang, J. Pang, and D. Lin, “Monocular 3d object detection with depth from motion,”
2022
Closest in time.
K. Han, Y. Wang, H. Chen, X. Chen, J. Guo, Z. Liu, Y. Tang, A. Xiao, C. Xu, Y. Xu
2022
Closest in time.
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick, “Masked autoencoders are scalable vision learners,” in
2022
Closest in time.
R. Qian, X. Lai, and X. Li, “3d object detection for autonomous driving: a survey,”
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
F. Yan, M. Nie, X. Cai, J. Han, H. Xu, Z. Yang, C. Ye, Y. Fu, M. B. Mi, and L. Zhang, “Once-3dlanes: Building monocular 3d lane detection,” in
2022
Closest in time.
T. Wang, W. Ji, S. Chen, G. Chongjian, E. Xie, and P. Luo, “DeepAccident: A large-scale accident dataset for multi-vehicle autonomous driving,” 2022
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
D. Rukhovich, A. Vorontsova, and A. Konushin, “Imvoxelnet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection,” in
2022
Closest in time.
2022
Closest in time.
B. Zhou and P. Krähenbühl, “Cross-view transformers for real-time map-view semantic segmentation,” in
2022
Closest in time.
Q. Li, Y. Wang, Y. Wang, and H. Zhao, “Hdmapnet: An online hd map construction and evaluation framework,” in
2022
Closest in time.
A. Saha, O. Mendez, C. Russell, and R. Bowden, “Translating images into maps,” in
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
T. Wang, Z. Xinge, J. Pang, and D. Lin, “Probabilistic and geometric depth: Detecting objects in perspective,” in
2022
Closest in time.
2022
Closest in time.
J. Huang and G. Huang, “BEVDet4D: Exploit temporal cues in multi-camera 3d object detection,”
2022
Closest in time.
L. Fan, Z. Pang, T. Zhang, Y.-X. Wang, H. Zhao, F. Wang, N. Wang, and Z. Zhang, “Embracing single stride 3d object detector with sparse transformer,” in
2022
Closest in time.
Y. Hu, Z. Ding, R. Ge, W. Shao, L. Huang, K. Li, and Q. Liu, “AFDetV2: Rethinking the necessity of the second stage for object detection from point clouds,”
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
X. Bai, Z. Hu, X. Zhu, Q. Huang, Y. Chen, H. Fu, and C.-L. Tai, “TransFusion: Robust lidar-camera fusion for 3d object detection with transformers,” in
2022
Closest in time.
Y. Li, A. W. Yu, T. Meng, B. Caine, J. Ngiam, D. Peng, J. Shen, B. Wu, Y. Lu, D. Zhou, Q. V. Le, A. Yuille, and M. Tan, “Deepfusion: Lidar-camera deep fusion for multi-modal 3d object detection,” in
2022
Closest in time.
2022
Closest in time.
Y. Wang, V. C. Guizilini, T. Zhang, Y. Wang, H. Zhao, and J. Solomon, “Detr3d: 3d object detection from multi-view images via 3d-to-2d queries,” in
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
T. Wang, Q. Lian, C. Zhu, X. Zhu, and W. Zhang, “MV-FCOS3D++: Multi-View camera-only 4d object detection with pretrained monocular backbones,”
2022
Closest in time.
N. Gosala and A. Valada, “Bird’s-eye-view panoptic segmentation using monocular frontal view images,”
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
(2022) Horizon Algorithms . [Online]. Available:
2022
Closest in time.
(2022) PhiGent: Technical Roadmap . [Online]. Available:
2022
Closest in time.
(2022) HAOMO AI DAY . [Online]. Available:
2022
Closest in time.
H. Wang, S. Shi, Z. Yang, R. Fang, Q. Qian, H. Li, B. Schiele, and L. Wang, “RBGNet: Ray-based grouping for 3d object detection,” in
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
(2022) ViDAR – Forward Visual Perception for Advanced Self-Driving . [Online]. Available:
2022
Closest in time.
F. Li, H. Zhang, S. Liu, J. Guo, L. M. Ni, and L. Zhang, “Dn-detr: Accelerate detr training by introducing query denoising,” in
2022
Closest in time.
P. Wang, A. Yang, R. Men, J. Lin, S. Bai, Z. Li, J. Ma, C. Zhou, J. Zhou, and H. Yang, “OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework,” in
2022
Closest in time.
2022
Closest in time.