Fetching the paper…
Reading the bibliography…
In this paper, we propose M$^2$BEV, a unified framework that jointly performs 3D object detection and map segmentation in the Birds Eye View~(BEV) space with multi-camera image inputs.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Monocular 3d object detection for autonomous driving
Xiaozhi Chen, Kaustav Kundu, Ziyu Zhang, Huimin Ma, Sanja Fidler, and Raquel Urtasun · 2016
Earlier work this paper cites.
Fast single shot detection and pose estimation
Patrick Poirson, Phil Ammirato, Cheng-Yang Fu, Wei Liu, Jana Kosecka, and Alexander C Berg · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
SSD-6D: Making rgb-based 3d detection and 6d pose estimation great again
Wadim Kehl, Fabian Manhardt, Federico Tombari, Slobodan Ilic, and Nassir Navab · 2017
Earlier work this paper cites.
Aggregated residual transformations for deep neural networks
Saining Xie, Ross Girshick, Piotr Dollár, Zhuowen Tu, and Kaiming He · 2017
Earlier work this paper cites.
Deformable convolutional networks
Jifeng Dai, Haozhi Qi, Yuwen Xiong, Yi Li, Guodong Zhang, Han Hu, and Yichen Wei · 2017
Earlier work this paper cites.
SECOND: Sparsely embedded convolutional detection
Yan Yan, Yuxing Mao, and Bo Li · 2018
Earlier work this paper cites.
MegDet: A large mini-batch object detector
Chao Peng, Tete Xiao, Zeming Li, Yuning Jiang, Xiangyu Zhang, Kai Jia, Gang Yu, and Jian Sun · 2018
Earlier work this paper cites.
Mixed precision training
Paulius Micikevicius, Sharan Narang, Jonah Alben, Gregory Diamos, Erich Elsen, David Garcia, Boris Ginsburg, Michael Houston, Oleksii Kuchaiev, Ganesh Venkatesh, et al · 2018
Earlier work this paper cites.
PointPillars: Fast encoders for object detection from point clouds
Alex H Lang, Sourabh Vora, Holger Caesar, Lubing Zhou, Jiong Yang, and Oscar Beijbom · 2019
Earlier work this paper cites.
PointRCNN: 3d object proposal generation and detection from point cloud
Shaoshuai Shi, Xiaogang Wang, and Hongsheng Li · 2019
Earlier work this paper cites.
Orthographic feature transform for monocular 3d object detection
Thomas Roddick, Alex Kendall, and Roberto Cipolla · 2019
Earlier work this paper cites.
Pseudo-lidar from visual depth estimation: Bridging the gap in 3d object detection for autonomous driving
Yan Wang, Wei-Lun Chao, Divyansh Garg, Bharath Hariharan, Mark Campbell, and Kilian Q Weinberger · 2019
Earlier work this paper cites.
Xingyi Zhou, Dequan Wang, and Philipp Krähenbühl · 2019
Cited alongside, same era.
MonoGRNet: A geometric reasoning network for monocular 3d object localization
Zengyi Qin, Jinglu Wang, and Yan Lu · 2019
Cited alongside, same era.
Disentangling monocular 3d object detection
Andrea Simonelli, Samuel Rota Bulo, Lorenzo Porzi, Manuel López-Antequera, and Peter Kontschieder · 2019
Cited alongside, same era.
Monocular 3d object detection and box fitting trained end-to-end using intersection-over-union loss
Eskil Jörgensen, Christopher Zach, and Fredrik Kahl · 2019
Cited alongside, same era.
FCOS: Fully convolutional one-stage object detection
Zhi Tian, Chunhua Shen, Hao Chen, and Tong He · 2019
Cited alongside, same era.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Jonah Philion and Sanja Fidler · 2020
Later among the works it cites.
PolarMask: Single shot instance segmentation with polar representation
Enze Xie, Peize Sun, Xiaoge Song, Wenhai Wang, Xuebo Liu, Ding Liang, Chunhua Shen, and Ping Luo · 2020
Later among the works it cites.
Learning to evaluate perception models using planner-centric metrics
Jonah Philion, Amlan Kar, and Sanja Fidler · 2020
Later among the works it cites.
The efficacy of neural planning metrics: A meta-analysis of pkl on nuscenes
Yiluan Guo, Holger Caesar, Oscar Beijbom, Jonah Philion, and Sanja Fidler · 2020
Later among the works it cites.
Center-based 3d object detection and tracking
Tianwei Yin, Xingyi Zhou, and Philipp Krahenbuhl · 2021
Later among the works it cites.
FCOS3D: Fully convolutional one-stage monocular 3d object detection
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FreeAnchor: Learning to match anchors for visual object detection
Xiaosong Zhang, Fang Wan, Chang Liu, Rongrong Ji, and Qixiang Ye · 2019
Cited alongside, same era.
Cascade R-CNN: high quality object detection and instance segmentation
Zhaowei Cai and Nuno Vasconcelos · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Cited alongside, same era.
PV-RCNN: Point-voxel feature set abstraction for 3d object detection
Shaoshuai Shi, Chaoxu Guo, Li Jiang, Zhe Wang, Jianping Shi, Xiaogang Wang, and Hongsheng Li · 2020
Cited alongside, same era.
nuScenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom · 2020
Cited alongside, same era.
Scalability in perception for autonomous driving: Waymo open dataset
Pei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard, Vijaysai Patnaik, Paul Tsui, James Guo, Yin Zhou, Yuning Chai, Benjamin Caine, et al · 2020
Cited alongside, same era.
Atlas: End-to-end 3d scene reconstruction from posed images
Zak Murez, Tarrence van As, James Bartolozzi, Ayan Sinha, Vijay Badrinarayanan, and Andrew Rabinovich · 2020
Cited alongside, same era.
Tai Wang, Xinge Zhu, Jiangmiao Pang, and Dahua Lin · 2021
Later among the works it cites.
DETR3D: 3d object detection from multi-view images via 3d-to-2d queries
Yue Wang, Vitor Guizilini, Tianyuan Zhang, Yilun Wang, Hang Zhao, and Justin Solomon · 2021
Later among the works it cites.
Probabilistic and geometric depth: Detecting objects in perspective
Tai Wang, Xinge Zhu, Jiangmiao Pang, and Dahua Lin · 2021
Later among the works it cites.
Deformable detr: Deformable transformers for end-to-end object detection
Xizhou Zhu, Weijie Su, Lewei Lu, Bin Li, Xiaogang Wang, and Jifeng Dai · 2021
Later among the works it cites.
NEAT: Neural attention fields for end-to-end autonomous driving
Kashyap Chitta, Aditya Prakash, and Andreas Geiger · 2021
Later among the works it cites.
Center-based radar and camera fusion for 3d object detection
R Nabati and H CenterFusion Qi · 2021
Later among the works it cites.
Is pseudo-lidar needed for monocular 3d object detection?
Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li, and Adrien Gaidon · 2021
Later among the works it cites.
Efficiently identifying task groupings for multi-task learning
Christopher Fifty, Ehsan Amid, Zhe Zhao, Tianhe Yu, Rohan Anil, and Chelsea Finn · 2021
Later among the works it cites.
ImVoxelNet: Image to voxels projection for monocular and multi-view general-purpose 3d object detection
Danila Rukhovich, Anna Vorontsova, and Anton Konushin · 2022
Closest in time.
Bird’s-eye-view panoptic segmentation using monocular frontal view images
Nikhil Gosala and Abhinav Valada · 2022
Closest in time.