Fetching the paper…
Reading the bibliography…
Transformer-based methods have swept the benchmarks on 2D and 3D detection on images.
Deep residual learning for image recognition
He Kaiming, Zhang Xiangyu, Ren Shaoqing, and Sun Jian · 2016
Earlier work this paper cites.
Sgdr: Stochastic gradient descent with warm restarts
Ilya Loshchilov and Frank Hutter · 2016
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
In defense of classical image processing: Fast depth completion on the cpu
Jason Ku, Ali Harakeh, and Steven L Waslander · 2018
Earlier work this paper cites.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Earlier work this paper cites.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Philion Jonah and Fidler Sanja · 2020
Earlier work this paper cites.
Generalized focal loss: Learning qualified and distributed bounding boxes for dense object detection
Xiang Li, Wenhai Wang, Lijun Wu, Shuo Chen, Xiaolin Hu, Jun Li, Jinhui Tang, and Jian Yang · 2020
Earlier work this paper cites.
Deformable detr: Deformable transformers for end-to-end object detection
Xizhou Zhu, Weijie Su, Lewei Lu, Bin Li, Xiaogang Wang, and Jifeng Dai · 2020
Earlier work this paper cites.
Fast convergence of detr with spatially modulated co-attention
Peng Gao, Minghang Zheng, Xiaogang Wang, Jifeng Dai, and Hongsheng Li · 2021
Earlier work this paper cites.
Bevdet: High-performance multi-camera 3d object detection in bird-eye-view
Huang Junjie, Huang Guan, Zhu Zheng, and Du Dalong · 2021
Earlier work this paper cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Earlier work this paper cites.
Conditional detr for fast training convergence
D. Meng, X. Chen, Z. Fan, G. Zeng, H. Li, Y. Yuan, L. Sun, and J. Wang · 2021
Earlier work this paper cites.
Is pseudo-lidar needed for monocular 3d object detection?
Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li, and Adrien Gaidon · 2021
Cited alongside, same era.
Rethinking transformer-based set prediction for object detection
Zhiqing Sun, Shengcao Cao, Yiming Yang, and Kris M Kitani · 2021
Cited alongside, same era.
FCOS3D: Fully convolutional one-stage monocular 3d object detection
Wang Tai, Zhu Xinge, Pang Jiangmiao, and Lin Dahua · 2021
Cited alongside, same era.
Probabilistic and Geometric Depth: Detecting objects in perspective
Wang Tai, Zhu Xinge, Pang Jiangmiao, and Lin Dahua · 2021
Cited alongside, same era.
Anchor detr: Query design for transformer-based object detection
Y. Wang, X. Zhang, T. Yang, and J. Sun · 2021
Cited alongside, same era.
Efficient detr: Improving end-to-end object detector with dense prior
Learning ego 3d representation as ray tracing
Jiachen Lu, Zheyuan Zhou, Xiatian Zhu, Hang Xu, and Li Zhang · 2022
Closest in time.
Focal-petr: Embracing foreground for efficient multi-camera 3d object detection
Shihao Wang, Xiaohui Jiang, and Ying Li · 2022
Closest in time.
Petrv2: A unified framework for 3d perception from multi-camera images
Liu Yingfei, Yan Junjie, Jia Fan, Li Shuailin, Gao Qi, Wang Tiancai, Zhang Xiangyu, and Sun Jian · 2022
Closest in time.
Petr: Position embedding transformation for multi-view 3d object detection
Liu Yingfei, Wang Tiancai, Zhang Xiangyu, and Sun Jian · 2022
Closest in time.
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Li Yinhao, Ge Zheng, Yu Guanyi, Yang Jinrong, Wang Zengran, Shi Yukang, Sun Jianjian, and Li Zeming · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhuyu Yao, Jiangbo Ai, Boxun Li, and Chi Zhang · 2021
Cited alongside, same era.
Futr3d: A unified sensor fusion framework for 3d detection
Xuanyao Chen, Tianyuan Zhang, Yue Wang, Yilun Wang, and Hang Zhao · 2022
Cited alongside, same era.
Spatialdetr: Robust scalable transformer-based 3d object detection from multi-view camera images with global cross-sensor attention
Simon Doll, Richard Schulz, Lukas Schneider, Viviane Benzin, Markus Enzweiler, and Hendrik PA Lensch · 2022
Cited alongside, same era.
Bevdet4d: Exploit temporal cues in multi-camera 3d object detection
Huang Junjie and Huang Guan · 2022
Cited alongside, same era.
Dn-detr: Accelerate detr training by introducing query denoising
F. Li, H. Zhang, S. Liu, J. Guo, L. M. Ni, and L. Zhang · 2022
Cited alongside, same era.
Dab-detr: Dynamic anchor boxes are better queries for detr
S. Liu, F. Li, H. Zhang, X. Yang, X. Qi, H. Su, J. Zhu, and L. Zhang · 2022
Cited alongside, same era.
Closest in time.
Detr3d: 3d object detection from multi-view images via 3d-to-2d queries
Wang Yue, Guizilini Vitor Campagnolo, Zhang Tianyuan, Wang Yilun, Zhao Hang, and Solomon Justin · 2022
Closest in time.
Sts: Surround-view temporal stereo for multi-view 3d detection
Wang Zengran, Min Chen, Ge Zheng, Li Yinhao, Li Zeming, Yang Hongyu, and Huang Di · 2022
Closest in time.
Accelerating detr convergence via semantic-aligned matching
G. Zhang, Z. Luo, Y. Yu, K. Cui, and S. Lu · 2022
Closest in time.
Dino: Detr with improved denoising anchor boxes for end-to-end object detection
H. Zhang, F. Li, S. Liu, L. Zhang, H. Su, J. Zhu, L. M. Ni, and H. Y. Shum · 2022
Closest in time.
Li Zhiqi, Wang Wenhai, Li Hongyang, Xie Enze, Sima Chonghao, Lu Tong, Yu Qiao, and Dai Jifeng · 2022
Closest in time.
Object as query: Equipping any 2d object detector with 3d detection ability
Zitian Wang, Zehao Huang, Jiahui Fu, Naiyan Wang, and Si Liu · 2023
Closest in time.