Fetching the paper…
Reading the bibliography…
Modern top-performing object detectors depend heavily on backbone networks, whose advances bring consistent performance gains through exploring more effective network structures.
A. Krogh and J. Vedelsby, “Neural network ensembles, cross validation, and active learning,” in NeurIPS , 1994
1994
Earlier work this paper cites.
2003
Earlier work this paper cites.
G. Brown, “Diversity in neural network ensembles,” Ph.D. dissertation, University of Birmingham, UK, 2004
2004
Earlier work this paper cites.
G. Brown, J. L. Wyatt, R. Harris, and X. Yao, “Diversity creation methods: a survey and categorisation,” Inf. Fusion , 2005
2005
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in CVPR , 2009
2009
Earlier work this paper cites.
2010
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in NeurIPS , 2012
2012
Earlier work this paper cites.
C. Zhang and Y. Ma, Ensemble machine learning: Methods and applications , 2012
2012
Earlier work this paper cites.
T. Lin, M. Maire, S. J. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: common objects in context,” in ECCV , 2014
2014
Earlier work this paper cites.
M. Liang and X. Hu, “Recurrent convolutional neural network for object recognition,” in CVPR , 2015
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in ICLR , 2015
2015
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. E. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in CVPR , 2015
2015
Earlier work this paper cites.
R. Girshick, “Fast r-cnn,” in ICCV , 2015
2015
Earlier work this paper cites.
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. E. Reed, C. Fu, and A. C. Berg, “SSD: single shot multibox detector,” in ECCV , 2016
2016
Earlier work this paper cites.
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi, “You only look once: Unified, real-time object detection,” in CVPR , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
L. Shen, Z. Lin, and Q. Huang, “Relay backpropagation for effective learning of deep convolutional neural networks,” in ECCV , 2016
2016
Earlier work this paper cites.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in CVPR , 2016
2016
Earlier work this paper cites.
S. Ren, K. He, R. B. Girshick, and J. Sun, “Faster R-CNN: towards real-time object detection with region proposal networks,” TPAMI , 2017
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. B. Girshick, “Mask R-CNN,” in ICCV , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Xie, R. Girshick, P. Dollár, Z. Tu, and K. He, “Aggregated residual transformations for deep neural networks,” in CVPR , 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017
2017
Earlier work this paper cites.
J. Dai, H. Qi, Y. Xiong, Y. Li, G. Zhang, H. Hu, and Y. Wei, “Deformable convolutional networks,” in ICCV , 2017
2017
Earlier work this paper cites.
T. Lin, P. Dollár, R. B. Girshick, K. He, B. Hariharan, and S. J. Belongie, “Feature pyramid networks for object detection,” in CVPR , 2017
2017
Earlier work this paper cites.
G. Huang, Z. Liu, L. Van Der Maaten, and K. Q. Weinberger, “Densely connected convolutional networks,” in CVPR , 2017
2017
Earlier work this paper cites.
S. Ren, K. He, R. B. Girshick, X. Zhang, and J. Sun, “Object detection networks on convolutional feature maps,” TPAMI , 2017
2017
Earlier work this paper cites.
J. Huang, V. Rathod, C. Sun, M. Zhu, A. Korattikara, A. Fathi, I. Fischer, Z. Wojna, Y. Song, S. Guadarrama, and K. Murphy, “Speed/accuracy trade-offs for modern convolutional object detectors,” in CVPR , 2017
2017
Earlier work this paper cites.
N. Bodla, B. Singh, R. Chellappa, and L. S. Davis, “Soft-NMS- improving object detection with one line of code,” in ICCV , 2017
2017
Cited alongside, same era.
Z. Cai and N. Vasconcelos, “Cascade R-CNN: delving into high quality object detection,” in CVPR , 2018
2018
Cited alongside, same era.
X. Zhang, X. Zhou, M. Lin, and J. Sun, “Shufflenet: An extremely efficient convolutional neural network for mobile devices,” in CVPR , 2018
2018
Cited alongside, same era.
Z. Li, C. Peng, G. Yu, X. Zhang, Y. Deng, and J. Sun, “Detnet: Design backbone for object detection,” in ECCV , 2018
2018
Cited alongside, same era.
S. Sun, J. Pang, J. Shi, S. Yi, and W. Ouyang, “Fishnet: A versatile backbone for image, region, and pixel level prediction,” in NeurIPS , 2018
2018
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in ECCV , 2020
2020
Later among the works it cites.
N. Wang, Y. Gao, H. Chen, P. Wang, Z. Tian, C. Shen, and Y. Zhang, “NAS-FCOS: fast neural architecture search for object detection,” in CVPR , 2020
2020
Later among the works it cites.
X. Du, T. Lin, P. Jin, G. Ghiasi, M. Tan, Y. Cui, Q. V. Le, and X. Song, “Spinenet: Learning scale-permuted backbone for recognition and localization,” in CVPR , 2020
2020
Later among the works it cites.
L. Yao, H. Xu, W. Zhang, X. Liang, and Z. Li, “SM-NAS: structural-to-modular neural architecture search for object detection,” in AAAI , 2020
2020
Later among the works it cites.
A. Kuznetsova, H. Rom, N. Alldrin, J. R. R. Uijlings, I. Krasin, J. Pont-Tuset, S. Kamali, S. Popov, M. Malloci, A. Kolesnikov, T. Duerig, and V. Ferrari, “The open images dataset V4,” IJCV , 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
O. Sagi and L. Rokach, “Ensemble learning: A survey,” WIDM , 2018
2018
Cited alongside, same era.
C. Peng, T. Xiao, Z. Li, Y. Jiang, X. Zhang, K. Jia, G. Yu, and J. Sun, “Megdet: A large mini-batch object detector,” in CVPR , 2018
2018
Cited alongside, same era.
S. Liu, L. Qi, H. Qin, J. Shi, and J. Jia, “Path aggregation network for instance segmentation,” in CVPR , 2018
2018
Cited alongside, same era.
M. Sandler, A. G. Howard, M. Zhu, A. Zhmoginov, and L. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” in CVPR , 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
G. Ghiasi, T. Lin, and Q. V. Le, “NAS-FPN: learning scalable feature pyramid architecture for object detection,” in CVPR , 2019
2019
Cited alongside, same era.
J. Pang, K. Chen, J. Shi, H. Feng, W. Ouyang, and D. Lin, “Libra r-cnn: Towards balanced learning for object detection,” in CVPR , 2019
2019
Cited alongside, same era.
2020
Later among the works it cites.
K. Kim and H. S. Lee, “Probabilistic anchor assignment with iou prediction for object detection,” in ECCV , A. Vedaldi, H. Bischof, T. Brox, and J. Frahm, Eds., 2020
2020
Later among the works it cites.
C. Jiang, H. Xu, W. Zhang, X. Liang, and Z. Li, “SP-NAS: serial-to-parallel backbone search for object detection,” in CVPR , 2020
2020
Later among the works it cites.
Y. Cao, J. Xu, S. Lin, F. Wei, and H. Hu, “Global context networks,” TPAMI , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Wu and K. He, “Group normalization,” IJCV , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
S. Gao, M. Cheng, K. Zhao, X. Zhang, M. Yang, and P. H. S. Torr, “Res2net: A new multi-scale backbone architecture,” TPAMI , 2021
2021
Closest in time.
W. Wang, E. Xie, X. Li, D. Fan, K. Song, D. Liang, T. Lu, P. Luo, and L. Shao, “Pyramid vision transformer: A versatile backbone for dense prediction without convolutions,” in ICCV , 2021
2021
Closest in time.
2021
Closest in time.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows,” in ICCV , 2021
2021
Closest in time.
J. Wang, K. Sun, T. Cheng, B. Jiang, C. Deng, Y. Zhao, D. Liu, Y. Mu, M. Tan, X. Wang, W. Liu, and B. Xiao, “Deep high-resolution representation learning for visual recognition,” TPAMI , 2021
2021
Closest in time.
P. Chen, M. Chang, J. Hsieh, and Y. Chen, “Parallel residual bi-fusion feature pyramid network for accurate single-shot object detection,” TIP , 2021
2021
Closest in time.
T. Liang, Y. Wang, Z. Tang, G. Hu, and H. Ling, “Opanas: One-shot path aggregation network architecture search for object detection,” in CVPR , 2021
2021
Closest in time.
L. Yao, R. Pi, H. Xu, W. Zhang, Z. Li, and T. Zhang, “Joint-detnas: Upgrade your detector with nas, pruning and dynamic distillation,” in CVPR , 2021
2021
Closest in time.
M. Chen, J. Fu, and H. Ling, “One-shot neural ensemble architecture search by diversity-guided search space shrinking,” in CVPR , 2021
2021
Closest in time.
X. Zhu, W. Su, L. Lu, B. Li, X. Wang, and J. Dai, “Deformable detr: Deformable transformers for end-to-end object detection,” in ICLR , 2021
2021
Closest in time.
C.-Y. Wang, A. Bochkovskiy, and H.-Y. M. Liao, “Scaled-yolov4: Scaling cross stage partial network,” in CVPR , 2021
2021
Closest in time.
G. Ghiasi, Y. Cui, A. Srinivas, R. Qian, T. Lin, E. D. Cubuk, Q. V. Le, and B. Zoph, “Simple copy-paste is a strong data augmentation method for instance segmentation,” in CVPR , 2021
2021
Closest in time.
S. Qiao, L. Chen, and A. L. Yuille, “Detectors: Detecting objects with recursive feature pyramid and switchable atrous convolution,” in CVPR , 2021
2021
Closest in time.
X. Dai, Y. Chen, B. Xiao, D. Chen, M. Liu, L. Yuan, and L. Zhang, “Dynamic head: Unifying object detection heads with attentions,” in CVPR , 2021
2021
Closest in time.
2021
Closest in time.
W. Wang, E. Xie, X. Li, D.-P. Fan, K. Song, D. Liang, T. Lu, P. Luo, and L. Shao, “Pvt v2: Improved baselines with pyramid vision transformer,” Computational Visual Media , 2022
2022
Closest in time.