D. Navon, “Forest before trees: The precedence of global features in visual perception,” Perception and Psychophysics , vol. 5, pp. 197–200, 1969
1969
Earlier work this paper cites.
S. Lazebnik, C. Schmid, and J. Ponce, “Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 2, 2006, pp. 2169–2178
2006
Earlier work this paper cites.
J. Hegdé, “Time course of visual perception: coarse-to-fine processing and beyond,” Progress in Neurobiology , vol. 84, no. 4, pp. 405–439, 2008
2008
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (voc) challenge,” International Journal of Computer Vision (IJCV) , vol. 88, no. 2, pp. 303–338, 2010
2010
Earlier work this paper cites.
L. Bazzani, H. Larochelle, V. Murino, J.-a. Ting, and N. D. Freitas, “Learning attentional policies for tracking and recognition in video with deep networks,” in Proceedings of the 28th International Conference on Machine Learning (ICML) , 2011, pp. 937–944
2011
Earlier work this paper cites.
Z.-H. Zhou, M.-L. Zhang, S.-J. Huang, and Y.-F. Li, “Multi-instance multi-label learning,” Artificial Intelligence , vol. 176, no. 1, pp. 2291–2320, 2012
2012
Earlier work this paper cites.
M.-L. Zhang and Z.-H. Zhou, “A review on multi-label learning algorithms,” IEEE Transactions on Knowledge and Data Engineering (TKDE) , vol. 26, no. 8, pp. 1819–1837, 2013
2013
Earlier work this paper cites.
A. V. Flevaris, A. Martínez, and S. A. Hillyard, “Attending to global versus local stimulus features modulates neural processing of low versus high spatial frequencies: an analysis with event-related brain potentials,” Frontiers in psychology , vol. 5, 2014
2014
Earlier work this paper cites.
R. Girshick, J. Donahue, T. Darrell, and J. Malik, “Rich feature hierarchies for accurate object detection and semantic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014, pp. 580–587
2014
Earlier work this paper cites.
C. L. Zitnick and P. Dollár, “Edge boxes: Locating object proposals from edges,” in European Conference on Computer Vision (ECCV) , 2014, pp. 391–405
2014
Earlier work this paper cites.
M.-M. Cheng, Z. Zhang, W.-Y. Lin, and P. Torr, “BING: Binarized normed gradients for objectness estimation at 300fps,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014, pp. 3286–3293
2014
Earlier work this paper cites.
V. Mnih, N. Heess, A. Graves et al. , “Recurrent models of visual attention,” in Advances in Neural Information Processing Systems (NIPS) , 2014, pp. 2204–2212
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in Advances in Neural Information Processing Systems (NIPS) , 2014, pp. 568–576
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in European Conference on Computer Vision (ECCV) , 2014, pp. 740–755
2014
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 1026–1034
2015
Earlier work this paper cites.
J. Shao, K. Kang, C. Change Loy, and X. Wang, “Deeply learned attributes for crowded scene understanding,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 4657–4666
2015
Earlier work this paper cites.
Z. Liu, P. Luo, X. Wang, and X. Tang, “Deep learning face attributes in the wild,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 3730–3738
2015
Earlier work this paper cites.