Fetching the paper…
Reading the bibliography…
Existing deep convolutional neural networks (CNNs) require a fixed-size (e.g., 224x224) input image.
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel, “Backpropagation applied to handwritten zip code recognition,” Neural computation , 1989
1989
Earlier work this paper cites.
J. Sivic and A. Zisserman, “Video google: a text retrieval approach to object matching in videos,” in ICCV , 2003
2003
Earlier work this paper cites.
D. G. Lowe, “Distinctive image features from scale-invariant keypoints,” IJCV , 2004
2004
Earlier work this paper cites.
K. Grauman and T. Darrell, “The pyramid match kernel: Discriminative classification with sets of image features,” in ICCV , 2005
2005
Earlier work this paper cites.
N. Dalal and B. Triggs, “Histograms of oriented gradients for human detection,” in CVPR , 2005
2005
Earlier work this paper cites.
S. Lazebnik, C. Schmid, and J. Ponce, “Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories,” in CVPR , 2006
2006
Earlier work this paper cites.
L. Fei-Fei, R. Fergus, and P. Perona, “Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,” CVIU , 2007
2007
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman, “The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results,” 2007
2007
Earlier work this paper cites.
J. C. van Gemert, J.-M. Geusebroek, C. J. Veenman, and A. W. Smeulders, “Kernel codebooks for scene categorization,” in ECCV , 2008
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in CVPR , 2009
2009
Earlier work this paper cites.
J. Yang, K. Yu, Y. Gong, and T. Huang, “Linear spatial pyramid matching using sparse coding for image classification,” in CVPR , 2009
2009
Earlier work this paper cites.
J. Wang, J. Yang, K. Yu, F. Lv, T. Huang, and Y. Gong, “Locality-constrained linear coding for image classification,” in CVPR , 2010
2010
Earlier work this paper cites.
F. Perronnin, J. Sánchez, and T. Mensink, “Improving the fisher kernel for large-scale image classification,” in ECCV , 2010
2010
Earlier work this paper cites.
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan, “Object detection with discriminatively trained part-based models,” PAMI , 2010
2010
Earlier work this paper cites.
K. E. van de Sande, J. R. Uijlings, T. Gevers, and A. W. Smeulders, “Segmentation as selective search for object recognition,” in ICCV , 2011
2011
Cited alongside, same era.
K. Chatfield, V. Lempitsky, A. Vedaldi, and A. Zisserman, “The devil is in the details: an evaluation of recent feature encoding methods,” in BMVC , 2011
2011
Cited alongside, same era.
A. Coates and A. Ng, “The importance of encoding versus training with sparse coding and vector quantization,” in ICML , 2011
2011
Cited alongside, same era.
C.-C. Chang and C.-J. Lin, “Libsvm: a library for support vector machines,” ACM Transactions on Intelligent Systems and Technology (TIST) , 2011
2011
Cited alongside, same era.
A. Krizhevsky, I. Sutskever, and G. Hinton, “Imagenet classification with deep convolutional neural networks,” in NIPS , 2012
2012
Cited alongside, same era.
2014
Closest in time.
R. Girshick, J. Donahue, T. Darrell, and J. Malik, “Rich feature hierarchies for accurate object detection and semantic segmentation,” in CVPR , 2014
2014
Closest in time.
2014
Closest in time.
A. S. Razavian, H. Azizpour, J. Sullivan, and S. Carlsson, “Cnn features off-the-shelf: An astounding baseline for recogniton,” in CVPR 2014, DeepVision Workshop , 2014
2014
Closest in time.
Y. Taigman, M. Yang, M. Ranzato, and L. Wolf, “Deepface: Closing the gap to human-level performance in face verification,” in CVPR , 2014
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Jegou, F. Perronnin, M. Douze, J. Sanchez, P. Perez, and C. Schmid, “Aggregating local image descriptors into compact codes,” TPAMI , vol. 34, no. 9, pp. 1704–1716, 2012
2012
Cited alongside, same era.
2013
Cited alongside, same era.
2013
Cited alongside, same era.
2013
Cited alongside, same era.
M. Lin, Q. Chen, and S. Yan, “Network in network,” arXiv:1312.4400 , 2013
2013
Cited alongside, same era.
Y. Jia, “Caffe: An open source convolutional architecture for fast feature embedding,” http://caffe.berkeleyvision.org/ , 2013
2013
Cited alongside, same era.
2013
Cited alongside, same era.
2014
Closest in time.
N. Zhang, M. Paluri, M. Ranzato, T. Darrell, and L. Bourdevr, “Panda: Pose aligned networks for deep attribute modeling,” in CVPR , 2014
2014
Closest in time.
2014
Closest in time.
C. L. Zitnick and P. Dollár, “Edge boxes: Locating object proposals from edges,” in ECCV , 2014
2014
Closest in time.
2014
Closest in time.
2014
Closest in time.
2014
Closest in time.
M. Oquab, L. Bottou, I. Laptev, J. Sivic et al. , “Learning and transferring mid-level image representations using convolutional neural networks,” in CVPR , 2014
2014
Closest in time.
C. Szegedy, A. Toshev, and D. Erhan, “Deep neural networks for object detection,” in NIPS , 2013
2014
Closest in time.