Fetching the paper…
Reading the bibliography…
In this paper we show that by carefully making good choices for various detailed but important factors in a visual recognition framework using deep learning features, one can achieve a simple, efficient, yet highly accurate image classification system.
Video Google: A text retrieval approach to object matching in videos
J. Sivic and A. Zisserman · 2003
Earlier work this paper cites.
Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories
S. Lazebnik, C. Schmid, and J. Ponce · 2006
Earlier work this paper cites.
The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results, 2007
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2007
Earlier work this paper cites.
Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories
L. Fei-Fei, R. Fergus, and P. Perona · 2007
Earlier work this paper cites.
Caltech-256 object category dataset, 2007
G. Griffin, A. Holub, and P. Perona · 2007
Earlier work this paper cites.
LIBLINEAR: A library for large linear classification
R.-E. Fan, K.-W. Chang, C.-J. Hsieh, X.-R. Wang, and C.-J. Lin · 2008
Earlier work this paper cites.
VLFeat: An open and portable library of computer vision algorithms, 2008
A. Vedaldi and B. Fulkerson · 2008
Earlier work this paper cites.
Recognizing indoor scenes
A. Quattoni and A. Torralba · 2009
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
H. Jégou, M. Douze, C. Schmid, and P. Pérez · 2010
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Earlier work this paper cites.
SUN database: Large-scale scene recognition from abbey to zoo
J. Xiao, J. Hays, K. A. Ehinger, A. Oliva, and A. Torralba · 2010
Cited alongside, same era.
CENTRIST: A visual descriptor for scene categorization
J. Wu and J. M. Rehg · 2011
Cited alongside, same era.
Human action recognition by learning bases of action attributes and parts
B. Yao, X. Jiang, A. Khosla, A. L. Lin, L. Guibas, and L. Fei-Fei · 2011
Cited alongside, same era.
Adaptive deconvolutional networks for mid and high level feature learning
M. D. Zeiler, G. W. Taylor, and R. Fergus · 2011
Cited alongside, same era.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Image classification with the fisher vector: Theory and practice
J. Sánchez, F. Perronnin, T. Mensink, and J. Verbeek · 2013
MatConvNet: Convolutional neural networks for MATLAB, 2014
A. Vedaldi and K. Lenc · 2014
Later among the works it cites.
Towards good practices for action video encoding
J. Wu, Y. Zhang, and W. Lin · 2014
Later among the works it cites.
Fisher kernel for deep neural activations
D. Yoo, S. Park, J.-Y. Lee, and I. S. Kweon · 2014
Later among the works it cites.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Later among the works it cites.
Learning deep features for scene recognition using places database
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva · 2014
Later among the works it cites.
Deep convolutional filter banks for texture recognition and segmentation
M. Cimpoi, S. Maji, and A. Vedaldi · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Cited alongside, same era.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Cited alongside, same era.
Multi-scale orderless pooling of deep convolutional activation features
Y. Gong, L. Wang, R. Guo, and S. Lazebnik · 2014
Cited alongside, same era.
Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2014
Cited alongside, same era.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich
Cited in the paper.
Closest in time.
The treasure beneath convolutional layers: Cross-convolutional-layer pooling for image classification
L. Liu, C. Shen, and A. van den Hengel · 2015
Closest in time.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Closest in time.
A discriminative CNN video representation for event detection
Z. Xu, Y. Yang, and A. G. Hauptmann · 2015
Closest in time.