Fetching the paper…
Reading the bibliography…
Training convolutional networks (CNN's) that fit on a single GPU with minibatch stochastic gradient descent has become effective in practice.
Adaptive mixtures of local experts
R. A. Jacobs, M. I. Jordan, S. J. Nowlan, and G. E. Hinton · 1991
Earlier work this paper cites.
Scaling large learning problems with hard parallel mixtures
R. Collobert, Y. Bengio, and S. Bengio · 2003
Earlier work this paper cites.
A visual vocabulary for flower classification
M. Nilsback and A. Zisserman · 2006
Earlier work this paper cites.
Recognizing indoor scenes
A. Quattoni and A. Torralba · 2009
Earlier work this paper cites.
Caltech-ucsd birds 200
P. Welinder, S. Branson, T. Mita, C. Wah, F. Schroff, S. Belongie, and P. Perona · 2010
Earlier work this paper cites.
Human action recognition by learning bases of action attributes and parts
B. Yao, X. Jiang, A. Khosla, A. L. Lin, L. J. Guibas, and F. Li · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Learning everything about anything: Webly-supervised visual concept learning
S. K. Divvala, A. Farhadi, and C. Guestrin · 2014
Cited alongside, same era.
CNN features off-the-shelf: an astounding baseline for recognition
A. S. Razavian, H. Azizpour, J. Sullivan, and S. Carlsson · 2014
Cited alongside, same era.
Self-informed neural network structure learning
D. Warde-Farley, A. Rabinovich, and D. Anguelov · 2014
Cited alongside, same era.
Webly supervised learning of convolutional networks
X. Chen and A. Gupta · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
G. E. Hinton, O. Vinyals, and J. Dean · 2015
Cited alongside, same era.
Network of experts for large-scale image categorization
K. Ahmed, M. H. Baig, and L. Torresani · 2016
Later among the works it cites.
Identity mappings in deep residual networks
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Learning visual features from large weakly supervised data
A. Joulin, L. van der Maaten, A. Jabri, and N. Vasilache · 2016
Later among the works it cites.
YFCC100M: the new data in multimedia research
B. Thomee, D. A. Shamma, G. Friedland, B. Elizalde, K. Ni, D. Poland, D. Borth, and L. Li · 2016
Later among the works it cites.
Sun database: Exploring a large collection of scene categories
J. Xiao, K. A. Ehinger, J. Hays, A. Torralba, and A. Oliva · 2016
Later among the works it cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer, 2017
N. Shazeer, A. Mirhoseini, K. Maziarz, A. Davis, Q. Le, G. Hinton, and J. Dean · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. S. Bernstein, A. C. Berg, and F. Li · 2015
Cited alongside, same era.
HD-CNN: hierarchical deep convolutional neural networks for large scale visual recognition
Z. Yan, H. Zhang, R. Piramuthu, V. Jagadeesh, D. DeCoste, W. Di, and Y. Yu · 2015
Cited alongside, same era.
Closest in time.