Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel, “Backpropagation Applied to Handwritten Zip Code Recognition,” Neural Computation , vol. 1, no. 4, pp. 541–551, 1989
1989
Earlier work this paper cites.
J. Sietsma and R. J. Dow, “Creating artificial neural networks that generalize,” Neural Networks , vol. 4, no. 1, pp. 67–79, 1991
1991
Earlier work this paper cites.
A. Krizhevsky, “Learning Multiple Layers of Features from Tiny Images,” Technical report, University of Toronto , pp. 1–60, 2009
2009
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified Linear Units Improve Restricted Boltzmann Machines,” in Proc. of the 27th International Conference on Machine Learning (ICML2010) , no. 3, 2010, pp. 807–814
2010
Earlier work this paper cites.
D. C. Ciresan, U. Meier, J. Masci, L. M. Gambardella, and U. JSchmidhuber, “Flexible, High Performance Convolutional Neural Networks for Image Classification,” in Proc. of the International Joint Conference on Artificial Intelligence (IJCAI 2011) , 2011, pp. 1237–1242
2011
Earlier work this paper cites.
D. Cireşan, U. Meier, and J. Schmidhuber, “Multi-column Deep Neural Networks for Image Classification,” in Proc. of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR2012) , 2012, pp. 3642–3649
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “ImageNet Classification with Deep Convolutional Neural Networks,” in Advances in Neural Information Processing Systems (NIPS2012) , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
G. E. Hinton, N. Srivastava, A. Krizhevsky, I. Sutskever, and R. R. Salakhutdinov, “Improving neural networks by preventing co-adaptation of feature detectors,” arXiv , pp. 1–18, 2012
2012
Earlier work this paper cites.
M. D. Zeiler and R. Fergus, “Visualizing and Understanding Convolutional Networks,” in Proc. of European Conference on Computer Vision (ECCV2014) , 2014, pp. 818–833
2014
Earlier work this paper cites.
P. Sermanet, D. Eigen, X. Zhang, M. Mathieu, R. Fergus, and Y. LeCun, “OverFeat: Integrated Recognition, Localization and Detection using Convolutional Networks,” in Proc. of International Conference on Learning Representations (ICLR2014) , 2014, pp. 1–16
2014
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, and C. V. Jan, “ImageNet Large Scale Visual Recognition Challenge,” arXiv , pp. 1–43, 2014
2014
Earlier work this paper cites.
T. Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in Proc. of European Conference on Computer Vision (ECCV2014) , 2014, pp. 740–755
2014
Earlier work this paper cites.
G. Hinton, O. Vinyals, and J. Dean, “Distilling the Knowledge in a Neural Network,” in Workshop on Advances in Neural Information Processing Systems (NIPS2014) , 2014, pp. 1–9
2014
Earlier work this paper cites.
R. Kiros, R. Salakhutdinov, and R. S. Zemel, “Unifying Visual-Semantic Embeddings with Multimodal Neural Language Models,” in Workshop on Advances in Neural Information Processing Systems (NIPS2014) , 2014, pp. 1–13
2014
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A Method for Stochastic Optimization,” arXiv , pp. 1–15, 2014
2014
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in Proc. of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR2015) , vol. 07-12-June, 2015, pp. 1–9
2015
Earlier work this paper cites.