Fetching the paper…
Reading the bibliography…
Training deep neural networks is known to require a large number of training samples.
Solutions of Ill-posed problems
A. N. Tikhonov and V. Y. Arsenin · 1977
Earlier work this paper cites.
Neocognitron: A new algorithm for pattern recognition tolerant of deformations and shifts in position
Kunihiko Fukushima and Sei Miyake · 1982
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel · 1989
Earlier work this paper cites.
Learning from hints in neural networks
Yaser S. Abu-Mostafa · 1990
Earlier work this paper cites.
Document image defect models
H. Baird · 1990
Earlier work this paper cites.
A method for learning from hints
Yaser S. Abu-Mostafa · 1992
Earlier work this paper cites.
Tangent Prop — a formalism for specifying selected invariances in an adaptive network
P. Simard, B. Victorri, Y. Lecun, and J. Denker · 1992
Earlier work this paper cites.
Hints and the vc dimension
Yaser S. Abu-Mostafa · 1993
Earlier work this paper cites.
Efficient Pattern Recognition Using a New Transformation Distance
Patrice Simard, Yann Le Cun, and J. Denker · 1993
Earlier work this paper cites.
Signature verification using a "siamese" time delay neural network
Jane Bromley, Isabelle Guyon, Yann LeCun, Eduard Säckinger, and Roopak Shah · 1994
Earlier work this paper cites.
Learning bayesian networks: The combination of knowledge and statistical data
David Heckerman, Dan Geiger, and David M. Chickering · 1995
Earlier work this paper cites.
Multitask learning
R. Caruana · 1997
Earlier work this paper cites.
Machine Learning
Thomas M. Mitchell · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann Lecun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Incorporating prior information in machine learning by creating virtual examples
P. Niyogi, F. Girosi, and T. Poggio · 1998
Earlier work this paper cites.
Hierarchical models of object recognition in cortex
Maximilian Riesenhuber and Tomaso Poggio · 1999
Earlier work this paper cites.
Learning with Kernels: Support Vector Machines, Regularization, Optimization, and Beyond
Bernhard Scholkopf and Alexander J. Smola · 2001
Earlier work this paper cites.
Best practices for convolutional neural networks applied to visual document analysis
Patrice Y. Simard, Dave Steinkraus, and John C. Platt · 2003
Cited alongside, same era.
Incorporating prior knowledge with weighted margin support vector machines
Xiaoyun Wu and Rohini Srihari · 2004
Cited alongside, same era.
Learning a similarity metric discriminatively, with application to face verification
Sumit Chopra, Raia Hadsell, and Yann LeCun · 2005
Cited alongside, same era.
Greedy layer-wise training of deep networks
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle · 2006
Cited alongside, same era.
Dimensionality reduction by learning an invariant mapping
Raia Hadsell, Sumit Chopra, and Yann LeCun · 2006
Cited alongside, same era.
A fast learning algorithm for deep belief nets
G. E. Hinton, S. Osindero, and Y. W. Teh · 2006
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E. Hinton · 2010
Later among the works it cites.
VQSVM: A case study for incorporating prior domain knowledge into inductive machine learning
Ting Yu, Simeon Simoff, and Tony Jan · 2010
Later among the works it cites.
Higher order contractive auto-encoder
Salah Rifai, Grégoire Mesnil, Pascal Vincent, Xavier Muller, Yoshua Bengio, Yann Dauphin, and Xavier Glorot · 2011
Later among the works it cites.
Contractive auto-encoders: Explicit invariance during feature extraction
Salah Rifai, Pascal Vincent, Xavier Muller, Xavier Glorot, and Yoshua Bengio · 2011
Later among the works it cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Later among the works it cites.
Deep learning via semi-supervised embedding
J. Weston, F. Ratle, H. Mobahi, and R. Collobert · 2012
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Efficient learning of sparse representations with an energy-based model
Marc’Aurelio Ranzato, Christopher S. Poultney, Sumit Chopra, and Yann LeCun · 2006
Cited alongside, same era.
Incorporating prior knowledge on features into learning
Eyal Krupka and Naftali Tishby · 2007
Cited alongside, same era.
Unsupervised learning of invariant feature hierarchies with applications to object recognition
Marc’Aurelio Ranzato, Fu-Jie Huang, Y-Lan Boureau, and Yann LeCun · 2007
Cited alongside, same era.
Incorporating Prior Domain Knowledge into Inductive Machine Learning: Its implementation in contemporary capital markets
Ting Yu, Tony Jan, Simeon Simoff, and John Debenham · 2007
Cited alongside, same era.
A unified architecture for natural language processing: deep neural networks with multitask learning
R. Collobert and J. Weston · 2008
Cited alongside, same era.
A unified architecture for natural language processing: deep neural networks with multitask learning
Ronan Collobert and Jason Weston · 2008
Cited alongside, same era.
Later among the works it cites.
ADADELTA: an adaptive learning rate method
Matthew D. Zeiler · 2012
Later among the works it cites.
Generating sequences with recurrent neural networks
Alex Graves · 2013
Later among the works it cites.
Towards deep neural network architectures robust to adversarial examples
Shixiang Gu and Luca Rigazio · 2014
Later among the works it cites.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Later among the works it cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Later among the works it cites.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott E. Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2014
Later among the works it cites.
Deep multi-task learning with evolving weights
S. Belharbi, R.Hérault, C. Chatelain, and S. Adam · 2016
Later among the works it cites.
Deep learning vector quantization
Harm De Vries, R Memisevic, and A Courville · 2016
Later among the works it cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Later among the works it cites.