Identity Crisis: Memorization and Generalization under Extreme Overparameterization, 2019
Original
Chiyuan Zhang, Samy Bengio, Moritz Hardt, and Yoram Singer · 1902
Earlier work this paper cites.
Regularization of inverse problems
Heinz Werner Engl, Martin Hanke, and Andreas Neubauer · 1996
Earlier work this paper cites.
A Fast Fixed-Point Algorithm for Independent Component Analysis
Aapo Hyvärinen and Erkki Oja · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Greedy Layer-Wise Training of Deep Networks
Yoshua Bengio, Pascal Lamblin, Dan Popovici, and Hugo Larochelle · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, Alex · 2009
Earlier work this paper cites.
Why Does Unsupervised Pre-training Help Deep Learning?
Dumitru Erhan, Yoshua Bengio, Aaron Courville, Pierre-Antoine Manzagol, Pascal Vincent, and Samy Bengio · 2010
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Scikit-learn: Machine Learning in Python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, Jake Vanderplas, Alexandre Passos, David Cournapeau, Matthieu Brucher, Matthieu Perrot, and Édouard Duchesnay · 2011
Earlier work this paper cites.
In search of the real inductive bias: On the role of implicit regularization in deep learning, 2014
Original
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2014
Earlier work this paper cites.