Probability in Banach spaces: Isoperimetry and processes
M. Ledoux and M. Talagrand · 1991
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
R. Tibshirani · 1996
Earlier work this paper cites.
Weak convergence and empirical processes
A. van der Vaart and J. Wellner · 1996
Earlier work this paper cites.
The sample complexity of pattern classification with neural networks: The size of the weights is more important than the size of the network
P. Bartlett · 1998
Earlier work this paper cites.
Empirical processes in M-estimation
S. van de Geer · 2000
Earlier work this paper cites.
Rademacher and Gaussian complexities: risk bounds and structural results
P. Bartlett and S. Mendelson · 2002
Earlier work this paper cites.
Stable signal recovery from incomplete and inaccurate measurements
E. Candès, J. Romberg, and T. Tao · 2006
Earlier work this paper cites.
Compressed sensing
D. Donoho · 2006
Earlier work this paper cites.
Neural network learning: Theoretical foundations
M. Anthony and P. Bartlett · 2009
Earlier work this paper cites.
Bounds for Rademacher processes via chaining
Original
J. Lederer · 2010
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. Hinton · 2010
Earlier work this paper cites.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. Dahl, A. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, and B. Kingsbury · 2012
Earlier work this paper cites.
Concentration inequalities: A nonasymptotic theory of independence
S. Boucheron, G. Lugosi, and P. Massart · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
A. Graves, A. Mohamed, and G. Hinton · 2013
Earlier work this paper cites.