Fetching the paper…
Reading the bibliography…
We present a generalization bound for feedforward neural networks in terms of the product of the spectral norm of the layers and the Frobenius norm of the weights.
Training a 3-node neural network is np-complete
Avrim L Blum and Ronald L Rivest · 1993
Earlier work this paper cites.
The sample complexity of pattern classification with neural networks: the size of the weights is more important than the size of the network
Peter L Bartlett · 1998
Earlier work this paper cites.
Some PAC-Bayesian theorems
David A McAllester · 1998
Earlier work this paper cites.
Almost linear vc dimension bounds for piecewise polynomial networks
Peter L Bartlett, Vitaly Maiorov, and Ron Meir · 1999
Earlier work this paper cites.
PAC-Bayesian model averaging
David A McAllester · 1999
Earlier work this paper cites.
(not) bounding the true error
John Langford and Rich Caruana · 2001
Cited alongside, same era.
Rademacher and gaussian complexities: Risk bounds and structural results
Peter L Bartlett and Shahar Mendelson · 2002
Cited alongside, same era.
Pac-bayes & margins
John Langford and John Shawe-Taylor · 2003
Cited alongside, same era.
Simplified pac-bayesian margin bounds
David McAllester · 2003
Cited alongside, same era.
User-friendly tail bounds for sums of random matrices
Joel A Tropp · 2012
Cited alongside, same era.
Spectrally-normalized margin bounds for neural networks
Peter Bartlett, Dylan J Foster, and Matus Telgarsky
Cited in the paper.
Spectrally-normalized margin bounds for neural networks
Peter Bartlett, Dylan J Foster, and Matus Telgarsky
Cited in the paper.
Path-SGD: Path-normalized optimization in deep neural networks
Behnam Neyshabur, Ruslan Salakhutdinov, and Nathan Srebro
Cited in the paper.
Norm-based capacity control in neural networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro
Cited in the paper.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang · 2016
Later among the works it cites.
Gintare Karolina Dziugaite and Daniel M Roy · 2017
Closest in time.
Nearly-tight vc-dimension bounds for piecewise linear neural networks
Nick Harvey, Chris Liaw, and Abbas Mehrabian · 2017
Closest in time.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nathan Srebro · 2017
Closest in time.
Understanding deep learning requires rethinking generalization
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Closest in time.