Fetching the paper…
Reading the bibliography…
Deep learning achieves remarkable generalization capability with overwhelming number of model parameters.
Information theory and statistical mechanics
Edwin T Jaynes · 1957
Earlier work this paper cites.
Neural nets with superlinear vc-dimension
Wolfgang Maass · 1994
Earlier work this paper cites.
A maximum entropy approach to natural language processing
Adam L Berger, Vincent J Della Pietra, and Stephen A Della Pietra · 1996
Earlier work this paper cites.
Statistical learning theory , volume 1
Vladimir Naumovich Vapnik and Vlamimir Vapnik · 1998
Earlier work this paper cites.
The information bottleneck method
Naftali Tishby, Fernando C. Pereira, and William Bialek · 1999
Earlier work this paper cites.
A comparison of algorithms for maximum entropy parameter estimation
Robert Malouf · 2002
Earlier work this paper cites.
Maximum entropy estimation for feature forests
Miyao Yusuke and Tsujii Jun’ichi · 2002
Cited alongside, same era.
Optimization, maxent models, and conditional estimation without magic
Christopher Manning and Dan Klein · 2003
Cited alongside, same era.
A tutorial on mm algorithms
David R Hunter and Kenneth Lange · 2004
Cited alongside, same era.
Using maximum entropy for automatic image annotation
Jiwoon Jeon and R Manmatha · 2004
Cited alongside, same era.
Norm-based capacity control in neural networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Kilian Q Weinberger, and Laurens van der Maaten · 2016
Later among the works it cites.
On the emergence of invariance and disentangling in deep representations
Alessandro Achille and Stefano Soatto · 2017
Closest in time.
Opening the black box of deep neural networks via information
Ravid Shwartz-Ziv and Naftali Tishby · 2017
Closest in time.
Towards understanding generalization of deep learning: Perspective of loss landscapes
Lei Wu, Zhanxing Zhu, et al · 2017
Closest in time.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Closest in time.