Fetching the paper…
Reading the bibliography…
It remains a puzzle that why deep neural networks (DNNs), with more parameters than samples, often generalize well.
1901
Earlier work this paper cites.
1901
Earlier work this paper cites.
1902
Earlier work this paper cites.
1902
Earlier work this paper cites.
1902
Earlier work this paper cites.
1904
Earlier work this paper cites.
1904
Earlier work this paper cites.
arXiv: 1905.07777. http://arxiv.org/abs/1905.07777
Zhang, Y., Xu, Z.-Q. J., Luo, T. & Ma, Z. (2019), ‘A type of generalization error induced by initialization in deep neural networks’, arXiv:1905.07777 [cs, stat] · 1905
Earlier work this paper cites.
Cybenko, G. (1989), ‘Approximation by superpositions of a sigmoidal function’, Mathematics of control, signals and systems
1989
Earlier work this paper cites.
Hochreiter, S. & Schmidhuber, J. (1995), Simplifying neural nets by discovering flat minima, in
1995
Earlier work this paper cites.
Bartlett, P. L., Maiorov, V. & Meir, R. (1999), Almost linear vc dimension bounds for piecewise polynomial networks, in
1999
Earlier work this paper cites.
Bartlett, P. L. & Mendelson, S. (2002), ‘Rademacher and gaussian complexities: Risk bounds and structural results’, Journal of Machine Learning Research
2002
Earlier work this paper cites.
Bousquet, O. & Elisseeff, A. (2002), ‘Stability and generalization’, Journal of machine learning research
2002
Earlier work this paper cites.
Xu, H. & Mannor, S. (2012), ‘Robustness and generalization’, Machine learning
2012
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Cited alongside, same era.
Shalev-Shwartz, S. & Ben-David, S. (2014), Understanding machine learning: From theory to algorithms
2014
Cited alongside, same era.
2015
Cited alongside, same era.
LeCun, Y., Bengio, Y. & Hinton, G. (2015), ‘Deep learning’, nature
2015
Cited alongside, same era.
2017
Later among the works it cites.
2018
Later among the works it cites.
Jacot, A., Gabriel, F. & Hongler, C. (2018), Neural tangent kernel: Convergence and generalization in neural networks, in
2018
Later among the works it cites.
Mei, S., Montanari, A. & Nguyen, P.-M. (2018), ‘A mean field view of the landscape of two-layer neural networks’, Proceedings of the National Academy of Sciences
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Cited alongside, same era.
Bartlett, P. L., Foster, D. J. & Telgarsky, M. J. (2017), Spectrally-normalized margin bounds for neural networks, in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
Poggio, T., Kawaguchi, K., Liao, Q., Miranda, B., Rosasco, L., Boix, X., Hidary, J. & Mhaskar, H. (2018), Theory of deep learning iii: the non-overfitting puzzle, Technical report, Technical report, CBMM memo 073
2018
Later among the works it cites.
2018
Later among the works it cites.
Rotskoff, G. & Vanden-Eijnden, E. (2018), Parameters as interacting particles: long time convergence and asymptotic error scaling of neural networks, in
2018
Later among the works it cites.
2018
Later among the works it cites.
Soudry, D., Hoffer, E., Nacson, M. S., Gunasekar, S. & Srebro, N. (2018), ‘The implicit bias of gradient descent on separable data’, Journal of Machine Learning Research
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.