Fetching the paper…
Reading the bibliography…
We study the regularisation induced in neural networks by Gaussian noise injections (GNIs).
Solutions of ill-posed problems / Andrey N. Tikhonov and Vasiliy Y. Arsenin ; translation editor, Fritz John
A N (Andrei Nikolaevich) Tikhonov · 1977
Earlier work this paper cites.
The Well-Calibrated Bayesian
A P Dawid · 1982
Earlier work this paper cites.
The comparison and evaluation of forecasters
Morris H. DeGroot and Stephen E. Fienberg · 1983
Earlier work this paper cites.
Biological Cybernetics Networks and the Best Approximation Property
F Girosi and T Poggio · 1990
Earlier work this paper cites.
Approximation capabilities of multilayer feedforward networks
Kurt Hornik · 1991
Earlier work this paper cites.
Lepton spectra as a measure of b quark polarization at LEP
Barbara Mele and Guido Altarelli · 1993
Earlier work this paper cites.
Functional Approximation by FeedForward Networks: A Least-Squares Approach to Generalization
Andrew R. Webb · 1994
Earlier work this paper cites.
Training with Noise is Equivalent to Tikhonov Regularization
Chris M. Bishop · 1995
Earlier work this paper cites.
A multivariate faa di bruno formula with applications
G. M. Constantine and T. H. Savits · 1996
Earlier work this paper cites.
Efficient BackProp , pages 9–48
Yann A. LeCun, Léon Bottou, Genevieve B. Orr, and Klaus-Robert Müller · 1998
Earlier work this paper cites.
On the mathematical foundations of learning
Felipe Cucker and Steve Smale · 2002
Earlier work this paper cites.
Analysis of Tikhonov regularization for function approximation by neural networks
Martin Burger and Andreas Neubauer · 2003
Earlier work this paper cites.
Predicting good probabilities with supervised learning
Alexandru Niculescu-Mizil and Rich Caruana · 2005
Earlier work this paper cites.
Dropout training as adaptive regularization
Stefan Wager, Sida Wang, and Percy S Liang · 2013
Earlier work this paper cites.
Analyzing noise in autoencoders and deep networks
Ben Poole, Jascha Sohl-Dickstein, and Surya Ganguli · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
On the inductive bias of dropout
David P. Helmbold and Philip M. Long · 2015
Cited alongside, same era.
Variational dropout and the local reparameterization trick
Diederik P. Kingma, Tim Salimans, and Max Welling · 2015
Cited alongside, same era.
Obtaining well calibrated probabilities using Bayesian Binning
Mahdi Pakdaman Naeini, Gregory F. Cooper, and Milos Hauskrecht · 2015
Cited alongside, same era.
Norm-based capacity control in neural networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2015
Cited alongside, same era.
Exponential expressivity in deep neural networks through transient chaos
Ben Poole, Subhaneil Lahiri, Maithra Raghu, Jascha Sohl-Dickstein, and Surya Ganguli · 2016
Cited alongside, same era.
Practical Gauss-Newton optimisation for deep learning
Aleksandar Botev, Hippolyt Ritter, and David Barber · 2017
Cited alongside, same era.
Improving DNN robustness to adversarial attacks using jacobian regularization
Daniel Jakubovitz and Raja Giryes · 2018
Later among the works it cites.
Second-order adversarial attack and certifiable robustness
Bai Li, Changyou Chen, Wenlin Wang, and Lawrence Carin · 2018
Later among the works it cites.
Empirical analysis of the hessian of over-parametrized neural networks
Levent Sagun, Utku Evci, V. Ugur Güney, Yann Dauphin, and Léon Bottou · 2018
Later among the works it cites.
Tangent Space Separability in Feedforward Neural Networks
Rita Aleksziev · 2019
Later among the works it cites.
On exact computation with an infinitely wide neural net
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, Russ R Salakhutdinov, and Ruosong Wang · 2019
Later among the works it cites.
On Lazy Training in Differentiable Programming
Lenaic Chizat, Edouard Oyallon, and Francis Bach · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sobolev training for neural networks
Wojciech Marian Czarnecki, Simon Osindero, Max Jaderberg, Grzegorz Swirszcz, and Razvan Pascanu · 2017
Cited alongside, same era.
Sharp minima can generalize for deep nets
Laurent Dinh, Razvan Pascanu, Samy Bengio, and Yoshua Bengio · 2017
Cited alongside, same era.
On calibration of modern neural networks
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q. Weinberger · 2017
Cited alongside, same era.
Principles of Riemannian geometry in neural networks
Michael Hauser and Asok Ray · 2017
Cited alongside, same era.
Three Factors Influencing Minima in SGD
Stanisław Jastrzȩbski, Zachary Kenton, Devansh Arpit, Nicolas Ballas, Asja Fischer, Yoshua Bengio, and Amos Storkey · 2017
Cited alongside, same era.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nathan Srebro · 2017
Cited alongside, same era.
Later among the works it cites.
Certified adversarial robustness via randomized smoothing
Jeremy Cohen, Elan Rosenfeld, and J. Zico Kolter · 2019
Later among the works it cites.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Jorge Nocedal, Ping Tak Peter Tang, Dheevatsa Mudigere, and Mikhail Smelyanskiy · 2019
Later among the works it cites.
Loss landscapes of regularized linear autoencoders
Daniel Kunin, Jonathan M. Bloom, Aleksandrina Goeva, and Cotton Seed · 2019
Later among the works it cites.
Variational bayesian dropout with a hierarchical prior
Yuhang Liu, Wenyong Dong, Lei Zhang, Dong Gong, and Qinfeng Shi · 2019
Later among the works it cites.
On the spectral bias of neural networks
Nasim Rahaman, Aristide Baratin, Devansh Arpit, Felix Draxler, Min Lin, Fred Hamprecht, Yoshua Bengio, and Aaron Courville · 2019
Later among the works it cites.
Dropout: Explicit Forms and Capacity Control
Raman Arora, Peter Bartlett, Poorya Mianjy, and Nathan Srebro · 2020
Closest in time.
A generalized neural tangent kernel analysis for two-layer neural networks, 2020
Zixiang Chen, Yuan Cao, Quanquan Gu, and Tong Zhang · 2020
Closest in time.
Try Depth Instead of Weight Correlations: Mean-field is a Less Restrictive Assumption for Deeper Networks
Sebastian Farquhar, Lewis Smith, and Yarin Gal · 2020
Closest in time.
The Implicit and Explicit Regularization Effects of Dropout
Colin Wei, Sham Kakade, and Tengyu Ma · 2020
Closest in time.