Fetching the paper…
Reading the bibliography…
We investigate the effect of explicitly enforcing the Lipschitz continuity of neural networks with respect to their inputs.
The sample complexity of pattern classification with neural networks: The size of the weights is more important than the size of the network
P.L. Bartlett · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Real mathematical analysis
Charles Chapman Pugh · 2002
Earlier work this paper cites.
Closed-form Dual Perturb and Combine for Tree-based Models
Pierre Geurts and Louis Wehenkel · 2005
Earlier work this paper cites.
Statistical Comparisons of Classifiers over Multiple Data Sets
Janez Demšar · 2006
Earlier work this paper cites.
Learning Multiple Layers of Features from Tiny Images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Robustness and generalization
Huan Xu and Shie Mannor · 2012
Earlier work this paper cites.
Regularization of neural networks using dropconnect
Li Wan, Matthew Zeiler, Sixin Zhang, Yann Le Cun, and Rob Fergus · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Understanding machine learning: From theory to algorithms
Shai Shalev-Shwartz and Shai Ben-David · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Cited alongside, same era.
Hyperopt: A Python library for model selection and hyperparameter optimization
James Bergstra, Brent Komer, Chris Eliasmith, Dan Yamins, and David D. Cox · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Variational dropout and the local reparameterization trick
Diederik P Kingma, Tim Salimans, and Max Welling · 2015
Cited alongside, same era.
Train faster, generalize better: Stability of stochastic gradient descent
Moritz Hardt, Ben Recht, and Yoram Singer · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Concrete dropout
Yarin Gal, Jiri Hron, and Alex Kendall · 2017
Later among the works it cites.
On the Properties of the Softmax Function with Application in Game Theory and Reinforcement Learning
Bolin Gao and Lacra Pavel · 2017
Later among the works it cites.
Implicit Regularization in Deep Learning
Behnam Neyshabur · 2017
Later among the works it cites.
Fashion-MNIST: A Novel Image Dataset for Benchmarking Machine Learning Algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Later among the works it cites.
Spectral Norm Regularization for Improving the Generalizability of Deep Learning
Yuichi Yoshida and Takeru Miyato · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Elementary Linear Algebra
Ron Larson · 2016
Cited alongside, same era.
Sergey Zagoruyko and Nikos Komodakis · 2016
Cited alongside, same era.
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Cited alongside, same era.
Lipschitz Properties for Deep Convolutional Networks
Radu Balan, Maneesh Singh, and Dongmian Zou · 2017
Cited alongside, same era.
Spectrally-normalized margin bounds for neural networks
Peter L Bartlett, Dylan J Foster, and Matus J Telgarsky · 2017
Cited alongside, same era.
Parseval Networks: Improving Robustness to Adversarial Examples
Moustapha Cisse, Piotr Bojanowski, Edouard Grave, Yann Dauphin, and Nicolas Usunier · 2017
Cited alongside, same era.
Size-Independent Sample Complexity of Neural Networks
Noah Golowich, Alexander Rakhlin, and Ohad Shamir · 2018
Closest in time.
Spectral Normalization for Generative Adversarial Networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida · 2018
Closest in time.
A PAC-Bayesian Approach to Spectrally-Normalized Margin Bounds for Neural Networks
Behnam Neyshabur, Srinadh Bhojanapalli, and Nathan Srebro · 2018
Closest in time.
The Singular Values of Convolutional Layers
Hanie Sedghi, Vineet Gupta, and Philip M. Long · 2018
Closest in time.
Yusuke Tsuzuku, Issei Sato, and Masashi Sugiyama · 2018
Closest in time.
MaxGain: Regularisation of Neural Networks by Constraining Activation Magnitudes
Henry Gouk, Bernhard Pfahringer, Eibe Frank, and Michael J. Cree · 2019
Closest in time.