Fetching the paper…
Reading the bibliography…
Each year, deep learning demonstrates new and improved empirical results with deeper and wider neural networks.
Robust locally weighted regression and smoothing scatterplots
William S Cleveland · 1979
Earlier work this paper cites.
Decision theoretic generalizations of the pac model for neural net and other learning applications
David Haussler · 1992
Earlier work this paper cites.
Bounds on the sample complexity of Bayesian learning using information theory and the VC dimension
David Haussler, Michael Kearns, and Robert E Schapire · 1994
Earlier work this paper cites.
Almost linear vc-dimension bounds for piecewise polynomial networks
Peter L Bartlett, Vitaly Maiorov, and Ron Meir · 1998
Earlier work this paper cites.
The information bottleneck method
Naftali Tishby, Fernando C Pereira, and William Bialek · 2000
Earlier work this paper cites.
Elements of Information Theory (Wiley Series in Telecommunications and Signal Processing)
Thomas M. Cover and Joy A. Thomas · 2006
Earlier work this paper cites.
Algorithms and SQ lower bounds for PAC learning one-hidden-layer ReLU networks
Ilias Diakonikolas, Daniel M. Kane, Vasilis Kontonis, and Nikos Zarifis · 2006
Earlier work this paper cites.
Superpolynomial lower bounds for learning one-layer neural networks using gradient descent, 2020
Surbhi Goel, Aravind Gollakota, Zhihan Jin, Sushrut Karmalkar, and Adam Klivans · 2006
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2010
Earlier work this paper cites.
On the computational efficiency of training neural networks
Roi Livni, Shai Shalev-Shwartz, and Ohad Shamir · 2014
Earlier work this paper cites.
In search of the real inductive bias: On the role of implicit regularization in deep learning
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2014
Earlier work this paper cites.
Generalization bounds for neural networks through tensor factorization
Majid Janzamin, Hanie Sedghi, and Anima Anandkumar · 2015
Earlier work this paper cites.
Patterns for learning with side information
Rico Jonschkowski, Sebastian Höfer, and Oliver Brock · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Norm-based capacity control in neural networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2015
Cited alongside, same era.
Rate-distortion bounds on bayes risk in supervised learning
Matthew Nokleby, Ahmad Beirami, and Robert Calderbank · 2016
Cited alongside, same era.
Spectrally-normalized margin bounds for neural networks
Peter Bartlett, Dylan Foster, and Matus Telgarsky · 2017
Cited alongside, same era.
Information-theoretic analysis of generalization capability of learning algorithms
Aolin Xu and Maxim Raginsky · 2017
Later among the works it cites.
Recovery guarantees for one-hidden-layer neural networks
Kai Zhong, Zhao Song, Prateek Jain, Peter L Bartlett, and Inderjit S Dhillon · 2017
Later among the works it cites.
Stronger generalization bounds for deep nets via a compression approach
Sanjeev Arora, Rong Ge, Behnam Neyshabur, and Yi Zhang · 2018
Later among the works it cites.
Size-independent sample complexity of neural networks
Noah Golowich, Alexander Rakhlin, and Ohad Shamir · 2018
Later among the works it cites.
Deterministic PAC-Bayesian generalization bounds for deep networks via generalizing noise-resilience
Vaishnavh Nagarajan and Zico Kolter · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gintare Karolina Dziugaite and Daniel M Roy · 2017
Cited alongside, same era.
Learning one-hidden-layer neural networks with landscape design
Rong Ge, Jason D Lee, and Tengyu Ma · 2017
Cited alongside, same era.
Nearly-tight VC-dimension bounds for piecewise linear neural networks
Nick Harvey, Christopher Liaw, and Abbas Mehrabian · 2017
Cited alongside, same era.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang · 2017
Cited alongside, same era.
Opening the black box of deep neural networks via information
Ravid Shwartz-Ziv and Naftali Tishby · 2017
Cited alongside, same era.
Cyclical learning rates for training neural networks
Leslie N. Smith · 2017
Cited alongside, same era.
A PAC-Bayesian approach to spectrally-normalized margin bounds for neural networks
Behnam Neyshabur, Srinadh Bhojanapalli, and Nathan Srebro
Cited in the paper.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
How much does your data exploration overfit? controlling bias via information usage
Daniel Russo and James Zou · 2019
Later among the works it cites.
Data-dependent sample complexity of deep neural networks via lipschitz augmentation
Colin Wei and Tengyu Ma · 2019
Later among the works it cites.
Guaranteed recovery of one-hidden-layer neural networks via cross entropy
Haoyu Fu, Yuejie Chi, and Yingbin Liang · 2020
Later among the works it cites.
Reinforcement learning, bit by bit
Xiuyuan Lu, Benjamin Van Roy, Vikranth Dwaracherla, Morteza Ibrahimi, Ian Osband, and Zheng Wen · 2021
Later among the works it cites.
Ian Osband, Zheng Wen, Mohammad Asghari, Morteza Ibrahimi, Xiyuan Lu, and Benjamin Van Roy · 2021
Later among the works it cites.