Fetching the paper…
Reading the bibliography…
An intriguing phenomenon observed during training neural networks is the spectral bias, which states that neural networks are biased towards learning less complex functions.
Frequency principle: Fourier analysis sheds light on deep neural networks
Xu, Z.-Q. J · 1901
Earlier work this paper cites.
Wide neural networks of any depth evolve as linear models under gradient descent
Lee, J · 1902
Earlier work this paper cites.
Linearized two-layers neural networks in high dimension
Ghorbani, B · 1904
Earlier work this paper cites.
A fine-grained spectral perspective on neural networks
Yang, G · 1907
Earlier work this paper cites.
Spectrum dependent learning curves in kernel regression and wide neural networks
Bordelon, B · 2002
Earlier work this paper cites.
Implicit regularization of random feature models
Jacot, A · 2002
Earlier work this paper cites.
Frequency bias in neural networks for input of non-uniform density
Basri, R · 2003
Earlier work this paper cites.
Learning theory estimates via integral operators and their approximations
Smale, S · 2007
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
Collobert, R · 2008
Earlier work this paper cites.
Deep neural tangent kernel and laplace kernel have the same rkhs
Chen, L · 2009
Earlier work this paper cites.
Kernel methods for deep learning
Cho, Y · 2009
Earlier work this paper cites.
Geometry on probability spaces
Smale, S · 2009
Earlier work this paper cites.
On learning with integral operators
Rosasco, L · 2010
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Vershynin, R · 2010
Earlier work this paper cites.
Spherical harmonics in p dimensions
Frye, C · 2012
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition
Hinton, G · 2012
Cited alongside, same era.
Learning polynomials with neural networks
Andoni, A · 2014
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K · 2015
Cited alongside, same era.
Deep residual learning for image recognition
He, K · 2016
Cited alongside, same era.
High-dimensional dynamics of generalization error in neural networks
Advani, M. S · 2017
Cited alongside, same era.
Breaking the curse of dimensionality with convex neural networks
Bach, F · 2017
Cited alongside, same era.
What can ResNet learn efficiently, going beyond kernels?
Allen-Zhu, Z · 2019
Closest in time.
A convergence theory for deep learning via over-parameterization
Allen-Zhu, Z · 2019
Closest in time.
The convergence rate of neural networks for learned functions of different frequencies
Basri, R · 2019
Closest in time.
On the inductive bias of neural tangent kernels
Bietti, A · 2019
Closest in time.
Generalization bounds of stochastic gradient descent for wide and deep neural networks
Cao, Y · 2019
Closest in time.
On lazy training in differentiable programming
Chizat, L · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Implicit regularization in matrix factorization
Gunasekar, S · 2017
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Zhang, C · 2017
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A · 2018
Cited alongside, same era.
Learning overparameterized neural networks via stochastic gradient descent on structured data
Li, Y · 2018
Cited alongside, same era.
Algorithmic regularization in over-parameterized matrix sensing and neural networks with quadratic activations
Li, Y · 2018
Cited alongside, same era.
Stochastic gradient descent on separable data: Exact convergence with a fixed learning rate
Nacson, M. S · 2018
Cited alongside, same era.
Nakkiran, P · 2019
Closest in time.
On the spectral bias of neural networks
Rahaman, N · 2019
Closest in time.
On learning over-parameterized neural networks: A functional approximation prospective
Su, L · 2019
Closest in time.
Gradient descent for one-hidden-layer neural networks: Polynomial convergence and sq lower bounds
Vempala, S · 2019
Closest in time.
Gradient descent optimizes over-parameterized deep ReLU networks
Zou, D · 2019
Closest in time.
An improved analysis of training over-parameterized deep neural networks
Zou, D · 2019
Closest in time.
Polylogarithmic width suffices for gradient descent to achieve arbitrarily small test error with shallow relu networks
Ji, Z · 2020
Closest in time.
Spherical harmonics and approximations on the unit sphere: an introduction
Atkinson, K · 2044
Closest in time.