Fetching the paper…
Reading the bibliography…
We propose a spectral-based approach to analyze how two-layer neural networks separate from linear methods in terms of approximating high-dimensional functions.
Uber die beste annaherung von funktionen einer gegebenen funktionenklasse
Andrei Nikolaevich Kolmogorov · 1936
Earlier work this paper cites.
Positive definite functions on spheres
IJ Schoenberg et al · 1942
Earlier work this paper cites.
Theory of reproducing kernels
Nachman Aronszajn · 1950
Earlier work this paper cites.
Dynamic Programming
Richard E. Bellman · 1957
Earlier work this paper cites.
On linear dimensionality of topological vector spaces
Andrei Nikolaevich Kolmogorov · 1958
Earlier work this paper cites.
Approximation of Functions, Athena Series
GG Lorentz · 1966
Earlier work this paper cites.
On n-dimensional diameters of compacts in a Hilbert space
Rais Sal’manovich Ismagilov · 1968
Earlier work this paper cites.
Neural net approximation
Andrew R Barron · 1992
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
Andrew R. Barron · 1993
Earlier work this paper cites.
Hinging hyperplanes for regression, classification, and function approximation
Leo Breiman · 1993
Earlier work this paper cites.
Constructive approximation , volume 303
Ronald A DeVore and George G Lorentz · 1993
Earlier work this paper cites.
Approximation and estimation bounds for artificial neural networks
Andrew R Barron · 1994
Earlier work this paper cites.
Nonlinear approximation by trigonometric sums
R.A. DeVore and V.N. Temlyakov · 1995
Earlier work this paper cites.
Estimates of the number of hidden units and variation with respect to half-spaces
Věra Kurková, Paul C Kainen, and Vladik Kreinovich · 1997
Earlier work this paper cites.
Uniform approximation by neural networks
Yuly Makovoz · 1998
Earlier work this paper cites.
Bounds on rates of variable-basis and neural-network approximation
Vera Kurková and Marcello Sanguineti · 2001
Earlier work this paper cites.
Regularization with dot-product kernels
Alex J Smola, Zoltan L Ovari, and Robert C Williamson · 2001
Earlier work this paper cites.
Comparison of worst case errors in linear and neural network approximation
Vera Kurková and Marcello Sanguineti · 2002
Earlier work this paper cites.
Random features for large-scale kernel machines
Ali Rahimi and Benjamin Recht · 2007
Cited alongside, same era.
Kernel methods for deep learning
Youngmin Cho and Lawrence K Saul · 2009
Cited alongside, same era.
NIST handbook of mathematical functions
Frank WJ Olver, Daniel W Lozier, Ronald F Boisvert, and Charles W Clark · 2010
Cited alongside, same era.
Spherical harmonics and approximations on the unit sphere: An introduction , volume 2044
Kendall Atkinson and Weimin Han · 2012
Cited alongside, same era.
A comparison between fixed-basis and variable-basis schemes for function approximation and functional optimization
Giorgio Gnecco · 2012
Cited alongside, same era.
N-widths in Approximation Theory , volume 7
Allan Pinkus · 2012
Cited alongside, same era.
Theory I: Deep networks and the curse of dimensionality
T Poggio and Q Liao · 2018
Later among the works it cites.
On the inductive bias of neural tangent kernels
Alberto Bietti and Julien Mairal · 2019
Later among the works it cites.
A priori estimates of the population risk for two-layer neural networks
Weinan E, Chao Ma, and Lei Wu · 2019
Later among the works it cites.
Bo Li, Shanshan Tang, and Haijun Yu · 2019
Later among the works it cites.
The generalization error of random features regression: Precise asymptotics and the double descent curve
Song Mei and Andrea Montanari · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jonathan W Siegel and Jinchao Xu · 2012
Cited alongside, same era.
On the computational efficiency of training neural networks
Roi Livni, Shai Shalev-Shwartz, and Ohad Shamir · 2014
Cited alongside, same era.
Gaussian error linear units (GELUs)
Dan Hendrycks and Kevin Gimpel · 2016
Cited alongside, same era.
Risk bounds for high-dimensional ridge function combinations including neural networks
Jason M Klusowski and Andrew R Barron · 2016
Cited alongside, same era.
ℓ 1 \ell_{1} -regularized neural networks are improperly learnable in polynomial time
Yuchen Zhang, Jason D Lee, and Michael I Jordan · 2016
Cited alongside, same era.
SGD learns the conjugate kernel class of the network
Amit Daniely · 2017
Cited alongside, same era.
A function space view of bounded norm infinite width ReLU nets: The multivariate case
Greg Ongie, Rebecca Willett, Daniel Soudry, and Nathan Srebro · 2019
Later among the works it cites.
Adaptivity of deep ReLU network for learning in Besov and mixed smooth Besov spaces: optimal rate and curse of dimensionality
Taiji Suzuki · 2019
Later among the works it cites.
On the power and limitations of random features for understanding neural networks
Gilad Yehudai and Ohad Shamir · 2019
Later among the works it cites.
A dynamical central limit theorem for shallow neural networks
Zhengdao Chen, Grant Rotskoff, Joan Bruna, and Eric Vanden-Eijnden · 2020
Later among the works it cites.
When hardness of approximation meets hardness of learning
Eran Malach and Shai Shalev-Shwartz · 2020
Later among the works it cites.
Deep equals shallow for ReLU networks in kernel regimes
Alberto Bietti and Francis Bach · 2021
Closest in time.
Deep neural tangent kernel and Laplace kernel have the same RKHS
Lin Chen and Sheng Xu · 2021
Closest in time.
The Barron space and the flow-induced function spaces for neural network models
Weinan E, Chao Ma, and Lei Wu · 2021
Closest in time.
Linearized two-layers neural networks in high dimension
Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, and Andrea Montanari · 2021
Closest in time.
Approximation spaces of deep neural networks
Rémi Gribonval, Gitta Kutyniok, Morten Nielsen, and Felix Voigtlaender · 2021
Closest in time.
A spectral analysis of dot-product kernels
Meyer Scetbon and Zaid Harchaoui · 2021
Closest in time.
Sharp lower bounds on the approximation rate of shallow neural networks
Jonathan W Siegel and Jinchao Xu · 2021
Closest in time.