Fetching the paper…
Reading the bibliography…
Understanding how convolutional neural networks (CNNs) can efficiently learn high-dimensional functions remains a fundamental challenge.
Xvi. functions of positive and negative type, and their connection the theory of integral equations
Mercer, J · 1909
Earlier work this paper cites.
Asymptotic behavior of the eigenvalues of certain integral equations
Widom, H · 1963
Earlier work this paper cites.
Recognition-by-components: a theory of human image understanding
Biederman, I · 1987
Earlier work this paper cites.
Spin glass theory and beyond: An Introduction to the Replica Method and Its Applications , volume 9
Mézard, M., Parisi, G., and Virasoro, M. A · 1987
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Regularization with dot-product kernels
Smola, A., Ovári, Z., and Williamson, R. C · 2000
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
Schölkopf, B., Smola, A. J., Bach, F., et al · 2002
Earlier work this paper cites.
Optimal rates for the regularized least-squares algorithm
Caponnetto, A. and De Vito, E · 2007
Earlier work this paper cites.
Random features for large-scale kernel machines
Rahimi, A. and Recht, B · 2007
Earlier work this paper cites.
Kernel methods for deep learning
Cho, Y. and Saul, L. K · 2009
Earlier work this paper cites.
Spherical harmonics and approximations on the unit sphere: an introduction , volume 2044
Atkinson, K. and Han, W · 2012
Earlier work this paper cites.
Invariant scattering convolution networks
Bruna, J. and Mallat, S · 2013
Earlier work this paper cites.
Spherical harmonics in p dimensions
Efthimiou, C. and Frye, C · 2014
Earlier work this paper cites.
Eigenvalues of dot-product kernels on the sphere
Azevedo, D. and Menegatto, V. A · 2015
Earlier work this paper cites.
Norm-based capacity control in neural networks
Neyshabur, B., Tomioka, R., and Srebro, N · 2015
Earlier work this paper cites.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
Daniely, A., Frostig, R., and Singer, Y · 2016
Earlier work this paper cites.
Breaking the curse of dimensionality with convex neural networks
Bach, F · 2017
Earlier work this paper cites.
Deep learning scaling is predictable, empirically
Hestness, J., Narang, S., Ardalani, N., Diamos, G., Jun, H., Kianinejad, H., Patwary, M., Ali, M., Yang, Y., and Zhou, Y · 2017
Earlier work this paper cites.
Deep neural networks as gaussian processes
Lee, J., Bahri, Y., Novak, R., Schoenholz, S. S., Pennington, J., and Sohl-Dickstein, J · 2017
Earlier work this paper cites.
Why and when can deep-but not shallow-networks avoid the curse of dimensionality: a review
Poggio, T., Mhaskar, H., Rosasco, L., Miranda, B., and Liao, Q · 2017
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A., Gabriel, F., and Hongler, C · 2018
Cited alongside, same era.
Gaussian processes and kernel methods: A review on connections and equivalences
Kanagawa, M., Hennig, P., Sejdinovic, D., and Sriperumbudur, B. K · 2018
Cited alongside, same era.
On the generalization of equivariance and convolution in neural networks to the action of compact groups
Kondor, R. and Trivedi, S · 2018
Cited alongside, same era.
Foundations of machine learning
Mohri, M., Rostamizadeh, A., and Talwalkar, A · 2018
Cited alongside, same era.
Building bayesian neural networks with blocks: On structure, interpretability and uncertainty
Zhou, H. H., Xiong, Y., and Singh, V · 2018
Cited alongside, same era.
On the sample complexity of learning under geometric stability
Bietti, A., Venturi, L., and Bruna, J · 2021
Later among the works it cites.
Spectral bias and task-model alignment explain generalization in kernel regression and infinitely wide neural networks
Canatar, A., Bordelon, B., and Pehlevan, C · 2021
Later among the works it cites.
Generalization error rates in kernel regression: The crossover from the noiseless to noisy regime
Cui, H., Loureiro, B., Krzakala, F., and Zdeborova, L · 2021
Later among the works it cites.
Locality defeats the curse of dimensionality in convolutional teacher-student scenarios
Favero, A., Cagnetta, F., and Wyart, M · 2021
Later among the works it cites.
Posterior contraction for deep gaussian process priors
Finocchio, G. and Schmidt-Hieber, J · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On exact computation with an infinitely wide neural net
Arora, S., Du, S. S., Hu, W., Li, Z., Salakhutdinov, R. R., and Wang, R · 2019
Cited alongside, same era.
On the inductive bias of neural tangent kernels
Bietti, A. and Mairal, J · 2019
Cited alongside, same era.
Fantastic generalization measures and where to find them
Jiang, Y., Neyshabur, B., Mobahi, H., Krishnan, D., and Bengio, S · 2019
Cited alongside, same era.
Wide neural networks of any depth evolve as linear models under gradient descent
Lee, J., Xiao, L., Schoenholz, S., Bahri, Y., Novak, R., Sohl-Dickstein, J., and Pennington, J · 2019
Cited alongside, same era.
Bayesian deep convolutional networks with many channels are gaussian processes
Novak, R., Xiao, L., Bahri, Y., Lee, J., Yang, G., Abolafia, D. A., Pennington, J., and Sohl-dickstein, J · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al · 2019
Cited alongside, same era.
High-dimensional statistics: A non-asymptotic viewpoint , volume 48
Wainwright, M. J · 2019
Cited alongside, same era.
Learning curves of generic features maps for realistic datasets with a teacher-student model
Loureiro, B., Gerbelot, C., Cui, H., Goldt, S., Krzakala, F., Mezard, M., and Zdeborová, L · 2021
Later among the works it cites.
Computational separation between convolutional and fully-connected networks
Malach, E. and Shalev-Shwartz, S · 2021
Later among the works it cites.
Learning with convolution and pooling operations in kernel methods
Misiakiewicz, T. and Mei, S · 2021
Later among the works it cites.
How isotropic kernels perform on simple invariants
Paccolat, J., Spigler, S., and Wyart, M · 2021
Later among the works it cites.
Relative stability toward diffeomorphisms indicates performance in deep nets
Petrini, L., Favero, A., Geiger, M., and Wyart, M · 2021
Later among the works it cites.
A spectral analysis of dot-product kernels
Scetbon, M. and Harchaoui, Z · 2021
Later among the works it cites.
The merged-staircase property: a necessary and nearly sufficient condition for sgd learning of sparse functions on two-layer neural networks
Abbe, E., Adsera, E. B., and Misiakiewicz, T · 2022
Closest in time.
Approximation and learning with deep convolutional models: a kernel perspective
Bietti, A · 2022
Closest in time.
On the spectral bias of convolutional neural tangent and gaussian process kernels
Geifman, A., Galun, M., Jacobs, D., and Basri, R · 2022
Closest in time.
On the inability of gaussian process regression to optimally learn compositional functions
Giordano, M., Ray, K., and Schmidt-Hieber, J · 2022
Closest in time.
Data-driven emergence of convolutional structure in neural networks
Ingrosso, A. and Goldt, S · 2022
Closest in time.
Failure and success of the spectral bias prediction for laplace kernel ridge regression: the case of low-dimensional data
Tomasini, U. M., Sclocchi, A., and Wyart, M · 2022
Closest in time.
Eigenspace restructuring: a principle of space and frequency in neural networks
Xiao, L · 2022
Closest in time.
Synergy and symmetry in deep learning: Interactions between the data, model, and inference algorithm
Xiao, L. and Pennington, J · 2022
Closest in time.
Norm-based generalization bounds for compositionally sparse neural networks
Galanti, T., Xu, M., Galanti, L., and Poggio, T · 2023
Closest in time.