Fetching the paper…
Reading the bibliography…
The empirical success of deep convolutional networks on tasks involving high-dimensional data such as images or audio suggests that they can efficiently approximate certain functions that are well-suited for such tasks.
Spline models for observational data , volume 59
Grace Wahba · 1990
Earlier work this paper cites.
Integral transforms, reproducing kernels and their applications , volume 369
Saburou Saitoh · 1997
Earlier work this paper cites.
Object recognition from local scale-invariant features
David G Lowe · 1999
Earlier work this paper cites.
Tensor product space anova models
Yi Lin · 2000
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
Bernhard Schölkopf and Alexander J Smola · 2001
Earlier work this paper cites.
Regularization with dot-product kernels
Alex J Smola, Zoltan L Ovari, and Robert C Williamson · 2001
Earlier work this paper cites.
Numerical operator calculus in higher dimensions
Gregory Beylkin and Martin J Mohlenkamp · 2002
Earlier work this paper cites.
On the mathematical foundations of learning
Felipe Cucker and Steve Smale · 2002
Earlier work this paper cites.
Distance-based classification with lipschitz functions
Ulrike von Luxburg and Olivier Bousquet · 2004
Earlier work this paper cites.
Mercer’s theorem, feature maps, and smoothing
Ha Quang Minh, Partha Niyogi, and Yuan Yao · 2006
Earlier work this paper cites.
Optimal rates for the regularized least-squares algorithm
Andrea Caponnetto and Ernesto De Vito · 2007
Earlier work this paper cites.
Kernel methods for deep learning
Youngmin Cho and Lawrence K Saul · 2009
Earlier work this paper cites.
A new scheme for the tensor representation
Wolfgang Hackbusch and Stefan Kühn · 2009
Earlier work this paper cites.
Tensor products of sobolev–besov spaces and applications to approximation from the hyperbolic cross
Winfried Sickel and Tino Ullrich · 2009
Earlier work this paper cites.
Aggregating local image descriptors into compact codes
Hervé Jégou, Florent Perronnin, Matthijs Douze, Jorge Sánchez, Patrick Pérez, and Cordelia Schmid · 2011
Earlier work this paper cites.
Group invariant scattering
Stéphane Mallat · 2012
Earlier work this paper cites.
Invariant scattering convolution networks
Joan Bruna and Stéphane Mallat · 2013
Earlier work this paper cites.
Image classification with the fisher vector: Theory and practice
Jorge Sánchez, Florent Perronnin, Thomas Mensink, and Jakob Verbeek · 2013
Earlier work this paper cites.
Spherical harmonics in p dimensions
Costas Efthimiou and Christopher Frye · 2014
Earlier work this paper cites.
Convolutional kernel networks
Julien Mairal, Piotr Koniusz, Zaid Harchaoui, and Cordelia Schmid · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Matthew D Zeiler and Rob Fergus · 2014
Earlier work this paper cites.
Convolutional rectifier networks as generalized tensor decompositions
Nadav Cohen and Amnon Shashua · 2016
Cited alongside, same era.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
Amit Daniely, Roy Frostig, and Yoram Singer · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
End-to-End Kernel Learning with Supervised Convolutional Kernel Networks
Julien Mairal · 2016
Cited alongside, same era.
Deep vs. shallow networks: An approximation theory perspective
Hrushikesh N Mhaskar and Tomaso Poggio · 2016
Cited alongside, same era.
Inductive bias of deep convolutional networks through pooling geometry
Nadav Cohen and Amnon Shashua · 2017
Greg Yang · 2019
Later among the works it cites.
Backward feature correction: How deep learning performs deep learning
Zeyuan Allen-Zhu and Yuanzhi Li · 2020
Later among the works it cites.
Towards understanding hierarchical learning: Benefits of neural representations
Minshuo Chen, Yu Bai, Jason D Lee, Tuo Zhao, Huan Wang, Caiming Xiong, and Richard Socher · 2020
Later among the works it cites.
Implicit bias of gradient descent for wide two-layer neural networks trained with the logistic loss
Lenaic Chizat and Francis Bach · 2020
Later among the works it cites.
On the similarity between the laplace and neural tangent kernels
Amnon Geifman, Abhay Yadav, Yoni Kasten, Meirav Galun, David Jacobs, and Ronen Basri · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Why and when can deep-but not shallow-networks avoid the curse of dimensionality: a review
Tomaso Poggio, Hrushikesh Mhaskar, Lorenzo Rosasco, Brando Miranda, and Qianli Liao · 2017
Cited alongside, same era.
Convexified convolutional neural networks
Y. Zhang, P. Liang, and M. J. Wainwright · 2017
Cited alongside, same era.
How many samples are needed to estimate a convolutional neural network?
Simon S Du, Yining Wang, Xiyu Zhai, Sivaraman Balakrishnan, Ruslan Salakhutdinov, and Aarti Singh · 2018
Cited alongside, same era.
Implicit bias of gradient descent on linear convolutional networks
Suriya Gunasekar, Jason D Lee, Daniel Soudry, and Nati Srebro · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Cited alongside, same era.
A mathematical theory of deep convolutional neural networks for feature extraction
Thomas Wiatowski and Helmut Bölcskei · 2018
Cited alongside, same era.
Later among the works it cites.
When do neural networks outperform kernel methods?
Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, and Andrea Montanari · 2020
Later among the works it cites.
Denoising and regularization via exploiting the structural bias of convolutional generators
Reinhard Heckel and Mahdi Soltanolkotabi · 2020
Later among the works it cites.
Finite versus infinite neural networks: an empirical study
Jaehoon Lee, Samuel Schoenholz, Jeffrey Pennington, Ben Adlam, Lechao Xiao, Roman Novak, and Jascha Sohl-Dickstein · 2020
Later among the works it cites.
Harmonic decompositions of convolutional networks
Meyer Scetbon and Zaid Harchaoui · 2020
Later among the works it cites.
Nonparametric regression using deep neural networks with relu activation function
Johannes Schmidt-Hieber et al · 2020
Later among the works it cites.
Neural kernels without tangents
Vaishaal Shankar, Alex Fang, Wenshuo Guo, Sara Fridovich-Keil, Jonathan Ragan-Kelley, Ludwig Schmidt, and Benjamin Recht · 2020
Later among the works it cites.
Learning Theory from First Principles (draft)
Francis Bach · 2021
Closest in time.
Deep equals shallow for ReLU networks in kernel regimes
Alberto Bietti and Francis Bach · 2021
Closest in time.
Deep neural tangent kernel and laplace kernel have the same rkhs
Lin Chen and Sheng Xu · 2021
Closest in time.
Locality defeats the curse of dimensionality in convolutional teacher-student scenarios
Alessandro Favero, Francesco Cagnetta, and Matthieu Wyart · 2021
Closest in time.
Why are convolutional nets more sample-efficient than fully-connected nets?
Zhiyuan Li, Yi Zhang, and Sanjeev Arora · 2021
Closest in time.
Computational separation between convolutional and fully-connected networks
Eran Malach and Shai Shalev-Shwartz · 2021
Closest in time.
Learning with invariances in random features and kernel models
Song Mei, Theodor Misiakiewicz, and Andrea Montanari · 2021
Closest in time.
Learning with convolution and pooling operations in kernel methods
Theodor Misiakiewicz and Song Mei · 2021
Closest in time.
The unreasonable effectiveness of patches in deep convolutional kernels methods
Louis Thiry, Michael Arbel, Eugene Belilovsky, and Edouard Oyallon · 2021
Closest in time.