Fetching the paper…
Reading the bibliography…
An interesting observation in artificial neural networks is their favorable generalization error despite typically being extremely overparameterized.
Functions of positive and negative type and their commection with the theory of integral equations
J Mercer · 1909
Earlier work this paper cites.
The sample complexity of pattern classification with neural networks: the size of the weights is more important than the size of the network
Peter L Bartlett · 1998
Earlier work this paper cites.
Elements of information theory
Thomas M Cover · 1999
Earlier work this paper cites.
On the inductive proof of Legendre addition theorem
Kamil Maleček and Zbyněk Nádeník · 2001
Earlier work this paper cites.
Rademacher and Gaussian complexities: Risk bounds and structural results
Peter L Bartlett and Shahar Mendelson · 2002
Earlier work this paper cites.
Learning bounds for kernel regression using effective data dimensionality
Tong Zhang · 2005
Earlier work this paper cites.
Random features for large-scale kernel machines
Ali Rahimi, Benjamin Recht, et al · 2007
Earlier work this paper cites.
Support vector machines
Ingo Steinwart and Andreas Christmann · 2008
Earlier work this paper cites.
Neural network learning: Theoretical foundations
Martin Anthony and Peter L Bartlett · 2009
Earlier work this paper cites.
Kernel methods for deep learning
Youngmin Cho and Lawrence Saul · 2009
Earlier work this paper cites.
Gaussian process optimization in the bandit setting: No regret and experimental design
Niranjan Srinivas, Andreas Krause, Sham Kakade, and Matthias W. Seeger · 2010
Earlier work this paper cites.
Applied Analysis
John K. Hunter and Bruno Nachtergaele · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Bayesian learning for neural networks , volume 118
Radford M Neal · 2012
Earlier work this paper cites.
Finite-time analysis of kernelised contextual bandits
Michal Valko, Nathan Korda, Rémi Munos, Ilias Flaounas, and Nello Cristianini · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Cited alongside, same era.
Norm-based capacity control in neural networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2015
Cited alongside, same era.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
Amit Daniely, Roy Frostig, and Yoram Singer · 2016
Cited alongside, same era.
Introduction to Fourier Analysis on Euclidean Spaces (PMS-32), Volume 32
Elias M Stein and Guido Weiss · 2016
Cited alongside, same era.
Gintare Karolina Dziugaite and Daniel M Roy · 2017
Cited alongside, same era.
Jax: composable transformations of python+ numpy programs
Matérn Gaussian processes on Riemannian manifolds
Viacheslav Borovitskiy, Alexander Terenin, Peter Mostowsky, and Marc Deisenroth · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Later among the works it cites.
Sparse gaussian processes with spherical harmonic features
Vincent Dutordoir, Nicolas Durrande, and James Hensman · 2020
Later among the works it cites.
Multiplicative filter networks
Rizal Fathony, Anit Kumar Sahu, Devin Willmott, and J Zico Kolter · 2020
Later among the works it cites.
On the similarity between the Laplace and neural tangent kernels
Amnon Geifman, Abhay Yadav, Yoni Kasten, Meirav Galun, David Jacobs, and Basri Ronen · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, and Skye Wanderman-Milne · 2018
Cited alongside, same era.
Gaussian process behaviour in wide deep neural networks
Alexander G. de G. Matthews, Jiri Hron, Mark Rowland, Richard E. Turner, and Zoubin Ghahramani · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clement Hongler · 2018
Cited alongside, same era.
Gaussian processes and kernel methods: A review on connections and equivalences
Motonobu Kanagawa, Philipp Hennig, Dino Sejdinovic, and Bharath K Sriperumbudur · 2018
Cited alongside, same era.
Deep neural networks as Gaussian processes
Jaehoon Lee, Jascha Sohl-dickstein, Jeffrey Pennington, Roman Novak, Sam Schoenholz, and Yasaman Bahri · 2018
Cited alongside, same era.
Gaussian process optimization with adaptive sketching: scalable and no regret
Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, and Lorenzo Rosasco · 2019
Cited alongside, same era.
On lazy training in differentiable programming
Lénaïc Chizat, Edouard Oyallon, and Francis Bach · 2019
Cited alongside, same era.
Chaoyue Liu, Libin Zhu, and Mikhail Belkin · 2020
Later among the works it cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng · 2020
Later among the works it cites.
Implicit neural representations with periodic activation functions
Vincent Sitzmann, Julien Martel, Alexander Bergman, David Lindell, and Gordon Wetzstein · 2020
Later among the works it cites.
Regularization matters: A nonparametric perspective on overparametrized neural network
Wenjia Wang, Tianyang Hu, Cong Lin, and Guang Cheng · 2020
Later among the works it cites.
On function approximation in reinforcement learning: Optimism in the face of large state spaces
Zhuoran Yang, Chi Jin, Zhaoran Wang, Mengdi Wang, and Michael I Jordan · 2020
Later among the works it cites.
Neural contextual bandits with UCB-based exploration
Dongruo Zhou, Lihong Li, and Quanquan Gu · 2020
Later among the works it cites.
Deep neural tangent kernel and Laplace kernel have the same RKHS
Lin Chen and Sheng Xu · 2021
Closest in time.
Quanquan Gu, Amin Karbasi, Khashayar Khosravi, Vahab Mirrokni, and Dongruo Zhou · 2021
Closest in time.
A short note on the relationship of information gain and eluder dimension
Kaixuan Huang, Sham M Kakade, Jason D Lee, and Qi Lei · 2021
Closest in time.
Neural Thompson sampling
Weitong ZHANG, Dongruo Zhou, Lihong Li, and Quanquan Gu · 2021
Closest in time.