Fetching the paper…
Reading the bibliography…
We study the average $\mbox{CV}_{loo}$ stability of kernel ridge-less regression and derive corresponding risk bounds.
Benign overfitting in linear regression
Peter L. Bartlett, Philip M. Long, Gábor Lugosi, and Alexander Tsigler · 1906
Earlier work this paper cites.
Distribution of eigenvalues for some sets of random matrices
V. A. Marchenko and L. A. Pastur · 1967
Earlier work this paper cites.
Generalized inversion of modified matrices
Carl Meyer · 1973
Earlier work this paper cites.
Interpolation of scattered data: distance matrices and conditionally positive definite functions
C. A. Micchelli · 1986
Earlier work this paper cites.
Stability and generalization
O. Bousquet and A. Elisseeff · 2001
Earlier work this paper cites.
Almost-everywhere algorithmic stability and generalization error
S. Kutin and P. Niyogi · 2002
Earlier work this paper cites.
A revisitation of formulae for the moore–penrose inverse of modified matrices
Jerzy K Baksalary, Oskar Maria Baksalary, and Götz Trenkler · 2003
Earlier work this paper cites.
General conditions for predictivity in learning theory
T. Poggio, R. Rifkin, S. Mukherjee, and P. Niyogi · 2004
Earlier work this paper cites.
Theory of classification: A survey of some recent advances
Stéphane Boucheron, Olivier Bousquet, and Gábor Lugosi · 2005
Earlier work this paper cites.
Learning theory: stability is sufficient for generalization and necessary and sufficient for consistency of empirical risk minimization
Sayan Mukherjee, Partha Niyogi, Tomaso Poggio, and Ryan Rifkin · 2006
Cited alongside, same era.
Support vector machines
Ingo Steinwart and Andreas Christmann · 2008
Cited alongside, same era.
The spectrum of kernel random matrices
Noureddine El Karoui · 2010
Cited alongside, same era.
Learnability, stability and uniform convergence
Shai Shalev-Shwartz, Ohad Shamir, Nathan Srebro, and Karthik Sridharan · 2010
Cited alongside, same era.
Statistics for high-dimensional data: methods, theory and applications
Peter Bühlmann and Sara Van De Geer · 2011
Cited alongside, same era.
Understanding Machine Learning: From Theory to Algorithms
Consistency of Interpolation with Laplace Kernels is a High-Dimensional Phenomenon
Alexander Rakhlin and Xiyu Zhai · 2018
Later among the works it cites.
Reconciling modern machine-learning practice and the classical bias–variance trade-off
Mikhail Belkin, Daniel Hsu, Siyuan Ma, and Soumik Mandal · 2019
Later among the works it cites.
Surprises in High-Dimensional Ridgeless Least Squares Interpolation
Trevor Hastie, Andrea Montanari, Saharon Rosset, and Ryan J. Tibshirani · 2019
Later among the works it cites.
On the Risk of Minimum-Norm Interpolants and Restricted Lower Isometry of Kernels
Tengyuan Liang, Alexander Rakhlin, and Xiyu Zhai · 2019
Later among the works it cites.
The generalization error of random features regression: Precise asymptotics and double descent curve
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shai Shalev-Shwartz and Shai Ben-David · 2014
Cited alongside, same era.
Learning with incremental iterative regularization
Lorenzo Rosasco and Silvia Villa · 2015
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Cited alongside, same era.
Song Mei and Andrea Montanari · 2019
Later among the works it cites.
Double descent in the condition number
T. Poggio, G. Kur, and A. Banburski · 2019
Later among the works it cites.
Just interpolate: Kernel “ridgeless” regression can generalize
Tengyuan Liang, Alexander Rakhlin, et al · 2020
Closest in time.
Stable foundations for learning
Tomaso Poggio · 2020
Closest in time.