Fetching the paper…
Reading the bibliography…
Linear regression is a classical paradigm in statistics.
A. Hoerl and R. Kennard, “Ridge regression: Biased estimation for nonorthogonal problems,” Technometrics , vol. 12, 1970
1970
Earlier work this paper cites.
L. G. Valiant, “A theory of the learnable,” Communications of the ACM , vol. 27, no. 11, pp. 1134–1142, 1984
1984
Earlier work this paper cites.
Y. M. Shtar’kov, “Universal sequential coding of single messages,” Problemy Peredachi Informatsii , vol. 23, no. 3, pp. 3–17, 1987
1987
Earlier work this paper cites.
V. Vapnik, “Principles of risk minimization for learning theory,” in Advances in neural information processing systems , 1992, pp. 831–838
1992
Earlier work this paper cites.
C. L. Lawson and R. J. Hanson, Solving least squares problems . Siam, 1995, vol. 15
1995
Earlier work this paper cites.
M. H. Hayes, “9.4: Recursive least squares,” Statistical Digital Signal Processing and Modeling , p. 541, 1996
1996
Earlier work this paper cites.
N. Merhav and M. Feder, “Universal prediction,” IEEE Transactions on Information Theory , vol. 44, no. 6, pp. 2124–2147, 1998
1998
Cited alongside, same era.
E. D. Sontag, “Vc dimension of neural networks,” NATO ASI Series F Computer and Systems Sciences , vol. 168, pp. 69–96, 1998
1998
Cited alongside, same era.
W. H. Press, S. A. Teukolsky, W. T. Vetterling, and B. P. Flannery, “Section 2.7. 1 sherman–morrison formula,” Numerical Recipes: The Art of Scientific Computing (3rd ed.). Cambridge University Press, New York , vol. 1, pp. 55–67, 2007
2007
Cited alongside, same era.
T. Roos and J. Rissanen, “On sequentially normalized maximum likelihood models,” 2008
2008
Cited alongside, same era.
T. Roos, T. Silander, P. Kontkanen, and P. Myllymaki, “Bayesian network structure learning using factorized nml universal models,” in Information Theory and Applications Workshop, 2008 . IEEE, 2008, pp. 272–276
G. James, D. Witten, T. Hastie, and R. Tibshirani, An introduction to statistical learning . Springer, 2013, vol. 112
2013
Later among the works it cites.
——, The nature of statistical learning theory . Springer science & business media, 2013
2013
Later among the works it cites.
C. Cardinali, “Observation influence diagnostic of a data assimilation system,” in Data Assimilation for Atmospheric, Oceanic and Hydrologic Applications (Vol. II) . Springer, 2013, pp. 89–110
2013
Later among the works it cites.
2018
Later among the works it cites.
K. Bibas, Y. Fogel, and M. Feder, “Deep pnml: Predictive normalized maximum likelihood for deep neural networks,” 2019
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2008
Cited alongside, same era.
H. Trevor, T. Robert, and F. JH, “The elements of statistical learning: data mining, inference, and prediction,” 2009
2009
Cited alongside, same era.