Fetching the paper…
Reading the bibliography…
In high-dimensional and/or non-parametric regression problems, regularization (or penalization) is used to control model complexity and induce desired structure.
Nelder, J. A. & Mead, R. (1965), ‘A simplex method for function minimization’, The computer journal
1965
Earlier work this paper cites.
Rellich, F. (1969), Perturbation theory of eigenvalue problems
1969
Earlier work this paper cites.
Wahba, G. (1981), ‘Spline interpolation and smoothing on the sphere’, SIAM Journal on Scientific and Statistical Computing
1981
Earlier work this paper cites.
Watson, G. A. (1992), ‘Characterization of the subdifferential of some matrix norms’, Linear algebra and its applications
1992
Earlier work this paper cites.
Donoho, D. L. & Johnstone, J. M. (1994), ‘Ideal spatial adaptation by wavelet shrinkage’, Biometrika
1994
Earlier work this paper cites.
Neal, R. M. (1996), ‘Bayesian learning for neural networks’
1996
Earlier work this paper cites.
Tibshirani, R. (1996), ‘Regression shrinkage and selection via the lasso’, Journal of the Royal Statistical Society. Series B (Methodological)
1996
Earlier work this paper cites.
Mammen, E., van de Geer, S. et al. (1997), ‘Locally adaptive regression splines’, The Annals of Statistics
1997
Earlier work this paper cites.
Larsen, J., Svarer, C., Andersen, L. N. & Hansen, L. K. (1998), Adaptive regularization in neural network modeling, in
1998
Earlier work this paper cites.
Fazel, M. (2002), Matrix rank minimization with applications, PhD thesis, PhD thesis, Stanford University
2002
Earlier work this paper cites.
Zou, H. & Hastie, T. (2003), ‘Regression shrinkage and selection via the elastic net, with applications to microarrays’, Journal of the Royal Statistical Society: Series B. v67
2003
Earlier work this paper cites.
Boyd, S. & Vandenberghe, L. (2004), Convex optimization
2004
Earlier work this paper cites.
Srebro, N. (2004), Learning with matrix factorizations, PhD thesis, Citeseer
2004
Cited alongside, same era.
Subramanian, A., Tamayo, P., Mootha, V. K., Mukherjee, S., Ebert, B. L., Gillette, M. A., Paulovich, A., Pomeroy, S. L., Golub, T. R., Lander, E. S. et al. (2005), ‘Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles’, Proceedings of the National Academy of Sciences of the United States of America
2005
Cited alongside, same era.
Tibshirani, R., Saunders, M., Rosset, S., Zhu, J. & Knight, K. (2005), ‘Sparsity and smoothness via the fused lasso’, Journal of the Royal Statistical Society: Series B (Statistical Methodology)
2005
Cited alongside, same era.
Burczynski, M. E., Peterson, R. L., Twine, N. C., Zuberek, K. A., Brodeur, B. J., Casciotti, L., Maganti, V., Reddy, P. S., Strahs, A., Immermann, F. et al. (2006), ‘Molecular classification of crohn’s disease and ulcerative colitis patients using transcriptional profiles in peripheral blood mononuclear cells’, The journal of molecular diagnostics
Bergstra, J. S., Bardenet, R., Bengio, Y. & Kégl, B. (2011), Algorithms for hyper-parameter optimization, in
2011
Later among the works it cites.
Bühlmann, P. & Van De Geer, S. (2011), Statistics for high-dimensional data: methods, theory and applications
2011
Later among the works it cites.
Candès, E. J., Li, X., Ma, Y. & Wright, J. (2011), ‘Robust principal component analysis?’, Journal of the ACM (JACM)
2011
Later among the works it cites.
Hutter, F., Hoos, H. H. & Leyton-Brown, K. (2011), Sequential model-based optimization for general algorithm configuration, in
2011
Later among the works it cites.
Tibshirani, R. J., Taylor, J. et al. (2011), ‘The solution path of the generalized lasso’, The Annals of Statistics
2011
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2006
Cited alongside, same era.
Yuan, M. & Lin, Y. (2006), ‘Model selection and estimation in regression with grouped variables’, Journal of the Royal Statistical Society: Series B (Statistical Methodology)
2006
Cited alongside, same era.
SIGKDD, A. & Netflix (2007), ‘Soft modelling by latent variables: the nonlinear iterative partial least squares (nipals) approach’, Proceedings of KDD Cup and Workshop
2007
Cited alongside, same era.
Yao, Y., Rosasco, L. & Caponnetto, A. (2007), ‘On early stopping in gradient descent learning’, Constructive Approximation
2007
Cited alongside, same era.
Foo, C.-s., Do, C. B. & Ng, A. Y. (2008), Efficient multiple hyperparameter learning for log-linear models, in
2008
Cited alongside, same era.
Tsybakov, A. (2008), Introduction to Nonparametric Estimation
2008
Cited alongside, same era.
Kim, S.-J., Koh, K., Boyd, S. & Gorinevsky, D. (2009), ‘ \ \backslash ell_1 trend filtering’, SIAM review
2009
Cited alongside, same era.
Lorbert, A. & Ramadge, P. J. (2010), Descent methods for tuning parameter refinement, in
2010
Cited alongside, same era.
Mazumder, R., Hastie, T. & Tibshirani, R. (2010), ‘Spectral regularization algorithms for learning large incomplete matrices’, Journal of machine learning research
2010
Cited alongside, same era.
Snoek, J., Larochelle, H. & Adams, R. P. (2012), Practical bayesian optimization of machine learning algorithms, in
2012
Later among the works it cites.
2013
Later among the works it cites.
Simon, N., Friedman, J., Hastie, T. & Tibshirani, R. (2013), ‘A sparse-group lasso’, Journal of Computational and Graphical Statistics
2013
Later among the works it cites.
2014
Later among the works it cites.
Maclaurin, D., Duvenaud, D. & Adams, R. P. (2015), Gradient-based hyperparameter optimization through reversible learning, in
2015
Later among the works it cites.
Tibshirani, R. J. (2015), ‘Degrees of freedom and model search’, Statistica Sinica
2015
Later among the works it cites.