Fetching the paper…
Reading the bibliography…
Many statistical estimators for high-dimensional linear regression are M-estimators, formed through minimizing a data-dependent square loss function plus a regularizer.
Estimation of the mean of a multivariate normal distribution
Stein, C. M. (1981) · 1981
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Tibshirani, R. (1996) · 1996
Earlier work this paper cites.
Atomic decomposition by basis pursuit
Chen, S. S., Donoho, D. L., and Saunders, M. A. (2001) · 2001
Earlier work this paper cites.
Variable selection via nonconcave penalized likelihood and its oracle properties
Fan, J. and Li, R. (2001) · 2001
Earlier work this paper cites.
The estimation of prediction error: covariance penalties and cross-validation
Efron, B. (2004) · 2004
Earlier work this paper cites.
Learning from examples as an inverse problem
Vito, E. D., Rosasco, L., Caponnetto, A., Giovannini, U. D., and Odone, F. (2005) · 2005
Earlier work this paper cites.
Boosting with early stopping: Convergence and consistency
Zhang, T. and Yu, B. (2005) · 2005
Earlier work this paper cites.
Compressed sensing
Donoho, D. L. (2006) · 2006
Earlier work this paper cites.
The adaptive lasso and its oracle properties
Zou, H. (2006) · 2006
Earlier work this paper cites.
The dantzig selector: Statistical estimation when p is much larger than n
Candes, E. and Tao, T. (2007) · 2007
Earlier work this paper cites.
Understanding implicit regularization in over-parameterized nonlinear statistical model
Fan, J., Yang, Z., and Yu, M. (2020) · 2007
Earlier work this paper cites.
On early stopping in gradient descent learning
Yao, Y., Rosasco, L., and Caponnetto, A. (2007) · 2007
Cited alongside, same era.
On the “degrees of freedom” of the lasso
Zou, H., Hastie, T., and Tibshirani, R. (2007) · 2007
Cited alongside, same era.
The restricted isometry property and its implications for compressed sensing
Candes, E. J. (2008) · 2008
Cited alongside, same era.
Sure independence screening for ultrahigh dimensional feature space
Fan, J. and Lv, J. (2008) · 2008
Cited alongside, same era.
Simultaneous analysis of lasso and dantzig selector
Bickel, P. J., Ritov, Y., Tsybakov, A. B., et al. (2009) · 2009
Cited alongside, same era.
Compressed sensing and best k-term approximation
Cohen, A., Dahmen, W., and DeVore, R. (2009) · 2009
High-dimensional statistics with a view toward applications in biology
Bühlmann, P., Kalisch, M., and Meier, L. (2014) · 2014
Later among the works it cites.
Early stopping and non-parametric regression: an optimal data-dependent stopping rule
Raskutti, G., Wainwright, M. J., and Yu, B. (2014) · 2014
Later among the works it cites.
Rare and weak effects in large-scale inference: methods and phase diagrams
Jin, J. and Ke, Z. T. (2016) · 2016
Later among the works it cites.
Gradient descent only converges to minimizers
Lee, J. D., Simchowitz, M., Jordan, M. I., and Recht, B. (2016) · 2016
Later among the works it cites.
Implicit regularization in matrix factorization
Gunasekar, S., Woodworth, B. E., Bhojanapalli, S., Neyshabur, B., and Srebro, N. (2017) · 2017
Later among the works it cites.
Lasso, fractional norm and structured sparse estimation using a hadamard product parametrization
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Information-theoretic limits on sparsity recovery in the high-dimensional and noisy setting
Wainwright, M. J. (2009) · 2009
Cited alongside, same era.
Regularization paths for generalized linear models via coordinate descent
Friedman, J., Hastie, T., and Tibshirani, R. (2010) · 2010
Cited alongside, same era.
Nearly unbiased variable selection under minimax concave penalty
Zhang, C.-H. (2010) · 2010
Cited alongside, same era.
Coordinate descent algorithms for nonconvex penalized regression, with applications to biological feature selection
Breheny, P. and Huang, J. (2011) · 2011
Cited alongside, same era.
Hanson-wright inequality and sub-gaussian concentration
Rudelson, M., Vershynin, R., et al. (2013) · 2013
Cited alongside, same era.
Hoff, P. D. (2017) · 2017
Later among the works it cites.
Characterizing implicit bias in terms of optimization geometry
Gunasekar, S., Lee, J., Soudry, D., and Srebro, N. (2018) · 2018
Later among the works it cites.
Algorithmic regularization in over-parameterized matrix sensing and neural networks with quadratic activations
Li, Y., Ma, T., and Zhang, H. (2018) · 2018
Later among the works it cites.
The implicit bias of gradient descent on separable data
Soudry, D., Hoffer, E., Nacson, M. S., Gunasekar, S., and Srebro, N. (2018) · 2018
Later among the works it cites.
Pathwise coordinate optimization for sparse learning: Algorithm and theory
Zhao, T., Liu, H., and Zhang, T. (2018) · 2018
Later among the works it cites.
Implicit regularization for optimal sparse recovery
Vaskevicius, T., Kanade, V., and Rebeschini, P. (2019) · 2019
Closest in time.