Fetching the paper…
Reading the bibliography…
In this paper, we present a generic framework that allows accelerating almost arbitrary non-accelerated deterministic and randomized algorithms for smooth convex optimization problems.
R.T. Rockafellar, Monotone operators and the proximal point algorithm , SIAM journal on control and optimization 14 (1976), pp. 877–898
1976
Earlier work this paper cites.
A.S. Nemirovsky and D.B. Yudin, Problem Complexity and Method Efficiency in Optimization , A Wiley-Interscience publication, Wiley, 1983
1983
Earlier work this paper cites.
B.T. Polyak, Introduction to optimization , Optimization Software, 1987
1987
Earlier work this paper cites.
Y. Hu, Y. Koren, and C. Volinsky, Collaborative filtering for implicit feedback datasets , in 2008 Eighth IEEE International Conference on Data Mining . Ieee, 2008, pp. 263–272
2008
Earlier work this paper cites.
Y. Nesterov and J.P. Vial, Confidence level solutions for stochastic programming , Automatica 44 (2008), pp. 1559–1568
2008
Earlier work this paper cites.
C.C. Chang and C.J. Lin, Libsvm: a library for support vector machines , ACM Transactions on Intelligent Systems and Technology (TIST) 2 (2011), p. 27
2011
Earlier work this paper cites.
Y. Nesterov, Efficiency of coordinate descent methods on huge-scale optimization problems , SIAM Journal on Optimization 22 (2012), pp. 341–362
2012
Earlier work this paper cites.
R.D. Monteiro and B.F. Svaiter, An accelerated hybrid proximal extragradient method for convex optimization and its implications to second-order methods , SIAM Journal on Optimization 23 (2013), pp. 1092–1125
2013
Earlier work this paper cites.
N. Parikh, S. Boyd, et al. , Proximal algorithms , Foundations and Trends® in Optimization 1 (2014), pp. 127–239
2014
Earlier work this paper cites.
S. Shalev-Shwartz and T. Zhang, Accelerated proximal stochastic dual coordinate ascent for regularized loss minimization , in International conference on machine learning . 2014, pp. 64–72
2014
Earlier work this paper cites.
S. Bubeck, Convex optimization: Algorithms and complexity , Foundations and Trends® in Machine Learning 8 (2015), pp. 231–357
2015
Earlier work this paper cites.
J.C. Duchi, M.I. Jordan, M.J. Wainwright, and A. Wibisono, Optimal rates for zero-order convex optimization: The power of two function evaluations , IEEE Trans. Information Theory 61 (2015), pp. 2788–2806
2015
Earlier work this paper cites.
H. Lin, J. Mairal, and Z. Harchaoui, A universal catalyst for first-order optimization , in Advances in neural information processing systems . 2015, pp. 3384–3392
2015
Earlier work this paper cites.
S.J. Wright, Coordinate descent algorithms , Mathematical Programming 151 (2015), pp. 3–34
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
H. Karimi, J. Nutini, and M. Schmidt, Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition , in Joint European Conference on Machine Learning and Knowledge Discovery in Databases . Springer, 2016, pp. 795–811
2016
Earlier work this paper cites.
B. Palaniappan and F. Bach, Stochastic variance reduction methods for saddle-point problems , in Advances in Neural Information Processing Systems . 2016, pp. 1416–1424
2016
Earlier work this paper cites.
A. Beck, First-order methods in optimization , Vol. 25, SIAM, 2017
2017
Earlier work this paper cites.
E. De Klerk, F. Glineur, and A.B. Taylor, On the worst-case complexity of the gradient method with exact line search for smooth strongly convex functions , Optimization Letters 11 (2017), pp. 1185–1199
2017
Earlier work this paper cites.
A. Gasnikov, Universal gradient descent , arXiv preprint arXiv:1711.00394 (2017)
2017
Earlier work this paper cites.
Y. Nesterov and S.U. Stich, Efficiency of the accelerated coordinate descent method on structured optimization problems , SIAM Journal on Optimization 27 (2017), pp. 110–123
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
O. Shamir, An optimal algorithm for bandit and zero-order convex optimization with two-point feedback , Journal of Machine Learning Research 18 (2017), pp. 52:1–52:11
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
A.C. Wilson, L. Mackey, and A. Wibisono, Accelerating Rescaled Gradient Descent: Fast Optimization of Smooth Functions , in Advances in Neural Information Processing Systems . 2019, pp. 13533–13543
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
P. Dvurechensky, A. Gasnikov, and A. Lagunovskaya, Parallel algorithms and probability of large deviation for stochastic convex optimization problems , Numerical Analysis and Applications 11 (2018), pp. 33–37
2018
Cited alongside, same era.
2018
Cited alongside, same era.
K. Mishchenko, F. Iutzeler, J. Malick, and M.R. Amini, A delay-tolerant proximal-gradient algorithm for distributed learning , in International Conference on Machine Learning . 2018, pp. 3587–3595
2018
Cited alongside, same era.
Y. Nesterov, Lectures on convex optimization , Vol. 137, Springer, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
B.E. Woodworth, J. Wang, A. Smith, B. McMahan, and N. Srebro, Graph oracle models, lower bounds, and gaps for parallel stochastic optimization , in Advances in neural information processing systems . 2018, pp. 8496–8506
2018
Cited alongside, same era.
N. Doikov and Y. Nesterov, Contracting proximal methods for smooth convex optimization , SIAM Journal on Optimization 30 (2020), pp. 3146–3169
2020
Closest in time.
2020
Closest in time.
D. Dvinskikh, S. Omelchenko, A. Gasnikov, and A. Tyurin, Accelerated Gradient Sliding for Minimizing a Sum of Functions , in Doklady Mathematics , Vol. 101. Springer, 2020, pp. 244–246
2020
Closest in time.
H. Hendrikx, F. Bach, and L. Massoulié, Dual-free stochastic decentralized optimization with variance reduction , Advances in Neural Information Processing Systems 33 (2020)
2020
Closest in time.
2020
Closest in time.
D. Kamzolov, P. Dvurechensky, and A. Gasnikov, Optimal Combination of Tensor Optimization Methods , in Optimization and Applications: 11th International Conference, OPTIMA 2020, Moscow, Russia, September 28–October 2, 2020, Proceedings . Springer Nature, p. 166
2020
Closest in time.
2020
Closest in time.
D. Kovalev, A. Salim, and P. Richtárik, Optimal and practical algorithms for smooth and strongly convex decentralized optimization , Advances in Neural Information Processing Systems 33 (2020)
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
T. Lin, C. Jin, and M. Jordan, On gradient descent ascent for nonconvex-concave minimax problems , in International Conference on Machine Learning . PMLR, 2020, pp. 6083–6093
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
J. Yang, S. Zhang, N. Kiyavash, and N. He, A catalyst framework for minimax optimization , Advances in Neural Information Processing Systems 33 (2020)
2020
Closest in time.
D. Dvinskikh, D. Kamzolov, A. Gasnikov, P. Dvurechensky, D. Pasechnyk, V. Matykhin, and A. Chernov, Accelerated meta-algorithm for convex optimization , Computational Mathematica and Mathematical Physics 61 (2021)
2021
Closest in time.
A. Gasnikov, Universal gradient descent , MCCME, Moscow, 2021
2021
Closest in time.
2021
Closest in time.
O. Fercoq and P. Richtárik, Accelerated, parallel, and proximal coordinate descent , SIAM Journal on Optimization 25 (2015), pp. 1997–2023
2023
Closest in time.
2034
Closest in time.