Fetching the paper…
Reading the bibliography…
In this paper, we propose a StochAstic Recursive grAdient algoritHm (SARAH), as well as its practical variant SARAH+, as a novel approach to the finite-sum minimization problems.
A stochastic approximation method
Robbins, Herbert and Monro, Sutton · 1951
Earlier work this paper cites.
Online learning and stochastic approximations
Bottou, Léon · 1998
Earlier work this paper cites.
Introductory lectures on convex optimization : a basic course
Nesterov, Yurii · 2004
Earlier work this paper cites.
A proximal stochastic gradient method with progressive variance reduction
Xiao, Lin and Zhang, Tong · 2004
Earlier work this paper cites.
Pegasos: Primal estimated sub-gradient solver for SVM
Shalev-Shwartz, Shai, Singer, Yoram, and Srebro, Nathan · 2007
Earlier work this paper cites.
A fast iterative shrinkage-thresholding algorithm for linear inverse problems
Beck, Amir and Teboulle, Marc · 2009
Earlier work this paper cites.
The Elements of Statistical Learning: Data Mining, Inference, and Prediction
Hastie, Trevor, Tibshirani, Robert, and Friedman, Jerome · 2009
Earlier work this paper cites.
Better mini-batch algorithms via accelerated gradient methods
Cotter, Andrew, Shamir, Ohad, Srebro, Nati, and Sridharan, Karthik · 2011
Earlier work this paper cites.
Pegasos: Primal estimated sub-gradient solver for SVM
Shalev-Shwartz, Shai, Singer, Yoram, Srebro, Nathan, and Cotter, Andrew · 2011
Cited alongside, same era.
A stochastic gradient method with an exponential convergence rate for finite training sets
Le Roux, Nicolas, Schmidt, Mark, and Bach, Francis · 2012
Cited alongside, same era.
Accelerating stochastic gradient descent using predictive variance reduction
Johnson, Rie and Zhang, Tong · 2013
Cited alongside, same era.
Semi-stochastic gradient descent methods
Konečný, Jakub and Richtárik, Peter · 2013
Cited alongside, same era.
Optimization with first-order surrogate functions
Mairal, Julien · 2013
Cited alongside, same era.
Stochastic dual coordinate ascent methods for regularized loss
Improved SVRG for Non-Strongly-Convex or Sum-of-Non-Convex Objectives
Allen-Zhu, Zeyuan and Yuan, Yang · 2016
Later among the works it cites.
Optimization methods for large-scale machine learning
Bottou, Léon, Curtis, Frank E, and Nocedal, Jorge · 2016
Later among the works it cites.
Mini-batch semi-stochastic gradient descent in the proximal setting
Konečný, Jakub, Liu, Jie, Richtárik, Peter, and Takáč, Martin · 2016
Later among the works it cites.
Stochastic variance reduction for nonconvex optimization
Reddi, Sashank J., Hefny, Ahmed, Sra, Suvrit, Póczos, Barnabás, and Smola, Alexander J · 2016
Later among the works it cites.
Minimizing finite sums with the stochastic average gradient
Schmidt, Mark, Le Roux, Nicolas, and Bach, Francis · 2016
Later among the works it cites.
Katyusha: The First Direct Acceleration of Stochastic Gradient Methods
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shalev-Shwartz, Shai and Zhang, Tong · 2013
Cited alongside, same era.
Mini-batch primal and dual methods for SVMs
Takáč, Martin, Bijral, Avleen Singh, Richtárik, Peter, and Srebro, Nathan · 2013
Cited alongside, same era.
SAGA: A fast incremental gradient method with support for non-strongly convex composite objectives
Defazio, Aaron, Bach, Francis, and Lacoste-Julien, Simon · 2014
Cited alongside, same era.
Allen-Zhu, Zeyuan · 2017
Closest in time.
A double incremental aggregated gradient method with linear convergence rate for large-scale optimization
Mokhtari, Aryan, Gürbüzbalaban, Mert, and Ribeiro, Alejandro · 2017
Closest in time.