Fetching the paper…
Reading the bibliography…
Contrary to the situation with stochastic gradient descent, we argue that when using stochastic methods with variance reduction, such as SDCA, SAG or SVRG, as well as their variants, it could be beneficial to reuse previously used samples instead of fresh samples, even when fresh samples are available.
Convergence of stochastic processes
Pollard, David · 1984
Earlier work this paper cites.
The tradeoffs of large scale learning
Bottou, Léon and Bousquet, Olivier · 2007
Earlier work this paper cites.
A dual coordinate descent method for large-scale linear SVM
Hsieh, Cho-Jui, Chang, Kai-Wei, Lin, Chih-Jen, Keerthi, S. Sathiya, and Sundararajan, S · 2008
Earlier work this paper cites.
Svm optimization: inverse dependence on training set size
Shalev-Shwartz, Shai and Srebro, Nathan · 2008
Earlier work this paper cites.
Fast rates for regularized objectives
Sridharan, Karthik, Srebro, Nathan, and Shalev-Shwartz, Shai · 2009
Earlier work this paper cites.
Libsvm: A library for support vector machines
Chang, Chih-Chung and Lin, Chih-Jen · 2011
Earlier work this paper cites.
Pegasos: primal estimated sub-gradient solver for SVM
Shalev-Shwartz, Shai, Singer, Yoram, Srebro, Nathan, and Cotter, Andrew · 2011
Earlier work this paper cites.
Stochastic gradient tricks
Bottou, Léon · 2012
Earlier work this paper cites.
Making gradient descent optimal for strongly convex stochastic optimization
Rakhlin, Alexander, Shamir, Ohad, and Sridharan, Karthik · 2012
Earlier work this paper cites.
Beneath the valley of the noncommutative arithmetic-geometric mean inequality: conjectures, case-studies, and consequences
Recht, Benjamin and Ré, Christopher · 2012
Earlier work this paper cites.
A stochastic gradient method with an exponential convergence rate for finite training sets
Roux, Nicolas Le, Schmidt, Mark W., and Bach, Francis · 2012
Earlier work this paper cites.
Accelerating stochastic gradient descent using predictive variance reduction
Johnson, Rie and Zhang, Tong · 2013
Cited alongside, same era.
Semi-stochastic gradient descent methods
Konečnỳ, Jakub and Richtárik, Peter · 2013
Cited alongside, same era.
Minimizing finite sums with the stochastic average gradient, 2013
Schmidt, Mark, Roux, Nicolas Le, and Bach, Francis · 2013
Cited alongside, same era.
Stochastic dual coordinate ascent methods for regularized loss
Shalev-Shwartz, Shai and Zhang, Tong · 2013
Cited alongside, same era.
Linear convergence with condition number independent access of full gradients
Zhang, Lijun, Mahdavi, Mehrdad, and Jin, Rong · 2013
Cited alongside, same era.
Stochastic proximal gradient descent with acceleration techniques
Stop wasting my gradients: Practical svrg
Babanezhad, Reza, Ahmed, Mohamed Osama, Virani, Alim, Schmidt, Mark, Konečnỳ, Jakub, and Sallinen, Scott · 2015
Later among the works it cites.
Stochastic dual coordinate ascent with adaptive probabilities
Csiba, Dominik, Qu, Zheng, and Richtarik, Peter · 2015
Later among the works it cites.
Averaged least-mean-squares: Bias-variance trade-offs and optimal sampling distributions
Défossez, Alexandre and Bach, Francis R · 2015
Later among the works it cites.
Why random reshuffling beats stochastic gradient descent
Gürbüzbalaban, Mert, Ozdaglar, Asu, and Parrilo, Pablo · 2015
Later among the works it cites.
Neighborhood watch: Stochastic gradient descent with neighbors
Hofmann, Thomas, Lucchi, Aurelien, and McWilliams, Brian · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nitanda, Atsushi · 2014
Cited alongside, same era.
Randomized dual coordinate ascent with arbitrary sampling
Qu, Zheng, Richtarik, Peter, and Zhang, Tong · 2014
Cited alongside, same era.
Accelerated proximal stochastic dual coordinate ascent for regularized loss minimization
Shalev-Shwartz, Shai and Zhang, Tong · 2014
Cited alongside, same era.
Stochastic dual coordinate ascent with alternating direction method of multipliers
Suzuki, Taiji · 2014
Cited alongside, same era.
A proximal stochastic gradient method with progressive variance reduction
Xiao, Lin and Zhang, Tong · 2014
Cited alongside, same era.
Stochastic optimization with importance sampling
Zhao, Peilin and Zhang, Tong · 2014
Cited alongside, same era.
Saga: A fast incremental gradient method with support for non-strongly convex composite objectives
Defazio, Aaron, Bach, Francis, and Lacoste-Julien, Simon
Cited in the paper.
Lan, Guanghui · 2015
Later among the works it cites.
Incremental majorization-minimization optimization with application to large-scale machine learning
Mairal, Julien · 2015
Later among the works it cites.
Local smoothness in variance reduced optimization
Vainsencher, Daniel, Liu, Han, and Zhang, Tong · 2015
Later among the works it cites.
Stochastic primal-dual coordinate method for regularized empirical risk minimization
Zhang, Yuchen and Xiao, Lin · 2015
Later among the works it cites.
Even faster accelerated coordinate descent using non-uniform sampling
Zhu, Zeyuan Allen, Qu, Zheng, Richtarik, Peter, and Yuan, Yang · 2015
Later among the works it cites.