Fetching the paper…
Reading the bibliography…
We consider the problem of minimizing the sum of two convex functions: one is the average of a large number of smooth component functions, and the other is a general convex function that admits a simple proximal mapping.
Convex Analysis
R. T. Rockafellar · 1970
Earlier work this paper cites.
Splitting algorithms for the sum of two nonlinear operators
P.-L. Lions and B. Mercier · 1979
Earlier work this paper cites.
Convergence rates in forward-backward splitting
G. H.-G. Chen and R. T. Rockafellar · 1997
Earlier work this paper cites.
A modified forward-backward splitting method for maximal monotone mappings
P. Tseng · 2000
Earlier work this paper cites.
RCV1: A new benchmark collection for text categorization research
D. D. Lewis, Y. Yang, T. Rose, and F. Li · 2004
Earlier work this paper cites.
Introductory Lectures on Convex Optimization: A Basic Course
Yu. Nesterov · 2004
Earlier work this paper cites.
A convergent incremental gradient method with a constant step size
D. Blatt, A. O. Hero, and H. Gauchman · 2007
Earlier work this paper cites.
Sido: A phamacology dataset
I. Guyon · 2008
Earlier work this paper cites.
A fast iterative shrinkage-threshold algorithm for linear inverse problems
A. Beck and M. Teboulle · 2009
Earlier work this paper cites.
Efficient online and batch learning using forward backward splitting
J. Duchi and Y. Singer · 2009
Earlier work this paper cites.
Accelerated gradient methods for stochastic optimization and online learning
C. Hu, J. T. Kwok, and W. Pan · 2009
Cited alongside, same era.
The Elements of Statistical Learning: Data Mining, Inference, and Prediction
T. Hastie, R. Tibshirani, and J. Friedman · 2009
Cited alongside, same era.
Sparse online learning via truncated gradient
J. Langford, L. Li, and T. Zhang · 2009
Cited alongside, same era.
Incremental gradient, subgradient, and proximal methods for convex optimization: a survey
D. P. Bertsekas · 2010
Cited alongside, same era.
Dual averaging methods for regularized stochastic learning and online optimization
L. Xiao · 2010
Cited alongside, same era.
Incremental proximal methods for large scale convex optimization
D. P. Bertsekas · 2011
Cited alongside, same era.
Proximal stochatic dual coordinate ascent
S. Shalev-Shwartz and T. Zhang · 2012
Later among the works it cites.
Covertype data set
J. A. Blackard, D. J. Dean, and C. W. Anderson · 2013
Later among the works it cites.
Tail bounds for stochastic approximation
M. P. Friedlander and G. Goh · 2013
Later among the works it cites.
Accelerating stochastic gradient descent using predictive variance reduction
R. Johnson and T. Zhang · 2013
Later among the works it cites.
Semi-stochastic gradient descent methods
J. Konečný and P. Richtárik · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
LIBSVM data: Classification, regression and multi-label
R.-E. Fan and C.-J. Lin · 2011
Cited alongside, same era.
Sample size selection in optimization methods for machine learning
R. H. Byrd, G. M. Chin, J. Nocedal, and Y. Wu · 2012
Cited alongside, same era.
Hybrid deterministic-stochastic methods for data fitting
M. P. Friedlander and M. Schmidt · 2012
Cited alongside, same era.
A stochastic gradient method with an exponential convergence rate for finite training sets
N. Le Roux, M. Schmidt, and F. Bach · 2012
Cited alongside, same era.
M. Mahdavi, L. Zhang, and R. Jin · 2013
Later among the works it cites.
Gradient methods for minimizing composite functions
Yu. Nesterov · 2013
Later among the works it cites.
Minimizing finite sums with the stochastic average gradient
M. Schmidt, N. Le Roux, and F. Bach · 2013
Later among the works it cites.
Stochastic dual coordinate ascent methods for regularized loss minimization
S. Shalev-Shwartz and T. Zhang · 2013
Later among the works it cites.
Linear convergence with condition number independent access of full gradients
L. Zhang, M. Mahdavi, and R. Jin · 2013
Later among the works it cites.