Fetching the paper…
Reading the bibliography…
In this paper, we consider the convex and non-convex composition problem with the structure $\frac{1}{n}\sum\nolimits_{i = 1}^n {{F_i}( {G( x )} )}$, where $G( x )=\frac{1}{n}\sum\nolimits_{j = 1}^n {{G_j}( x )} $ is the inner function, and $F_i(\cdot)$ is the outer function.
Reinforcement learning: An introduction
Richard S Sutton, Andrew G Barto, et al · 1998
Earlier work this paper cites.
Stochastic neighbor embedding
Geoffrey E Hinton and Sam T Roweis · 2003
Earlier work this paper cites.
Learning nonlinear image manifolds by global alignment of local linear models
Jakob Verbeek · 2006
Earlier work this paper cites.
Alternating direction method of multipliers
Stephen Boyd · 2011
Earlier work this paper cites.
Spherical stochastic neighbor embedding of hyperspectral data
Dalton Lunga and Okan Ersoy · 2013
Earlier work this paper cites.
Accelerating stochastic gradient descent using predictive variance reduction
Rie Johnson and Tong Zhang · 2013
Earlier work this paper cites.
Stochastic dual coordinate ascent methods for regularized loss minimization
Shai Shalev-Shwartz and Tong Zhang · 2013
Earlier work this paper cites.
Introductory lectures on convex optimization: A basic course
Yurii Nesterov · 2013
Earlier work this paper cites.
Accelerated proximal stochastic dual coordinate ascent for regularized loss minimization
Shai Shalev-Shwartz and Tong Zhang · 2014
Earlier work this paper cites.
An accelerated proximal coordinate gradient method
Qihang Lin, Zhaosong Lu, and Lin Xiao · 2014
Earlier work this paper cites.
A proximal stochastic gradient method with progressive variance reduction
Lin Xiao and Tong Zhang · 2014
Cited alongside, same era.
Learning the information divergence
Onur Dikmen, Zhirong Yang, and Erkki Oja · 2015
Cited alongside, same era.
Silhouette analysis for human action recognition based on supervised temporal t-sne and incremental learning
Jian Cheng, Haijun Liu, Feng Wang, Hongsheng Li, and Ce Zhu · 2015
Cited alongside, same era.
Hashing on nonlinear manifolds
Fumin Shen, Chunhua Shen, Qinfeng Shi, Anton van den Hengel, Zhenmin Tang, and Heng Tao Shen · 2015
Cited alongside, same era.
An accelerated randomized proximal coordinate gradient method and its application to regularized empirical risk minimization
Qihang Lin, Zhaosong Lu, and Lin Xiao · 2015
Cited alongside, same era.
Finite-sum composition optimization via variance reduced gradient descent
Xiangru Lian, Mengdi Wang, and Ji Liu · 2017
Later among the works it cites.
Duality-free methods for stochastic composition optimization
Liu Liu, Ji Liu, and Dacheng Tao · 2017
Later among the works it cites.
Variance reduced methods for non-convex composition optimization
Liu Liu, Ji Liu, and Dacheng Tao · 2017
Later among the works it cites.
Less than a single pass: Stochastically controlled stochastic gradient
Lihua Lei and Michael Jordan · 2017
Later among the works it cites.
Non-convex finite-sum optimization via scsg methods
Lihua Lei, Cheng Ju, Jianbo Chen, and Michael I Jordan · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ji Liu, Mengdi Wang, and Ethan Fang · 2016
Cited alongside, same era.
Accelerated gradient methods for nonconvex nonlinear and stochastic programming
Saeed Ghadimi and Guanghui Lan · 2016
Cited alongside, same era.
Stochastic variance reduction for nonconvex optimization
Sashank J Reddi, Ahmed Hefny, Suvrit Sra, Barnabas Poczos, and Alex Smola · 2016
Cited alongside, same era.
Stochastic compositional gradient descent: algorithms for minimizing compositions of expected-value functions
Mengdi Wang, Ethan X Fang, and Han Liu · 2017
Cited alongside, same era.
Zeyuan Allen-Zhu · 2017
Later among the works it cites.
Katyusha: The first direct acceleration of stochastic gradient methods
Zeyuan Allen-Zhu · 2017
Later among the works it cites.
Fast stochastic variance reduced admm for stochastic composition optimization
Yue Yu and Longbo Huang · 2017
Later among the works it cites.
Stochastic zeroth-order optimization via variance reduction method
Liu Liu, Minhao Cheng, Cho-Jui Hsieh, and Dacheng Tao · 2018
Closest in time.