Fetching the paper…
Reading the bibliography…
The question of how to incorporate curvature information in stochastic approximation methods is challenging.
A stochastic approximation method
H. Robbins and S. Monro · 1951
Earlier work this paper cites.
Some global convergence properties of a variable metric algorithm for minimization without exact line searches
M.J.D. Powell · 1976
Earlier work this paper cites.
Practical Methods of Optimization
R. Fletcher · 1987
Earlier work this paper cites.
On the limited memory bfgs method for large scale optimization
D. C. Liu and J. Nocedal · 1989
Earlier work this paper cites.
Natural gradient works efficiently in learning
Shun-Ichi Amari · 1998
Earlier work this paper cites.
A statistical study of on-line learning
Noboru Murata · 1998
Earlier work this paper cites.
Numerical Optimization
Jorge Nocedal and Stephen Wright · 1999
Earlier work this paper cites.
Adaptive natural gradient learning algorithms for various stochastic models
Hyeyoung Park, S-I Amari, and Kenji Fukumizu · 2000
Earlier work this paper cites.
Convergence rate of incremental subgradient algorithms
Angelia Nedić and Dimitri Bertsekas · 2001
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
David D Lewis, Yiming Yang, Tony G Rose, and Fan Li · 2004
Cited alongside, same era.
A stochastic approximation algorithm with step-size adaptation
Alexander Plakhov and Pedro Cruz · 2004
Cited alongside, same era.
Stochastic simulation: Algorithms and analysis
Søren Asmussen and Peter W Glynn · 2007
Cited alongside, same era.
Topmoumoute online natural gradient algorithm
Nicolas L Roux, Pierre-Antoine Manzagol, and Yoshua Bengio · 2007
Cited alongside, same era.
A stochastic quasi-newton method for online convex optimization
Nicol Schraudolph, Jin Yu, and Simon Günter · 2007
Cited alongside, same era.
The tradeoffs of large scale learning
Leon Bottou and Olivier Bousquet · 2008
Cited alongside, same era.
A coordinate gradient descent method for nonsmooth separable minimization
P. Tseng and S. Yun · 2009
Later among the works it cites.
A fast natural Newton method
Nicolas L Roux and Andrew W Fitzgibbon · 2010
Later among the works it cites.
On the use of stochastic Hessian information in unconstrained optimization
R.H Byrd, G. M Chin, W. Neveitt, and J. Nocedal · 2011
Later among the works it cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer · 2011
Later among the works it cites.
Sample size selection in optimization methods for machine learning
R. H. Byrd, G. M. Chin, J. Nocedal, and Y. Wu · 2012
Later among the works it cites.
On stochastic gradient and subgradient methods with adaptive steplength sequences
Farzad Yousefian, Angelia Nedić, and Uday V Shanbhag · 2012
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
SGD-QN: Careful quasi-Newton stochastic gradient descent
Antoine Bordes, Léon Bottou, and Patrick Gallinari · 2009
Cited alongside, same era.
Robust stochastic approximation approach to stochastic programming
Arkadi Nemirovski, Anatoli Juditsky, Guanghui Lan, and Alexander Shapiro · 2009
Cited alongside, same era.
Variable metric stochastic approximation theory
Peter Sunehag, Jochen Trumpf, SVN Vishwanathan, and Nicol Schraudolph · 2009
Cited alongside, same era.
Later among the works it cites.
Non-strongly-convex smooth stochastic approximation with convergence rate o (1/n)
F. Bach and E. Moulines · 2013
Later among the works it cites.
Regularized stochastic BFGS algorithm, 2013
Aryan Mokhtari and Alejandro Ribeiro · 2013
Later among the works it cites.
Parallel boosting with momentum
Indraneel Mukherjee, Kevin Canini, Rafael Frongillo, and Yoram Singer · 2013
Later among the works it cites.