Fetching the paper…
Reading the bibliography…
We undertake a precise study of the asymptotic and non-asymptotic properties of stochastic approximation procedures with Polyak-Ruppert averaging for solving a linear system $\bar{A} \theta = \bar{b}$.
A stochastic approximation method
Herbert Robbins and Sutton Monro · 1951
Earlier work this paper cites.
Monotone operators associated with saddle-functions and minimax problems
R Tyrrell Rockafellar · 1970
Earlier work this paper cites.
Integral inequalities for convex functions of operators on martingales
Donald L Burkholder, Burgess J Davis, and Richard F Gundy · 1972
Earlier work this paper cites.
On tail probabilities for martingales
David A Freedman et al · 1975
Earlier work this paper cites.
Martingale Limit Theory and Its Application
Peter Hall and Christopher C Heyde · 1980
Earlier work this paper cites.
Problem Complexity and Method Efficiency in Optimization
Arkadii Nemirovskii and David Borisovich Yudin · 1983
Earlier work this paper cites.
Efficient estimations from a slowly convergent robbins-monro process
David Ruppert · 1988
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Richard S Sutton · 1988
Earlier work this paper cites.
Parallel and Distributed Computation: Numerical Methods
Dimitris P Bertsekas and John N Tsitsiklis · 1989
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
Boris T Polyak and Anatoli B Juditsky · 1992
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Dynamic Programming and Stochastic Control
Dimitris P. Bertsekas · 1995
Earlier work this paper cites.
On average versus discounted reward temporal-difference learning
John N Tsitsiklis and Benjamin Van Roy · 2002
Earlier work this paper cites.
Stochastic Approximation and Recursive Algorithms and Applications
Harold Kushner and G George Yin · 2003
Earlier work this paper cites.
Stochastic approximation
Tze Leung Lai · 2003
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Mark L. Puterman · 2005
Earlier work this paper cites.
Foundations of Modern Probability
Olav Kallenberg · 2006
Earlier work this paper cites.
Stochastic Approximation: A Dynamical Systems Viewpoint
Vivek S Borkar · 2008
Earlier work this paper cites.
Optimal Transport: Old and New
Cédric Villani · 2008
Cited alongside, same era.
Robust stochastic approximation approach to stochastic programming
Arkadi Nemirovski, Anatoli Juditsky, Guanghui Lan, and Alexander Shapiro · 2009
Cited alongside, same era.
Curvature, concentration and error estimates for Markov chain Monte Carlo
Aldéric Joulin and Yann Ollivier · 2010
Cited alongside, same era.
Non-asymptotic analysis of stochastic approximation algorithms for machine learning
Eric Moulines and Francis R Bach · 2011
Cited alongside, same era.
Information-theoretic lower bounds on the oracle complexity of stochastic convex optimization
Alekh Agarwal, Peter L Bartlett, Pradeep Ravikumar, and Martin J Wainwright · 2012
Cited alongside, same era.
Adaptive Algorithms and Stochastic Approximations
Albert Benveniste, Michel Métivier, and Pierre Priouret · 2012
Harder, better, faster, stronger convergence rates for least-squares regression
Aymeric Dieuleveut, Nicolas Flammarion, and Francis Bach · 2017
Later among the works it cites.
Parallelizing stochastic gradient descent for least squares regression: mini-batching, averaging, and model misspecification
Prateek Jain, Praneeth Netrapalli, Sham M Kakade, Rahul Kidambi, and Aaron Sidford · 2017
Later among the works it cites.
Stochastic gradient descent as approximate Bayesian inference
Stephan Mandt, Matthew D Hoffman, and David M Blei · 2017
Later among the works it cites.
A finite time analysis of temporal difference learning with linear function approximation
Jalaj Bhandari, Daniel Russo, and Raghav Singal · 2018
Later among the works it cites.
Statistical sparse online regression: A diffusion approximation perspective
Jianqing Fan, Wenyan Gong, Chris Junchi Li, and Qiang Sun · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Making gradient descent optimal for strongly convex stochastic optimization
Alexander Rakhlin, Ohad Shamir, and Karthik Sridharan · 2012
Cited alongside, same era.
A stochastic gradient method with an exponential convergence rate for finite training sets
Nicolas L Roux, Mark Schmidt, and Francis R Bach · 2012
Cited alongside, same era.
Accelerating stochastic gradient descent using predictive variance reduction
Rie Johnson and Tong Zhang · 2013
Cited alongside, same era.
Differential Equations and Dynamical Systems
Lawrence Perko · 2013
Cited alongside, same era.
Stochastic gradient descent for non-smooth optimization: Convergence results and optimal averaging schemes
Ohad Shamir and Tong Zhang · 2013
Cited alongside, same era.
Saga: A fast incremental gradient method with support for non-strongly convex composite objectives
Aaron Defazio, Francis Bach, and Simon Lacoste-Julien · 2014
Cited alongside, same era.
Accelerating stochastic gradient descent for least squares regression
Prateek Jain, Sham M Kakade, Rahul Kidambi, Praneeth Netrapalli, and Aaron Sidford · 2018
Later among the works it cites.
Linear stochastic approximation: How far does constant step-size and iterate averaging go?
Chandrashekar Lakshminarayanan and Csaba Szepesvari · 2018
Later among the works it cites.
Statistical inference using SGD
Tianyang Li, Liu Liu, Anastasios Kyrillidis, and Constantine Caramanis · 2018
Later among the works it cites.
Weijie J Su and Yuancheng Zhu · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto · 2018
Later among the works it cites.
Making the last iterate of SGD information theoretically optimal
Prateek Jain, Dheeraj Nagaraj, and Praneeth Netrapalli · 2019
Later among the works it cites.
Non-asymptotic analysis of biased stochastic approximation scheme
Belhal Karimi, Blazej Miasojedow, Éric Moulines, and Hoi-To Wai · 2019
Later among the works it cites.
Statistical inference for the population landscape via moment-adjusted stochastic gradients
Tengyuan Liang and Weijie J Su · 2019
Later among the works it cites.
Ashwin Pananjady and Martin J Wainwright · 2019
Later among the works it cites.
High-Dimensional Statistics: A Non-Asymptotic Viewpoint
Martin J Wainwright · 2019
Later among the works it cites.
Martin J Wainwright · 2019
Later among the works it cites.
Variance-reduced Q-learning is minimax optimal
Martin J Wainwright · 2019
Later among the works it cites.