Fetching the paper…
Reading the bibliography…
In this paper, we obtain the Berry-Esseen bound for multivariate normal approximation for the Polyak-Ruppert averaged iterates of the linear stochastic approximation (LSA) algorithm with decreasing step size.
Fourier analysis of distribution functions. A mathematical study of the Laplace-Gaussian law
Carl-Gustav Esseen · 1945
Earlier work this paper cites.
Sums of Independent Random Variables
V. Petrov · 1975
Earlier work this paper cites.
The bayesian bootstrap
Donald B Rubin · 1981
Earlier work this paper cites.
Exact Convergence Rates in Some Martingale Central Limit Theorems
E. Bolthausen · 1982
Earlier work this paper cites.
Quadratic mean and almost-sure convergence of unbounded stochastic approximation algorithms with correlated observations
E. Eweda and O. Macchi · 1983
Earlier work this paper cites.
Problem complexity and method efficiency in optimization
Arkadij Semenovič Nemirovskij and David Borisovich Yudin · 1983
Earlier work this paper cites.
Vladimir Mikhailovich Zolotarev · 1984
Earlier work this paper cites.
Efficient estimations from a slowly convergent robbins-monro process
David Ruppert · 1988
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
R. S Sutton · 1988
Earlier work this paper cites.
Bootstrap methods: another look at the jackknife
Bradley Efron · 1992
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
Boris T Polyak and Anatoli B Juditsky · 1992
Earlier work this paper cites.
On the generation of markov decision processes
TW Archibald, KIM McKinnon, and LC Thomas · 1995
Earlier work this paper cites.
Weak convergence and empirical processes
A. W. Van Der Vaart and J. A. Wellner · 1996
Earlier work this paper cites.
An analysis of temporal-difference learning with function approximation
J. N. Tsitsiklis and B. Van Roy · 1997
Earlier work this paper cites.
On a perturbation approach for the analysis of stochastic tracking algorithms
Rafik Aguech, Eric Moulines, and Pierre Priouret · 2000
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer · 2002
Earlier work this paper cites.
Stochastic approximation and recursive algorithms and applications
Harold Kushner and G George Yin · 2003
Earlier work this paper cites.
Mathematical statistics
Jun Shao · 2003
Earlier work this paper cites.
Stochastic Approximation: A Dynamical Systems Viewpoint
Vivek S Borkar · 2008
Earlier work this paper cites.
Advanced Mathematical Tools for Automatic Control Engineers: Deterministic Techniques
A. S. Poznyak · 2008
Earlier work this paper cites.
Robust stochastic approximation approach to stochastic programming
Arkadi Nemirovski, Anatoli Juditsky, Guanghui Lan, and Alexander Shapiro · 2009
Earlier work this paper cites.
Non-asymptotic analysis of stochastic approximation algorithms for machine learning
Eric Moulines and Francis Bach · 2011
Earlier work this paper cites.
Fundamentals of Stein’s method
Nathan Ross · 2011
Earlier work this paper cites.
Freedman’s inequality for matrix martingales
Joel A. Tropp · 2011
Earlier work this paper cites.
Adaptive algorithms and stochastic approximations
A. Benveniste, M. Métivier, and P. Priouret · 2012
Earlier work this paper cites.
Ergodic mirror descent
John C Duchi, Alekh Agarwal, Mikael Johansson, and Michael I Jordan · 2012
Cited alongside, same era.
An optimal method for stochastic composite optimization
Guanghui Lan · 2012
Cited alongside, same era.
Sharp Martingale and Semimartingale Inequalities
A. Osekowski · 2012
Cited alongside, same era.
Making gradient descent optimal for strongly convex stochastic optimization
Alexander Rakhlin, Ohad Shamir, and Karthik Sridharan · 2012
Cited alongside, same era.
Non-strongly-convex smooth stochastic approximation with convergence rate o(1/n)
F. Bach and E. Moulines · 2013
Cited alongside, same era.
Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors
Victor Chernozhukov, Denis Chetverikov, and Kengo Kato · 2013
Cited alongside, same era.
Bootstrap confidence sets for spectral projectors of sample covariance
Alexey Naumov, Vladimir Spokoiny, and Vladimir Ulyanov · 2019
Later among the works it cites.
Statistical inference for model parameters in stochastic gradient descent
Xi Chen, Jason D. Lee, Xin T. Tong, and Yichen Zhang · 2020
Later among the works it cites.
On linear stochastic approximation: Fine-grained Polyak-Ruppert and non-asymptotic concentration
Wenlong Mou, Chris Junchi Li, Martin J Wainwright, Peter L Bartlett, and Michael I Jordan · 2020
Later among the works it cites.
Tight high probability bounds for linear stochastic approximation with fixed stepsize
Alain Durmus, Eric Moulines, Alexey Naumov, Sergey Samsonov, Kevin Scaman, and Hoi-To Wai · 2021
Later among the works it cites.
On the stability of random matrix product with markovian noise: Application to linear stochastic approximation and td learning
Alain Durmus, Eric Moulines, Alexey Naumov, Sergey Samsonov, and Hoi-To Wai · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The nature of statistical learning theory
Vladimir Vapnik · 2013
Cited alongside, same era.
Off-policy learning with eligibility traces: a survey
Matthieu Geist, Bruno Scherrer, et al · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Learning to optimize via posterior sampling
Daniel Russo and Benjamin Van Roy · 2014
Cited alongside, same era.
Central limit theorems for stochastic approximation with controlled Markov chain dynamics
G. Fort · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Cited alongside, same era.
Matrix concentration for products
De Huang, Jonathan Niles-Weed, Joel A Tropp, and Rachel Ward · 2021
Later among the works it cites.
Hpc resources of the higher school of economics
PS Kostenetskiy, RA Chulkevich, and VI Kozyrev · 2021
Later among the works it cites.
Statistical estimation and online inference via local sgd
Xiang Li, Jiadong Liang, Xiangyu Chang, and Zhihua Zhang · 2022
Later among the works it cites.
Multivariate normal approximation on the wiener space: new bounds in the convex distance
Ivan Nourdin, Giovanni Peccati, and Xiaochuan Yang · 2022
Later among the works it cites.
Berry–Esseen bounds for multivariate nonlinear statistics with applications to M-estimators and stochastic gradient descent algorithms
Qi-Man Shao and Zhuo-Song Zhang · 2022
Later among the works it cites.
Bounding Kolmogorov distances through Wasserstein and related integral probability metrics
Robert E Gaunt and Siqi Li · 2023
Later among the works it cites.
Bias and extrapolation in markovian linear stochastic approximation with constant stepsizes
Dongyan Huo, Yudong Chen, and Qiaomin Xie · 2023
Later among the works it cites.
Online statistical inference for nonlinear stochastic approximation with Markovian data
Xiang Li, Jiadong Liang, and Zhihua Zhang · 2023
Later among the works it cites.
A statistical analysis of Polyak-Ruppert averaged Q-learning
Xiang Li, Wenhao Yang, Jiadong Liang, Zhihua Zhang, and Michael I Jordan · 2023
Later among the works it cites.
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
Gandharv Patil, LA Prashanth, Dheeraj Nagaraj, and Doina Precup · 2023
Later among the works it cites.
Online bootstrap inference for policy evaluation in reinforcement learning
Pratik Ramprasad, Yuantong Li, Zhuoran Yang, Zhaoran Wang, Will Wei Sun, and Guang Cheng · 2023
Later among the works it cites.
Online Covariance Matrix Estimation in Stochastic Gradient Descent
Xi Chen Wanrong Zhu and Wei Biao Wu · 2023
Later among the works it cites.
Online Bootstrap Inference with Nonconvex Stochastic Gradient Descent Estimator
Yanjie Zhong, Todd Kuffner, and Soumendra Lahiri · 2023
Later among the works it cites.
Finite-time high-probability bounds for Polyak-Ruppert averaged iterates of linear stochastic approximation
Alain Durmus, Eric Moulines, Alexey Naumov, and Sergey Samsonov · 2024
Closest in time.
Quantitative limit theorems and bootstrap approximations for empirical spectral projectors
Moritz Jirak and Martin Wahl · 2024
Closest in time.
Fast inference for quantile regression with tens of millions of observations
Sokbae Lee, Yuan Liao, Myung Hwan Seo, and Youngki Shin · 2024
Closest in time.
High-probability sample complexities for policy evaluation with linear function approximation
Gen Li, Weichen Wu, Yuejie Chi, Cong Ma, Alessandro Rinaldo, and Yuting Wei · 2024
Closest in time.
Optimal and instance-dependent guarantees for markovian linear stochastic approximation
Wenlong Mou, Ashwin Pananjady, Martin J Wainwright, and Peter L Bartlett · 2024
Closest in time.
Improved High-Probability Bounds for the Temporal Difference Learning Algorithm via Exponential Stability
Sergey Samsonov, Daniil Tiapkin, Alexey Naumov, and Eric Moulines · 2024
Closest in time.
R Srikant · 2024
Closest in time.