Finite-time error bounds for linear stochastic approximation andtd learning
Rayadurgam Srikant and Lei Ying · 2019
Later among the works it cites.
Sample-optimal parametric q-learning using linearly additive features
Lin Yang and Mengdi Wang · 2019
Later among the works it cites.
Finite-sample analysis for sarsa with linear function approximation
Shaofeng Zou, Tengyu Xu, and Yingbin Liang · 2019
Later among the works it cites.
Model-based reinforcement learning with a generative model is minimax optimal
Alekh Agarwal, Sham Kakade, and Lin F Yang · 2020
Later among the works it cites.
Least squares regression with markovian data: Fundamental limits and algorithms
Original
Guy Bresler, Prateek Jain, Dheeraj Nagaraj, Praneeth Netrapalli, and Xian Wu · 2020
Later among the works it cites.
A new convergent variant of q-learning with linear function approximation
Diogo Carvalho, Francisco S. Melo, and Pedro Santos · 2020
Later among the works it cites.
Finite-sample analysis of contractive stochastic approximation using smooth convex envelopes
Original
Zaiwei Chen, Siva Theja Maguluri, Sanjay Shakkottai, and Karthikeyan Shanmugam · 2020
Later among the works it cites.
Convergence rates of accelerated markov gradient descent with applications in reinforcement learning
Original
Thinh T Doan, Lam M Nguyen, Nhan H Pham, and Justin Romberg · 2020
Later among the works it cites.
A theoretical analysis of deep q-learning
Jianqing Fan, Zhaoran Wang, Yuchen Xie, and Zhuoran Yang · 2020
Later among the works it cites.
Provably efficient reinforcement learning with linear function approximation
Chi Jin, Zhuoran Yang, Zhaoran Wang, and Michael I Jordan · 2020
Later among the works it cites.
Sample complexity of asynchronous q-learning: Sharper analysis and variance reduction
Gen Li, Yuting Wei, Yuejie Chi, Yuantao Gu, and Yuxin Chen · 2020
Later among the works it cites.
Finite-time analysis of asynchronous stochastic approximation and q q -learning
Guannan Qu and Adam Wierman · 2020
Later among the works it cites.
Momentum q-learning with finite-sample convergence guarantee
Original
Bowen Weng, Huaqing Xiong, Lin Zhao, Yingbin Liang, and Wei Zhang · 2020
Later among the works it cites.
A finite-time analysis of q-learning with neural network function approximation
Pan Xu and Quanquan Gu · 2020
Later among the works it cites.
Learning near optimal policies with low inherent bellman error
Andrea Zanette, Alessandro Lazaric, Mykel Kochenderfer, and Emma Brunskill · 2020
Later among the works it cites.
Is q-learning minimax optimal? a tight sample complexity analysis
Original
Gen Li, Changxiao Cai, Yuxin Chen, Yuantao Gu, Yuting Wei, and Yuejie Chi · 2021
Closest in time.