On kernelized multi-armed bandits
Chowdhury, S. R · 2017
Later among the works it cites.
SGD learns the conjugate kernel class of the network
Daniely, A · 2017
Later among the works it cites.
Unifying PAC and regret: Uniform PAC bounds for episodic reinforcement learning
Dann, C · 2017
Later among the works it cites.
Contextual decision processes with low bellman rank are pac-learnable
Jiang, N · 2017
Later among the works it cites.
Mastering the game of Go without human knowledge
Silver, D · 2017
Later among the works it cites.
Efficient reinforcement learning in deterministic systems with value function generalization
Wen, Z · 2017
Later among the works it cites.
Frequentist coverage and sup-norm convergence rate in gaussian process regression
Original
Yang, Y · 2017
Later among the works it cites.
A note on lazy training in supervised differentiable programming
Original
Chizat, L · 2018
Later among the works it cites.
On oracle-efficient PAC RL with rich observations
Dann, C · 2018
Later among the works it cites.
Streaming kernel regression with provably adaptive mean, variance, and regularization
Durand, A · 2018
Later among the works it cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A · 2018
Later among the works it cites.
Is Q-learning provably efficient?
Jin, C · 2018
Later among the works it cites.
Learning overparameterized neural networks via stochastic gradient descent on structured data
Li, Y · 2018
Later among the works it cites.
Randomized prior functions for deep reinforcement learning
Osband, I · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
Sutton, R. S · 2018
Later among the works it cites.
High-Dimensional Probability: An Introduction with Applications in Data Science
Vershynin, R · 2018
Later among the works it cites.
Deep reinforcement learning for NLP
Wang, W. Y · 2018
Later among the works it cites.
How SGD selects the global minima in over-parameterized learning: A dynamical stability perspective
Wu, L · 2018
Later among the works it cites.
Stochastic gradient descent optimizes over-parameterized deep ReLU networks
Original
Zou, D · 2018
Later among the works it cites.
Convergence of adversarial training in overparametrized neural networks
Gao, R · 2019
Later among the works it cites.
Bandit Algorithms
Lattimore, T · 2019
Later among the works it cites.
Towards understanding the role of over-parametrization in generalization of neural networks
Neyshabur, B · 2019
Later among the works it cites.
Worst-case regret bounds for exploration via randomized value functions
Russo, D · 2019
Later among the works it cites.
No-regret learning in unknown games with correlated payoffs
Sessa, P. G · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Vinyals, O · 2019
Later among the works it cites.