Multi-armed bandit, dynamic environments and meta-bandits, 2006
C Hartland, S Gelly, N Baskiotis, O Teytaud, and M Sebag · 2006
Cited alongside, same era.
Adapting to a Stochastically Changing Environment: The Dynamic Multi-Armed Bandits Problem, 2007
Aleksandrs Slivkins and Eli Upfal · 2007
Cited alongside, same era.
Adapting to a changing environment: The Brownian restless bandits
A. Slivkins and E. Upfal · 2008
Cited alongside, same era.
UCB revisited: Improved regret bounds for the stochastic multi-armed bandit problem
Peter Auer and Ronald Ortner · 2010
Cited alongside, same era.
A modern Bayesian look at the multi-armed bandit
Steven L. Scott · 2010
Cited alongside, same era.
An Empirical Evaluation of Thompson Sampling
Olivier Chapelle and Lihong Li · 2011
Cited alongside, same era.
The KL-UCB Algorithm for Bounded Stochastic Bandits and Beyond
Aurelien Garivier and Olivier Cappe · 2011
Cited alongside, same era.
On Upper-Confidence Bound Policies for Non-Stationary Bandit Problems
Aurelien Garivier and Eric Moulines · 2011
Cited alongside, same era.
Thompson sampling for dynamic multi-armed bandits
Neha Gupta, Ole Christoffer Granmo, and Ashok Agrawala · 2011
Cited alongside, same era.
Analysis of Thompson Sampling for the multi-armed bandit problem
Shipra Agrawal and Navin Goyal · 2012
Cited alongside, same era.
On Bayesian upper confidence bounds for bandit problems
Emilie Kaufmann, Olivier Cappé, and Aurélien Garivier · 2012
Cited alongside, same era.