Fetching the paper…
Reading the bibliography…
In the Best-$K$ identification problem (Best-$K$-Arm), we are given $N$ stochastic bandit arms with unknown reward distributions.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer · 2002
Earlier work this paper cites.
The sample complexity of exploration in the multi-armed bandit problem
Shie Mannor and John N Tsitsiklis · 2004
Earlier work this paper cites.
Action elimination and stopping conditions for the multi-armed bandit and reinforcement learning problems
Eyal Even-Dar, Shie Mannor, and Yishay Mansour · 2006
Earlier work this paper cites.
Best arm identification in multi-armed bandits
Jean-Yves Audibert and Sébastien Bubeck · 2010
Earlier work this paper cites.
Efficient selection of multiple bandit arms: Theory and practice
Shivaram Kalyanakrishnan and Peter Stone · 2010
Earlier work this paper cites.
Multi-bandit best arm identification
Victor Gabillon, Mohammad Ghavamzadeh, Alessandro Lazaric, and Sébastien Bubeck · 2011
Earlier work this paper cites.
Regret analysis of stochastic and nonstochastic multi-armed bandit problems
Sébastien Bubeck, Nicolo Cesa-Bianchi, et al · 2012
Earlier work this paper cites.
Best arm identification: A unified approach to fixed budget and fixed confidence
Victor Gabillon, Mohammad Ghavamzadeh, and Alessandro Lazaric · 2012
Earlier work this paper cites.
Pac subset selection in stochastic multi-armed bandits
Shivaram Kalyanakrishnan, Ambuj Tewari, Peter Auer, and Peter Stone · 2012
Earlier work this paper cites.
Multiple identifications in multi-armed bandits
Sébastien Bubeck, Tengyao Wang, and Nitin Viswanathan · 2013
Cited alongside, same era.
Information complexity in bandit subset selection
Emilie Kaufmann and Shivaram Kalyanakrishnan · 2013
Cited alongside, same era.
Almost optimal exploration in multi-armed bandits
Zohar Karnin, Tomer Koren, and Oren Somekh · 2013
Cited alongside, same era.
Combinatorial pure exploration of multi-armed bandits
Shouyuan Chen, Tian Lin, Irwin King, Michael R Lyu, and Wei Chen · 2014
Cited alongside, same era.
lil’ucb: An optimal exploration algorithm for multi-armed bandits
Kevin Jamieson, Matthew Malloy, Robert Nowak, and Sébastien Bubeck · 2014
Cited alongside, same era.
Best-arm identification algorithms for multi-armed bandits in the fixed confidence setting
Kevin Jamieson and Robert Nowak · 2014
Cited alongside, same era.
On the complexity of best arm identification in multi-armed bandit models
Emilie Kaufmann, Olivier Cappé, and Aurélien Garivier · 2015
Later among the works it cites.
Pure exploration of multi-armed bandit under matroid constraints
Lijie Chen, Anupam Gupta, and Jian Li · 2016
Later among the works it cites.
Tight (lower) bounds for the fixed budget best arm identification bandit problem
Alexandra Carpentier and Andrea Locatelli · 2016
Later among the works it cites.
Towards instance optimal bounds for best arm identification
Lijie Chen, Jian Li, and Mingda Qiao · 2016
Later among the works it cites.
Optimal best arm identification with fixed confidence
Aurélien Garivier and Emilie Kaufmann · 2016
Later among the works it cites.
Improved learning complexity in combinatorial pure exploration bandits
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Optimal pac multiple arm identification with applications to crowdsourcing
Yuan Zhou, Xi Chen, and Jian Li · 2014
Cited alongside, same era.
On the optimal sample complexity for best arm identification
Lijie Chen and Jian Li · 2015
Cited alongside, same era.
On top-k selection in multi-armed bandits and hidden bipartite graphs
Wei Cao, Jian Li, Yufei Tao, and Zhize Li · 2015
Cited alongside, same era.
Victor Gabillon, Alessandro Lazaric, Mohammad Ghavamzadeh, Ronald Ortner, and Peter Bartlett · 2016
Later among the works it cites.
Nearly instance optimal sample complexity bounds for top-k arm selection
Lijie Chen, Jian Li, and Mingda Qiao · 2017
Closest in time.
The simulator: Understanding adaptive sampling in the moderate-confidence regime
Max Simchowitz, Kevin Jamieson, and Benjamin Recht · 2017
Closest in time.