Fetching the paper…
Reading the bibliography…
Monte Carlo Tree Search (MCTS) efficiently balances exploration and exploitation in tree search based on count-derived uncertainty.
An analysis of alpha-beta pruning
Donald E Knuth and Ronald W Moore · 1975
Earlier work this paper cites.
The B* tree search algorithm: A best-first proof procedure
Hans Berliner · 1981
Earlier work this paper cites.
Exploiting graph properties of game trees
Aske Plaat, Jonathan Schaeffer, Wim Pijls, and Arie De Bruin · 1996
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer · 2002
Earlier work this paper cites.
Efficient selectivity and backup operators in Monte-Carlo tree search
Rémi Coulom · 2006
Earlier work this paper cites.
Bandit based monte-carlo planning
Levente Kocsis and Csaba Szepesvári · 2006
Earlier work this paper cites.
On the parallelization of UCT
Tristan Cazenave and Nicolas Jouandeau · 2007
Earlier work this paper cites.
Monte-Carlo Tree Search: A New Framework for Game AI
Guillaume Chaslot, Sander Bakkes, Istvan Szita, and Pieter Spronck · 2008
Earlier work this paper cites.
Monte-Carlo tree search solver
Mark HM Winands, Yngvi Björnsson, and Jahn-Takeshi Saito · 2008
Earlier work this paper cites.
Best arm identification in multi-armed bandits
Jean-Yves Audibert and Sébastien Bubeck · 2010
Cited alongside, same era.
Score bounded Monte-Carlo tree search
Tristan Cazenave and Abdallah Saffidine · 2010
Cited alongside, same era.
Time management for Monte-Carlo tree search applied to the game of Go
Shih-Chieh Huang, Remi Coulom, and Shun-Shii Lin · 2010
Cited alongside, same era.
Multi-armed bandits with episode context
Christopher D Rosin · 2011
Cited alongside, same era.
A survey of monte carlo tree search methods
Cameron B Browne, Edward Powley, Daniel Whitehouse, Simon M Lucas, Peter I Cowling, Philipp Rohlfshagen, Stephen Tavener, Diego Perez, Spyridon Samothrakis, and Simon Colton · 2012
Cited alongside, same era.
Bayesian inference in monte-carlo tree search
Gerald Tesauro, VT Rajan, and Richard Segal · 2012
Vizdoom: A doom-based ai research platform for visual reinforcement learning
Michał Kempka, Marek Wydmuch, Grzegorz Runc, Jakub Toczek, and Wojciech Jaśkowski · 2016
Later among the works it cites.
Generalization and Exploration via Randomized Value Functions
Ian Osband, Benjamin Van Roy, and Zheng Wen · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
Monte-carlo tree search by best arm identification
Emilie Kaufmann and Wouter M Koolen · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
From bandits to Monte-Carlo Tree Search: The optimistic principle applied to optimization and planning
Rémi Munos et al · 2014
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Thomas M Moerland, Joost Broekens, and Catholijn M Jonker · 2018
Later among the works it cites.
Reinforcement learning: An Introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Exploration by distributional reinforcement learning
Yunhao Tang and Shipra Agrawal · 2018
Later among the works it cites.