Fetching the paper…
Reading the bibliography…
Competing with top human players in the ancient game of Go has been a long-term goal of artificial intelligence.
Temporal difference learning of position evaluation in the game of go
Schraudolph, Nicol N, Dayan, Peter, and Sejnowski, Terrence J · 1994
Earlier work this paper cites.
The integration of a priori knowledge into a go playing neural network
Enzenberger, Markus · 1996
Earlier work this paper cites.
Evolving neural networks to play go
Richards, Norman, Moriarty, David E, and Miikkulainen, Risto · 1998
Earlier work this paper cites.
Actor-critic algorithms
Konda, Vijay R and Tsitsiklis, John N · 1999
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Sutton, Richard S, McAllester, David A, Singh, Satinder P, Mansour, Yishay, et al · 1999
Earlier work this paper cites.
Bandit based monte-carlo planning
Kocsis, Levente and Szepesvári, Csaba · 2006
Earlier work this paper cites.
Mimicking go experts with convolutional neural networks
Sutskever, Ilya and Nair, Vinod · 2008
Cited alongside, same era.
Reinforcement learning and simulation-based search
Silver, David · 2009
Cited alongside, same era.
Fuego—an open-source framework for board games and go engine based on monte carlo tree search
Enzenberger, Markus, Müller, Martin, Arneson, Broderick, and Segal, Richard · 2010
Cited alongside, same era.
Pachi: State of the art open source go program
Baudis, Petr and Gailly, Jean-loup · 2012
Cited alongside, same era.
A survey of monte carlo tree search methods
Browne, Cameron B, Powley, Edward, Whitehouse, Daniel, Lucas, Simon M, Cowling, Peter, Rohlfshagen, Philipp, Tavener, Stephen, Perez, Diego, Samothrakis, Spyridon, Colton, Simon, et al · 2012
Cited alongside, same era.
Training deep convolutional neural networks to play go
Clark, Christopher and Storkey, Amos · 2015
Closest in time.
Adaptive playouts in monte-carlo tree search with policy-gradient reinforcement learning
Graf, Tobias and Platzner, Marco · 2015
Closest in time.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2015
Closest in time.
Move evaluation in go using deep convolutional neural networks
Maddison, Chris J, Huang, Aja, Sutskever, Ilya, and Silver, David · 2015
Closest in time.
Mastering the game of go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J., Guez, Arthur, Sifre, Laurent, van den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, Dieleman, Sander, Grewe, Dominik, Nham, John, Kalchbrenner, Nal, Sutskever, Ilya, Lillicrap, Timothy, Leach, Madeleine, Kavukcuoglu, Koray, Graepel, Thore, and Hassabis, Demis · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…