Fetching the paper…
Reading the bibliography…
DeepMind's recent spectacular success in using deep convolutional neural nets and machine learning to build superhuman level agents --- e.g.
Computers and automata
Claude E. Shannon · 1953
Earlier work this paper cites.
Mathematical games
Martin Gardner · 1957
Earlier work this paper cites.
Hex ist PSPACE-vollständig
Stefan Reisch · 1981
Earlier work this paper cites.
Temporal difference learning and td-gammon
Gerald Tesauro · 1995
Earlier work this paper cites.
Reinforcement Learning : An Introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
The game of hex: An automatic theorem proving approach to game programming
Vadim V Anshelevich · 2000
Cited alongside, same era.
Wolve wins hex tournament
Broderick Arneson, Ryan Hayward, and Philip Henderson · 2008
Cited alongside, same era.
Monte Carlo Tree Search in Hex
Broderick Arneson, Ryan B. Hayward, and Philip Henderson · 2010
Cited alongside, same era.
Theano: a CPU and GPU math expression compiler
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio · 2010
Cited alongside, same era.
Theano: new features and speed improvements
Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, James Bergstra, Ian J. Goodfellow, Arnaud Bergeron, Nicolas Bouchard, and Yoshua Bengio · 2012
Cited alongside, same era.
Lecture 6.5—RmsProp: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Later among the works it cites.
Mohex wins Hex tournament
Ryan B. Hayward · 2013
Later among the works it cites.
Mohex 2.0: A pattern-based mcts hex player
Shih-Chieh Huang, Broderick Arneson, Ryan B. Hayward, Martin Müller, and Jakub Pawlewicz · 2014
Later among the works it cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…