Fetching the paper…
Reading the bibliography…
An almost-perfect chess playing agent has been a long standing challenge in the field of Artificial Intelligence.
Zur theorie der gesellschaftsspiele
John von Neumann · 1928
Earlier work this paper cites.
Xxii. programming a computer for playing chess
Claude E. Shannon · 1950
Earlier work this paper cites.
Computer science as empirical inquiry: Symbols and search
Allen Newell and Herbert A. Simon · 1976
Earlier work this paper cites.
Computer chess bad—-human chess worse
Fred Hapgood · 1982
Earlier work this paper cites.
Deep blue system overview
Feng-hsiung Hsu, Murray S. Campbell, and A. Joseph Hoane, Jr · 1995
Earlier work this paper cites.
Behind Deep Blue: Building the Computer That Defeated the World Chess Champion
Feng-Hsiung Hsu · 2002
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
Peter Auer · 2003
Earlier work this paper cites.
Bandit based monte-carlo planning
Levente Kocsis and Csaba Szepesvári · 2006
Earlier work this paper cites.
Bandit algorithms for tree search
Pierre-Arnaud Coquelin and Rémi Munos · 2007
Earlier work this paper cites.
https://chessprogramming.wikispaces.com/Centipawns , 2009
Centi pawn scale - chess programming wiki · 2009
Earlier work this paper cites.
Multi-armed bandits with episode context
Christopher D. Rosin · 2011
Cited alongside, same era.
https://en.chessbase.com/post/reconstructing-turing-s-paper-machine , 2012
Turing’s paper machine - garry kasparov · 2012
Cited alongside, same era.
Classification-based approximate policy iteration: Experiments and extended discussions
Amir-massoud Farahmand, Doina Precup, André da Motta Salles Barreto, and Mohammad Ghavamzadeh · 2014
Cited alongside, same era.
http://stockfishchess.org , 2015
Stockfish - an open source chess engine · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeffrey Dean · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Searching for activation functions
Prajit Ramachandran, Barret Zoph, and Quoc V. Le · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy P. Lillicrap, Karen Simonyan, and Demis Hassabis · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Later among the works it cites.
One pixel attack for fooling deep neural networks
Jiawei Su, Danilo Vasconcellos Vargas, and Sakurai Kouichi · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Giraffe: Using deep reinforcement learning to play chess
Matthew Lai · 2015
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Cited alongside, same era.
Thinking fast and slow with deep learning and tree search
Thomas Anthony, Zheng Tian, and David Barber · 2017
Cited alongside, same era.
On fair comparision between Stockfish and AlphaZero by tord romstad · 2018
Closest in time.
Learning to search with mctsnets
Arthur Guez, Théophane Weber, Ioannis Antonoglou, Karen Simonyan, Oriol Vinyals, Daan Wierstra, Rémi Munos, and David Silver · 2018
Closest in time.
Memory augmented control networks
Arbaaz Khan, Clark Zhang, Nikolay Atanasov, Konstantinos Karydis, Vijay Kumar, and Daniel D. Lee · 2018
Closest in time.
On the convergence of adam and beyond
Sashank J. Reddi, Satyen Kale, and Sanjiv Kumar · 2018
Closest in time.
Elf opengo
Yuandong Tian, Jerry Ma*, Qucheng Gong*, Shubho Sengupta, Zhuoyuan Chen, and C. Lawrence Zitnick · 2018
Closest in time.