Fetching the paper…
Reading the bibliography…
Since AlphaGo and AlphaGo Zero have achieved breakground successes in the game of Go, the programs have been generalized to solve other tasks.
1903
Earlier work this paper cites.
Iwata S, Kasai T. The Othello game on an n × \times n board is PSPACE-complete. Theoretical Computer Science. 123
1994
Earlier work this paper cites.
Heinz E A: New self-play results in computer chess. International Conference on Computers and Games. Springer, Berlin, Heidelberg. pp. 262–276 (2000)
2000
Earlier work this paper cites.
Albers P C H, Vries H. Elo-rating as a tool in the sequential estimation of dominance strengths. Animal Behaviour 489–495 (2001)
2001
Earlier work this paper cites.
Schneider M O, Rosa J L G: Neural connect 4-A connectionist approach to the game. Neural Networks SBRN 2002. Proceedings. VII Brazilian Symposium on. IEEE pp. 236–241 (2002)
2002
Earlier work this paper cites.
Birattari M, Stützle T, Paquete L, et al. A racing algorithm for configuring metaheuristics. Proceedings of the 4th Annual Conference on Genetic and Evolutionary Computation. Morgan Kaufmann Publishers Inc. 11-18 (2002)
2002
Earlier work this paper cites.
Coulom R. Whole-history rating: A Bayesian rating system for players of time-varying strength. International Conference on Computers and Games. Springer, Berlin, Heidelberg, 113–124, 2008
2008
Earlier work this paper cites.
Wiering M A: Self-Play and Using an Expert to Learn to Play Backgammon with Temporal Difference Learning. Journal of Intelligent Learning Systems and Applications 2
2010
Earlier work this paper cites.
Browne C B, Powley E, Whitehouse D, et al: A survey of monte carlo tree search methods. IEEE Transactions on Computational Intelligence and AI in games 4
2012
Earlier work this paper cites.
Van Der Ree M, Wiering M: Reinforcement learning in the game of Othello: Learning against a fixed opponent and learning from self-play. In Adaptive Dynamic Programming And Reinforcement Learning. pp. 108–115 (2013)
2013
Cited alongside, same era.
B Ruijl, J Vermaseren, A Plaat, J Herik: Combining Simulated Annealing and Monte Carlo Tree Search for Expression Simplification. In: Béatrice Duval, H. Jaap van den Herik, Stéphane Loiseau, Joaquim Filipe. Proceedings of the 6th International Conference on Agents and Artificial Intelligence 2014, vol. 1, pp. 724–731. SciTePress, Setúbal, Portugal (2014)
2014
Cited alongside, same era.
Kingma D P, Ba J: Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)
2014
Cited alongside, same era.
Srivastava N, Hinton G, Krizhevsky A, et al: Dropout: a simple way to prevent neural networks from overfitting. The Journal of Machine Learning Research. 15
2014
Wang F Y, Zhang J J, Zheng X, et al: Where does AlphaGo go: From church-turing thesis to AlphaGo thesis and beyon. IEEE/CAA Journal of Automatica Sinica 3
2016
Later among the works it cites.
Fu M C: AlphaGo and Monte Carlo tree search: the simulation optimization perspective. Proceedings of the 2016 Winter Simulation Conference. IEEE Press pp. 659–670 (2016)
2016
Later among the works it cites.
Tao J, Wu L, Hu X: Principle Analysis on AlphaGo and Perspective in Military Application of Artificial Intelligence. Journal of Command and Control 2
2016
Later among the works it cites.
Zhang Z: When doctors meet with AlphaGo: potential application of machine learning to clinical medicine. Annals of translational medicine 4
2016
Later among the works it cites.
Silver D, Schrittwieser J, Simonyan K, et al: Mastering the game of go without human knowledge. Nature 550
2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Schmidhuber J: Deep learning in neural networks: An overview. Neural networks 61
2015
Cited alongside, same era.
Clark C, Storkey A. Training deep convolutional neural networks to play go. International Conference on Machine Learning. pp. 1766–1774 (2015)
2015
Cited alongside, same era.
Mnih V, Kavukcuoglu K, Silver D, et al: Human-level control through deep reinforcement learning. Nature 518
2015
Cited alongside, same era.
Ioffe S, Szegedy C: Batch normalization: accelerating deep network training by reducing internal covariate shift. Proceedings of the 32nd International Conference on International Conference on Machine Learning-Volume 37. pp. 448–456 (2015)
2015
Cited alongside, same era.
Silver D, Huang A, Maddison C J, et al: Mastering the game of Go with deep neural networks and tree search. Nature 529
2016
Cited alongside, same era.
Surag Nair, https://github.com/suragnair/alpha-zero-general
Cited in the paper.
Later among the works it cites.
2017
Later among the works it cites.
Granter S R, Beck A H, Papke Jr D J: AlphaGo, deep learning, and the future of the human microscopist. Archives of pathology & \& laboratory medicine 141
2017
Later among the works it cites.
2018
Later among the works it cites.