Fetching the paper…
Reading the bibliography…
During the development of AlphaGo, its many hyper-parameters were tuned with Bayesian optimization multiple times.
The application of Bayesian methods for seeking the extremum
J Moćkus, V Tiesis, and A Źilinskas · 1978
Earlier work this paper cites.
Gaussian Processes for Machine Learning
Carl Edward Rasmussen and Chris Williams · 2006
Earlier work this paper cites.
Whole-history rating: A bayesian rating system for players of time-varying strength
Rémi Coulom · 2008
Earlier work this paper cites.
Eric Brochu, Vlad M Cora, and Nando De Freitas · 2010
Earlier work this paper cites.
Time management for Monte-Carlo tree search applied to the game of Go
Shih-Chieh Huang, Remi Coulom, and Shun-Shii Lin · 2010
Cited alongside, same era.
Practical Bayesian optimization of machine learning algorithms
Jasper Snoek, Hugo Larochelle, and Ryan P. Adams · 2012
Cited alongside, same era.
Taking the human out of the loop: A review of Bayesian optimization
Bobak Shahriari, Kevin Swersky, Ziyu Wang, Ryan P. Adams, and Nando de Freitas · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Later among the works it cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, Yutian Chen, Timothy Lillicrap, Fan Hui, Laurent Sifre, George van den Driessche, Thore Graepel, and Demis Hassabis · 2017
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…