Some studies in machine learning using the game of checkers
A. L. Samuel · 1959
Earlier work this paper cites.
Learning and development in neural networks: The importance of starting small
J. L. Elman · 1993
Earlier work this paper cites.
Noise and the reality gap: The use of simulation in evolutionary robotics
N. Jakobi, P. Husbands, and I. Harvey · 1995
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Novelty or surprise?
A. Barto, M. Mirolli, and G. Baldassarre · 2013
Earlier work this paper cites.
Deterministic policy gradient algorithms
D. Silver, G. Lever, N. Heess, T. Degris, D. Wierstra, and M. Riedmiller · 2014
Earlier work this paper cites.
Poppy: open-source, 3d printed and fully-modular robotic platform for science, art and education
M. Lapeyre · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2014
Earlier work this paper cites.
Concrete problems in AI safety
Original
D. Amodei, C. Olah, J. Steinhardt, P. F. Christiano, J. Schulman, and D. Mané · 2016
Earlier work this paper cites.
Unifying count-based exploration and intrinsic motivation, 2016
M. G. Bellemare, S. Srinivasan, G. Ostrovski, T. Schaul, D. Saxton, and R. Munos · 2016
Earlier work this paper cites.
Mastering the game of Go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis · 2016
Earlier work this paper cites.