Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning methods have achieved state of the art performance in learning control policies for the games in the Atari 2600 domain.
Technical note: Q-learning
Christopher J. C. H. Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Reinforcement learning for robots using neural networks
Long-Ji Lin · 1993
Earlier work this paper cites.
Introduction to reinforcement learning
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Reinforcement learning of local shape in the game of go
David Silver, Richard Sutton, and Martin Muller · 2007
Earlier work this paper cites.
Theano: a CPU and GPU math expression compiler
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio · 2010
Earlier work this paper cites.
Theano: new features and speed improvements
Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, James Bergstra, Ian J. Goodfellow, Arnaud Bergeron, Nicolas Bouchard, and Yoshua Bengio · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G. Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
Imitating human playing styles in Super Mario Bros
Juan Ortega, Noor Shaker, Julian Togelius, and Georgios N. Yannakakis · 2013
Cited alongside, same era.
Temporal abstraction in monte carlo tree search
Mostafa Vafadost · 2013
Cited alongside, same era.
Frame skip is a powerful parameter for learning to play atari
Alex Braylan, Mark Hollenbeck, Elliot Meyerson, and Risto Miikkulainen · 2015
Cited alongside, same era.
Deep reinforcement learning in parametrized action space
Matthew Hausknecht and Peter Stone · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Language understanding for text-based games using deep reinforcement learning
Narasimhan, Tejas Kulkarni, and Regina Barzilay · 2015
Later among the works it cites.
Tom Schaul, John Quan, Ioannis Antonoglou, and David Silver · 2015
Later among the works it cites.
Dueling network architectures for deep reinforcement learning
Ziyu Wang, Schaul, Matteo Hessel, Hado van Hasselt, Marc Lanctot, and Nando de Freitas · 2015
Later among the works it cites.
Learning to communicate to solve riddles with deep distributed recurrent q-networks
Jakob N. Foerster, Yannis M. Assael, Nando de Freitas, and Shimon Whiteson · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Cited alongside, same era.
Volodymyr Mnih, Adria Puigdom enech Badia, Mehdi Mirza, Alex Graves, Tim Harley, Timothy P. Lillicrap, David Silver, and Koray Kavukcuoglu · 2016
Closest in time.