Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning (RL) recently emerged as one of the most competitive approaches for learning in sequential decision making problems with fully observable environments, e.g., computer Go.
Christopher J. C. H. Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Reinforcement learning for robots using neural networks
Long-Ji Lin · 1993
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
A novel orthogonal nmf-based belief compression for pomdps
Xin Li, William Kwok-Wai Cheung, Jiming Liu, and Zhili Wu · 2007
Earlier work this paper cites.
SARSOP: efficient point-based POMDP planning by approximating optimally reachable belief spaces
Hanna Kurniawati, David Hsu, and Wee Sun Lee · 2008
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G. Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller · 2013
Cited alongside, same era.
A survey of point-based pomdp solvers
Guy Shani, Joelle Pineau, and Robert Kaplow · 2013
Cited alongside, same era.
Deep reinforcement learning with pomdps
Maxim Egorov · 2015
Cited alongside, same era.
Deep recurrent q-learning for partially observable mdps
Matthew J. Hausknecht and Peter Stone · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Kavukcuoglu Koray, Silver David, Rusu Andrei A, Veness Joel, Bellemare Marc G, Graves Alex, Riedmiller Martin, Fidjeland Andreas K, Ostrovski Georg, et al · 2015
Later among the works it cites.
Incentivizing exploration in reinforcement learning with deep predictive models
Bradly C. Stadie, Sergey Levine, and Pieter Abbeel · 2015
Later among the works it cites.
Learning to communicate to solve riddles with deep distributed recurrent q-networks
Jakob N. Foerster, Yannis M. Assael, Nando de Freitas, and Shimon Whiteson · 2016
Later among the works it cites.
Playing FPS games with deep reinforcement learning
Guillaume Lample and Devendra Singh Chaplot · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Timothy P. Lillicrap, Jonathan J. Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Cited alongside, same era.
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.