Fetching the paper…
Reading the bibliography…
Deep learning has enabled traditional reinforcement learning methods to deal with high-dimensional problems.
J. N. Sitsiklis and B. V. Roy. An analysis of temporal-difference learning with function approximation
1997
Earlier work this paper cites.
R. S. Sutton and A. G. Barto. Reinforcement Learning: An Introduction. The MIT Press, Cambridge, MA, 1998
1998
Earlier work this paper cites.
V. R. Konda and J. N. Tsitsiklis. Actor-critic algorithms. In
2000
Earlier work this paper cites.
T. G. Dietterich. Hierarchical reinforcement learning with the maxq value function decomposition
2000
Earlier work this paper cites.
N. Chentanez, A. G. Barto, and S. P. Singh. Intrinsically motivated reinforcement learning. In
2005
Earlier work this paper cites.
H. V. Hasselt. Double q-learning. In
2010
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling. The arcade learning environment: An evaluation platform for general agents
2013
Earlier work this paper cites.
R. Veerabhadrappa, A. Bhatti, C. P. Lim, T. T. Nguyen, S. J. Tye, P. Monaghan, and S. Nahavandi. Statistical modelling of artificial neural network for sorting temporally synchronous spikes. In
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al. Human-level control through deep reinforcement learning
2015
Cited alongside, same era.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich. Going deeper with convolutions. In
2015
Cited alongside, same era.
T. Schaul, J. Quan, I. Antonoglou, and D. Silver. Prioritized experience replay. In
2016
Cited alongside, same era.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu. Asynchronous methods for deep reinforcement learning. In
2016
Cited alongside, same era.
Z. Wang, T. Schaul, M. Hessel, H. Hasselt, M. Lanctot, and N. Freitas. Dueling network architectures for deep reinforcement learning. In
2016
S. Sukhbaatar and R. Fergus. Learning multiagent communication with backpropagation. In
2016
Later among the works it cites.
N. D. Nguyen, T. Nguyen, and S. Nahavandi. System design perspective for human-level agents using deep reinforcement learning: a survey
2017
Later among the works it cites.
2017
Later among the works it cites.
P. F. Christiano, J. Leike, T. Brown, M. Martic, S. Legg, and D. Amodei. Deep reinforcement learning from human preferences. In
2017
Later among the works it cites.
J. K. Gupta, M. Egorov, and M. Kochenderfer. Cooperative multi-agent control using deep reinforcement learning. In
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
B. Zhou, A. Khosla, A. Lapedriza, A. Oliva, and A. Torralba. Learning deep features for discriminative localization. In
2016
Cited alongside, same era.
T. D. Kulkarni, K. R. Narasimhan, A. Saeedi, and J. B. Tenenbaum. Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation. In
2016
Cited alongside, same era.
J. Foerster, Y. M. Assael, N. de Freitas, and S. Whiteson. Learning to communicate with deep multi-agent reinforcement learning. In
2016
Cited alongside, same era.
2018
Closest in time.
2018
Closest in time.
A. Khatami, M. Babaie, H. R. Tizhoosh, A. Khosravi, T. Nguyen, and S. Nahavandi. A sequential search-space shrinking using CNN transfer learning and a radon projection pool for medical image retrieval
2018
Closest in time.