Fetching the paper…
Reading the bibliography…
A critical component to enabling intelligent reasoning in partially observable environments is memory.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Reinforcement Learning: an Introduction
R. Sutton and A. Barto · 1998
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey J Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. H. Cho, and Y. Bengio · 2015
Earlier work this paper cites.
Deep recurrent q-learning for partially observable mdps
M. Hausknecht and P. Stone · 2015
Earlier work this paper cites.
Spatial transformer networks
M. Jaderberg, K. Simonyan, and A. Zisserman · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Cited alongside, same era.
End-to-end memory networks
S. Sukhbaatar, A. Szlam, J. Weston, and R. Fergus · 2015
Cited alongside, same era.
Playing doom with slam-augmented deep reinforcement learning
S. Bhatti, A. Desmaison, O. Miksik, N. Nardelli, N. Siddharth, and P. H. S. Torr · 2016
Cited alongside, same era.
Hybrid computing using a neural network with dynamic external memory
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-Barwińska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, A. P. Badia, K. M. Hermann, Y. Zwols, G. Ostrovski, A. Cain, H. King, C. Summerfield, P. Blunsom, K. Kavukcuoglu, and D. Hassabis · 2016
Cited alongside, same era.
Neural gpus learn algorithms
L. Kaiser and I. Sutskever · 2016
Cited alongside, same era.
End-to-end training of deep visuomotor policies
Sergey Levine, Chelsea Finn, Trevor Darrell, and Pieter Abbeel · 2016
Later among the works it cites.
Key-value memory networks for directly reading documents
A. Miller, A. Fisch, J. Dodge, A. Karimi, A. Bordes, and J. Weston · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Harley, T. P. Lillicrap, D. Silver, and K. Kavukcuoglu · 2016
Later among the works it cites.
Control of memory, active perception, and action in minecraft
J. Oh, V. Chockalingam, S. Singh, and H. Lee · 2016
Later among the works it cites.
Scaling memory-augmented neural networks with sparse reads and writes
J. W. Rae, J. J. Hunt, T. Harley, I. Danihelka, A. Senior, G. Wayne, A. Graves, and T. Lillicrap · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
ViZDoom: A Doom-based AI research platform for visual reinforcement learning
Michał Kempka, Marek Wydmuch, Grzegorz Runc, Jakub Toczek, and Wojciech Jaśkowski · 2016
Cited alongside, same era.
Neural random-access machines
K. Kurach, M. Andrychowicz, and I. Sutskever · 2016
Cited alongside, same era.
Playing fps games with deep reinforcement learning
G. Lample and D. S. Chaplot · 2016
Cited alongside, same era.
Value iteration networks
Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, and Pieter Abbeel · 2016
Later among the works it cites.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Closest in time.