Fetching the paper…
Reading the bibliography…
The Neural Turing Machine (NTM) is more expressive than all previously considered models because of its external memory.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, Ronald J · 1992
Earlier work this paper cites.
Scaling internal-state policy-gradient methods for pomdps
Aberdeen, Douglas and Baxter, Jonathan · 2002
Earlier work this paper cites.
Policy gradient reinforcement learning for fast quadrupedal locomotion
Kohl, Nate and Stone, Peter · 2004
Earlier work this paper cites.
Optimal ordered problem solver
Schmidhuber, Jürgen · 2004
Earlier work this paper cites.
Policy gradient methods for robotics
Peters, Jan and Schaal, Stefan · 2006
Earlier work this paper cites.
Curriculum learning
Bengio, Yoshua, Louradour, Jérôme, Collobert, Ronan, and Weston, Jason · 2009
Cited alongside, same era.
Self-delimiting neural networks
Schmidhuber, Juergen · 2012
Cited alongside, same era.
Playing atari with deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Graves, Alex, Antonoglou, Ioannis, Wierstra, Daan, and Riedmiller, Martin · 2013
Cited alongside, same era.
Multiple object recognition with visual attention
Ba, Jimmy, Mnih, Volodymyr, and Kavukcuoglu, Koray · 2014
Cited alongside, same era.
Recurrent models of visual attention
Mnih, Volodymyr, Heess, Nicolas, Graves, Alex, et al · 2014
Cited alongside, same era.
Graves, Alex, Wayne, Greg, and Danihelka, Ivo
Cited in the paper.
Graves, Alex, Wayne, Greg, and Danihelka, Ivo
Cited in the paper.
Zaremba, Wojciech and Sutskever, Ilya · 2014
Later among the works it cites.
Learning to transduce with unbounded memory
Grefenstette, Edward, Hermann, Karl Moritz, Suleyman, Mustafa, and Blunsom, Phil · 2015
Closest in time.
Inferring algorithmic patterns with stack-augmented recurrent nets
Joulin, Armand and Mikolov, Tomas · 2015
Closest in time.
End-to-end training of deep visuomotor policies
Levine, Sergey, Finn, Chelsea, Darrell, Trevor, and Abbeel, Pieter · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sukhbaatar, Sainbayar, Szlam, Arthur, Weston, Jason, and Fergus, Rob · 2015
Closest in time.