Multi-armed bandits for intelligent tutoring systems
Benjamin Clement, Didier Roy, Pierre-Yves Oudeyer, and Manuel Lopes · 2015
Cited alongside, same era.
Neural gpus learn algorithms
Original
Łukasz Kaiser and Ilya Sutskever · 2015
Cited alongside, same era.
Grid long short-term memory
Original
Nal Kalchbrenner, Ivo Danihelka, and Alex Graves · 2015
Cited alongside, same era.
End-to-end training of deep visuomotor policies, 2015
Original
Sergey Levine, Chelsea Finn, Trevor Darrell, and Pieter Abbeel · 2015
Cited alongside, same era.
Continuous control with deep reinforcement learning
Original
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, and et al. Alex Graves · 2015
Cited alongside, same era.
Neural programmer-interpreters
Original
Scott Reed and Nando De Freitas · 2015
Cited alongside, same era.
Trust region policy optimization
John Schulman, Sergey Levine, Pieter Abbeel, Michael I Jordan, and Philipp Moritz · 2015
Cited alongside, same era.
Incentivizing exploration in reinforcement learning with deep predictive models
Original
Bradly C Stadie, Sergey Levine, and Pieter Abbeel · 2015
Cited alongside, same era.
Reinforcement learning through asynchronous advantage actor-critic on a gpu
Mohammad Babaeizadeh, Iuri Frosio, Stephen Tyree, Jason Clemons, and Jan Kautz · 2016
Cited alongside, same era.
Unifying count-based exploration and intrinsic motivation
Marc Bellemare, Sriram Srinivasan, Georg Ostrovski, Tom Schaul, David Saxton, and Remi Munos · 2016
Cited alongside, same era.
Hybrid computing using a neural network with dynamic external memory
Alex Graves, Greg Wayne, Malcolm Reynolds, Tim Harley, Ivo Danihelka, Agnieszka Grabska-Barwińska, and et al. Sergio Gómez Colmenarejo · 2016
Cited alongside, same era.