Fetching the paper…
Reading the bibliography…
Models that can simulate how environments change in response to actions can be used by agents to plan and act efficiently.
Intuitive physics
M. McCloskey · 1983
Earlier work this paper cites.
Gradient-based learning algorithms for recurrent networks and their computational complexity
R. J. Williams and D. Zipser · 1995
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
A sensorimotor account of vision and visual consciousness
J. K. O’Regan and A. Noë · 2001
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
P. Auer, N. Cesa-Bianchi, and P. Fischer · 2002
Earlier work this paper cites.
Predictive representations of state
M. L. Littman, R. S. Sutton, and S. Singh · 2002
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
P.-Y. Oudeyer, F. Kaplan, and V. V. Hafner · 2007
Earlier work this paper cites.
Hippocampal contributions to control: The third way
M. Lengyel and P. Dayan · 2008
Earlier work this paper cites.
Reinforcement learning in the brain
Y. Niv · 2009
Cited alongside, same era.
Causality
J. Pearl · 2009
Cited alongside, same era.
The Arcade Learning Environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013
Cited alongside, same era.
Generating sequences with recurrent neural networks
A. Graves · 2013
Cited alongside, same era.
Torcs: The open racing car simulator, v1.3.5
B. Wymann, E. Espié, C. Guionneau, C. Dimitrakakis, R. Coulom, and A. Sumner · 2013
Cited alongside, same era.
Model regularization for stable sample rollouts
E. Talvitie · 2014
Cited alongside, same era.
Spatio-temporal video autoencoder with differentiable memory
V. Patraucean, A. Handa, and R. Cipolla · 2015
Later among the works it cites.
Unsupervised learning of video representations using LSTMs
N. Srivastava, E. Mansimov, and R. Salakhutdinov · 2015
Later among the works it cites.
Learning to filter with predictive state inference machines
W. Sun, A. Venkatraman, B. Boots, and J. A. Bagnell · 2015
Later among the works it cites.
From pixels to torques: Policy learning with deep dynamical models
N. Wahlström, T. B. Schön, and M. P. Deisenroth · 2015
Later among the works it cites.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scheduled sampling for sequence prediction with recurrent neural networks
S. Bengio, O. Vinyals, N. Jaitly, and N. Shazeer · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Cited alongside, same era.
Action-conditional video prediction using deep networks in Atari games
J. Oh, X. Guo, H. Lee, R. L. Lewis, and S. P. Singh · 2015
Cited alongside, same era.
Empirical evaluation of rectified activations in convolutional network
B. Xu, N. Wang, T. Chen, and M. Li · 2015
Later among the works it cites.
C. Beattie, J. Z. Leibo, D. Teplyashin, T. Ward, M. Wainwright, H. Küttler, A. Lefrancq, S. Green, V. Valdés, A. Sadik, J. Schrittwieser, K. Anderson, S. York, M. Cant, A. Cain, A. Bolton, S. Gaffney, H. King, D. Hassabis, S. Legg, and S. Petersen · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. Puigdomènech Badia, M. Mirza, A. Graves, T. P Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Later among the works it cites.
Recurrent neural network regularization
W. Zaremba, I. Sutskever, and O. Vinyals · 2048
Closest in time.