Fetching the paper…
Reading the bibliography…
Humans spend a remarkable fraction of waking life engaged in acts of "mental time travel".
The effect of the introduction of reward upon the maze performance of rats
Blodgett, H. C · 1929
Earlier work this paper cites.
A note on measurement of utility
Samuelson, P. A · 1937
Earlier work this paper cites.
Cognitive maps in rats and men
Tolman, E. C · 1948
Earlier work this paper cites.
The chess machine: an example of dealing with a complex task by adaptation
Newell, A · 1955
Earlier work this paper cites.
Some studies in machine learning using the game of checkers
Samuel, A. L · 1959
Earlier work this paper cites.
Steps toward artificial intelligence
Minsky, M · 1961
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J · 1992
Earlier work this paper cites.
Neural networks and the bias/variance dilemma
Geman, S., Bienenstock, E. & Doursat, R · 1992
Earlier work this paper cites.
Endowments and assets: The anthropology of wealth and the economics of intrahousehold allocation
Guyer, J · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction (MIT press, 1998)
Sutton, R. S., Barto, A. G. et al · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Sutton, R. S., McAllester, D. A., Singh, S. P. & Mansour, Y · 2000
Earlier work this paper cites.
Infinite-horizon policy-gradient estimation
Baxter, J. & Bartlett, P. L · 2001
Earlier work this paper cites.
Time discounting and time preference: A critical review
Frederick, S., Loewenstein, G. & O’Donoghue, T · 2002
Earlier work this paper cites.
The dawn of human culture (Wiley New York, 2002)
Klein, R. G. & Edgar, B · 2002
Earlier work this paper cites.
Dopamine-dependent facilitation of ltp induction in hippocampal ca1 by exposure to spatial novelty
Li, S., Cullen, W. K., Anwyl, R. & Rowan, M. J · 2003
Earlier work this paper cites.
Delaying execution of intentions: Overcoming the costs of interruptions
McDaniel, M. A., Einstein, G. O., Graham, T. & Rall, E · 2004
Cited alongside, same era.
Dopamine d1/d5 receptors gate the acquisition of novel information through hippocampal long-term potentiation and long-term depression
Lemon, N. & Manahan-Vaughan, D · 2006
Cited alongside, same era.
Visual long-term memory has a massive storage capacity for object details
Brady, T. F., Konkle, T., Alvarez, G. A. & Oliva, A · 2008
Cited alongside, same era.
Signal-to-noise ratio analysis of policy gradient algorithms
Roberts, J. W. & Tedrake, R · 2009
Cited alongside, same era.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Simonyan, K., Vedaldi, A. & Zisserman, A · 2013
Cited alongside, same era.
High-dimensional continuous control using generalized advantage estimation
Schulman, J., Moritz, P., Levine, S., Jordan, M. & Abbeel, P · 2015
Later among the works it cites.
Optimizing Expectations: From Deep Reinforcement Learning to Stochastic Computation Graphs
Schulman, J · 2016
Later among the works it cites.
Hybrid computing using a neural network with dynamic external memory
Graves, A. et al · 2016
Later among the works it cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S. & Sun, J · 2016
Later among the works it cites.
A guide to convolution arithmetic for deep learning
Dumoulin, V. & Visin, F · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Speech recognition with deep recurrent neural networks
Graves, A., Mohamed, A.-r. & Hinton, G · 2013
Cited alongside, same era.
Training recurrent neural networks
Sutskever, I · 2013
Cited alongside, same era.
The Recursive Mind: The Origins of Human Language, Thought, and Civilization-Updated Edition (Princeton University Press, 2014)
Corballis, M. C · 2014
Cited alongside, same era.
Bias in natural actor-critic algorithms
Thomas, P · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K. & Bengio, Y · 2014
Cited alongside, same era.
Graves, A., Wayne, G. & Danihelka, I · 2014
Cited alongside, same era.
The cifar-10 dataset
Krizhevsky, A., Nair, V. & Hinton, G · 2014
Cited alongside, same era.
Mnih, V. et al · 2016
Later among the works it cites.
Beattie, C. et al · 2016
Later among the works it cites.
Neuroscience-inspired artificial intelligence
Hassabis, D., Kumaran, D., Summerfield, C. & Botvinick, M · 2017
Later among the works it cites.
Unsupervised predictive memory in a goal-directed agent
Wayne, G. et al · 2018
Closest in time.
Been there, done that: Meta-learning with episodic recall
Ritter, S. et al · 2018
Closest in time.
Sparse attentive backtracking: Temporal creditassignment through reminding
Ke, N. R. et al · 2018
Closest in time.
Rudder: Return decomposition for delayed rewards
Arjona-Medina, J. A., Gillhofer, M., Widrich, M., Unterthiner, T. & Hochreiter, S · 2018
Closest in time.
The Book of Why: The New Science of Cause and Effect (Basic Books, 2018)
Pearl, J. & Mackenzie, D · 2018
Closest in time.
Psychlab: a psychology laboratory for deep reinforcement learning agents
Leibo, J. Z. et al · 2018
Closest in time.