Fetching the paper…
Reading the bibliography…
As humans we are driven by a strong desire for seeking novelty in our world.
Motivation reconsidered: The concept of competence
White, R. W · 1959
Earlier work this paper cites.
Intrinsic motivation and self-determination in human behavior
Deci, E. and Ryan, R. M · 1985
Earlier work this paper cites.
A survey of algorithmic methods for partially observed markov decision processes
Lovejoy, W. S · 1991
Earlier work this paper cites.
Curious model-building control systems
Schmidhuber, J · 1991
Earlier work this paper cites.
Optimal experience: Psychological studies of flow in consciousness
Csikszentmihalyi, M. and Csikszentmihalyi, I. S · 1992
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Exact and approximate algorithms for partially observable Markov decision processes
Cassandra, A. R · 1998
Earlier work this paper cites.
Mesolimbocortical and nigrostriatal dopamine responses to salient non-reward events
Horvitz, J. C · 2000
Earlier work this paper cites.
Reward, motivation, and reinforcement learning
Dayan, P. and Balleine, B. W · 2002
Earlier work this paper cites.
Dopamine: generalization and bonuses
Kakade, S. and Dayan, P · 2002
Earlier work this paper cites.
Predictive representations of state
Littman, M. L., Sutton, R. S., and Singh, S · 2002
Earlier work this paper cites.
Merriam-Webster’s collegiate dictionary
Merriam-Webster, I · 2004
Earlier work this paper cites.
Curiosity and the pleasures of learning: Wanting and liking new information
Litman, J · 2005
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
Oudeyer, P.-Y., Kaplan, F., and Hafner, V. V · 2007
Earlier work this paper cites.
How can we define intrinsic motivation?
Oudeyer, P.-Y. and Kaplan, F · 2008
Earlier work this paper cites.
Bayesian surprise attracts human attention
Itti, L. and Baldi, P · 2009
Earlier work this paper cites.
Simple algorithmic theory of subjective beauty, novelty, surprise, interestingness, attention, curiosity, creativity, art, science, music, jokes
Schmidhuber, J · 2009
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
Schmidhuber, J · 2010
Earlier work this paper cites.
Exploration in model-based reinforcement learning by empirically estimating learning progress
Lopes, M., Lang, T., Toussaint, M., and Oudeyer, P.-Y · 2012
Earlier work this paper cites.
Information-seeking, curiosity, and attention: computational and neural mechanisms
Gottlieb, J., Oudeyer, P.-Y., Lopes, M., and Baranes, A · 2013
Cited alongside, same era.
The predictive mind
Hohwy, J · 2013
Cited alongside, same era.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2013
Cited alongside, same era.
Learning and exploration in action-perception loops
Little, D. Y.-J. and Sommer, F. T · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Cited alongside, same era.
Curiosity driven reinforcement learning for motion planning on humanoids
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Later among the works it cites.
Conditional image generation with pixelcnn decoders
van den Oord, A., Kalchbrenner, N., Espeholt, L., Vinyals, O., Graves, A., et al · 2016
Later among the works it cites.
Surprise-based intrinsic motivation for deep reinforcement learning
Achiam, J. and Sastry, S · 2017
Later among the works it cites.
A nice surprise? predictive processing and the active pursuit of novelty
Clark, A · 2017
Later among the works it cites.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
Moravčík, M., Schmid, M., Burch, N., Lisỳ, V., Morrill, D., Bard, N., Davis, T., Waugh, K., Johanson, M., and Bowling, M · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Frank, M., Leitner, J., Stollenga, M., Förster, A., and Schmidhuber, J · 2014
Cited alongside, same era.
Computational psychiatry: the brain as a phantastic organ
Friston, K. J., Stephan, K. E., Montague, R., and Dolan, R. J · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Cited alongside, same era.
Draw: A recurrent neural network for image generation
Gregor, K., Danihelka, I., Graves, A., Rezende, D. J., and Wierstra, D · 2015
Cited alongside, same era.
The psychology and neuroscience of curiosity
Kidd, C. and Hayden, B. Y · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
Cited alongside, same era.
Trust region policy optimization
Schulman, J., Levine, S., Abbeel, P., Jordan, M., and Moritz, P · 2015
Cited alongside, same era.
Ostrovski, G., Bellemare, M. G., Oord, A. v. d., and Munos, R · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction
Pathak, D., Agrawal, P., Efros, A., and Darrell, T · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al · 2017
Later among the works it cites.
Amos, B., Dinh, L., Cabi, S., Rothörl, T., Colmenarejo, S. G., Muldal, A., Erez, T., Tassa, Y., de Freitas, N., and Denil, M · 2018
Later among the works it cites.
Large-scale study of curiosity-driven learning
Burda, Y., Edwards, H., Pathak, D., Storkey, A., Darrell, T., and Efros, A. A · 2018
Later among the works it cites.
Neural predictive belief representations
Guo, Z. D., Azar, M. G., Piot, B., Pires, B. A., Pohlen, T., and Munos, R · 2018
Later among the works it cites.
Ha, D. and Schmidhuber, J · 2018
Later among the works it cites.
Learning to play with intrinsically-motivated self-aware agents
Haber, N., Mrowca, D., Fei-Fei, L., and Yamins, D. L · 2018
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning
Hessel, M., Modayil, J., Van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., and Silver, D · 2018
Later among the works it cites.
Representation learning with contrastive predictive coding
Oord, A. v. d., Li, Y., and Vinyals, O · 2018
Later among the works it cites.
Observe and look further: Achieving consistent performance on atari
Pohlen, T., Piot, B., Hester, T., Azar, M. G., Horgan, D., Budden, D., Barth-Maron, G., van Hasselt, H., Quan, J., Večerík, M., et al · 2018
Later among the works it cites.
Model-based active exploration
Shyam, P., Jaśkowski, W., and Gomez, F · 2018
Later among the works it cites.
Recurrent experience replay in distributed reinforcement learning
Kapturowski, S., Ostrovski, G., Dabney, W., Quan, J., and Munos, R · 2019
Closest in time.