Fetching the paper…
Reading the bibliography…
We consider a setting in which the objective is to learn to navigate in a controlled Markov process (CMP) where transition probabilities may abruptly change.
A possibility for implementing curiosity and boredom in model-building neural controllers
Jürgen Schmidhuber · 1991
Earlier work this paper cites.
Intrinsically motivated reinforcement learning
Satinder P. Singh, Andrew G. Barto, and Nuttapong Chentanez · 2004
Earlier work this paper cites.
Experts in a Markov decision process
Eyal Even-dar, Sham M Kakade, and Yishay Mansour · 2005
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
P-Y. Oudeyer, F. Kaplan, and V.V. Hafner · 2007
Earlier work this paper cites.
What is intrinsic motivation? a typology of computational approaches
Pierre-Yves Oudeyer and Frederic Kaplan · 2007
Earlier work this paper cites.
R-IAC: Robust intrinsically motivated exploration and active learning
A. Baranes and P.-Y. Oudeyer · 2009
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
J. Schmidhuber · 2010
Earlier work this paper cites.
Intrinsically motivated reinforcement learning: An evolutionary perspective
Satinder P. Singh, Richard L. Lewis, Andrew G. Barto, and Jonathan Sorg · 2010
Earlier work this paper cites.
Autonomous exploration for navigating in MDPs
Shiau Hong Lim and Peter Auer · 2012
Cited alongside, same era.
Exploration in model-based reinforcement learning by empirically estimating learning progress
Manuel Lopes, Tobias Lang, Marc Toussaint, and Pierre-Yves Oudeyer · 2012
Cited alongside, same era.
Online learning in Markov decision processes with adversarially chosen transition probability distributions
Yasin Abbasi, Peter L Bartlett, Varun Kanade, Yevgeny Seldin, and Csaba Szepesvari · 2013
Cited alongside, same era.
Information-seeking, curiosity, and attention: computational and neural mechanisms
Jacqueline Gottlieb, Pierre-Yves Oudeyer, Manuel Lopes, and Adrien Baranes · 2013
Cited alongside, same era.
Reinforcement learning in robotics: A survey
Jens Kober, J. Andrew Bagnell, and Jan Peters · 2013
Cited alongside, same era.
Incentivizing exploration in reinforcement learning with deep predictive models
Count-based exploration with neural density models
Georg Ostrovski, Marc G Bellemare, Aäron van den Oord, and Rémi Munos · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Later among the works it cites.
Learning to play with intrinsically-motivated, self-aware agents
Nick Haber, Damian Mrowca, Stephanie Wang, Li F Fei-Fei, and Daniel L Yamins · 2018
Later among the works it cites.
Mohammad Gheshlaghi Azar, Bilal Piot, Bernardo A. Pires, Jean-Bastien Grill, Florent Altché, and Rémi Munos · 2019
Closest in time.
Large-scale study of curiosity-driven learning
Yuri Burda, Harrison Edwards, Deepak Pathak, Amos J. Storkey, Trevor Darrell, and Alexei A. Efros · 2019
Closest in time.
Provably efficient maximum entropy exploration
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bradly C. Stadie, Sergey Levine, and Pieter Abbeel · 2015
Cited alongside, same era.
Variational information maximizing exploration
Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel · 2016
Cited alongside, same era.
Surprise-based intrinsic motivation for deep reinforcement learning
Joshua Achiam and Shankar Sastry · 2017
Cited alongside, same era.
Elad Hazan, Sham Kakade, Karan Singh, and Abby Van Soest · 2019
Closest in time.
Deep reinforcement learning robot for search and rescue applications: Exploration in unknown cluttered environments
F. Niroui, K. Zhang, Z. Kashino, and G. Nejat · 2019
Closest in time.
Variational regret bounds for reinforcement learning
Ronald Ortner, Pratik Gajane, , and Peter Auer · 2019
Closest in time.