Fetching the paper…
Reading the bibliography…
Intrinsically motivated artificial agents learn advantageous behavior without externally-provided rewards.
Reinforcement learning in continuous time and space
Kenji Doya · 2000
Earlier work this paper cites.
Intrinsically motivated learning of hierarchical collections of skills
Andrew G Barto, Satinder Singh, and Nuttapong Chentanez · 2004
Earlier work this paper cites.
Intrinsically motivated reinforcement learning
Nuttapong Chentanez, Andrew G. Barto, and Satinder P. Singh · 2005
Earlier work this paper cites.
Empowerment: a universal agent-centric measure of control
A. S. Klyubin, D. Polani, and C. L. Nehaniv · 2005
Earlier work this paper cites.
All else being equal be empowered
Alexander S Klyubin, Daniel Polani, and Chrystopher L Nehaniv · 2005
Earlier work this paper cites.
Optimal control theory
Emanuel Todorov · 2006
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
Jürgen Schmidhuber · 2010
Earlier work this paper cites.
Empowerment for continuous agent—environment systems
Tobias Jung, Daniel Polani, and Peter Stone · 2011
Earlier work this paper cites.
Information theory of decisions and actions
Naftali Tishby and Daniel Polani · 2011
Earlier work this paper cites.
Elements of information theory
Thomas M Cover and Joy A Thomas · 2012
Earlier work this paper cites.
Intrinsic motivation and reinforcement learning
Andrew G Barto · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning, 2013
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Approximation of empowerment in the continuous domain
Christoph Salge, Cornelius Glackin, and Daniel Polani · 2013
Earlier work this paper cites.
Causal entropic forces
Alexander D Wissner-Gross and Cameron E Freer · 2013
Cited alongside, same era.
Empowerment–an introduction
Christoph Salge, Cornelius Glackin, and Daniel Polani · 2014
Cited alongside, same era.
Sapiens: A Brief History of Humankind
Harari Yuval, Noah · 2014
Cited alongside, same era.
Variational information maximisation for intrinsically motivated reinforcement learning
Shakir Mohamed and Danilo Jimenez Rezende · 2015
Cited alongside, same era.
Openai gym, 2016
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Intrinsically motivated general companion npcs via coupled empowerment maximisation
Christian Guckelsberger, Christoph Salge, and Simon Colton · 2016
Proximal policy optimization algorithms, 2017
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Control capacity of partially observable dynamic systems in continuous time
Stas Tiomkin, Daniel Polani, and Naftali Tishby · 2017
Later among the works it cites.
Variational option discovery algorithms, 2018
Joshua Achiam, Harrison Edwards, Dario Amodei, and Pieter Abbeel · 2018
Later among the works it cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Later among the works it cites.
Nonlinear dynamics and chaos with student solutions manual: With applications to physics, biology, chemistry, and engineering
Steven H Strogatz · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Intrinsic motivation, curiosity, and learning: Theory and applications in educational technologies
P-Y Oudeyer, Jacqueline Gottlieb, and Manuel Lopes · 2016
Cited alongside, same era.
Deep variational information bottleneck
Alex Alemi, Ian Fischer, Josh Dillon, and Kevin Murphy · 2017
Cited alongside, same era.
Openai baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov · 2017
Cited alongside, same era.
Variational intrinsic control
Karol Gregor, Danilo Jimenez Rezende, and Daan Wierstra · 2017
Cited alongside, same era.
Vime: Variational information maximizing exploration, 2017
Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel · 2017
Cited alongside, same era.
Curiosity-driven exploration by self-supervised prediction, 2017
Deepak Pathak, Pulkit Agrawal, Alexei A. Efros, and Trevor Darrell · 2017
Cited alongside, same era.
Reinforcement learning: An introduction , volume 1
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Diversity is all you need: Learning skills without a reward function
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2019
Later among the works it cites.
Unsupervised real-time control through variational empowerment
Maximilian Karl, Philip Becker-Ehmck, Maximilian Soelch, Djalel Benbouzid, Patrick van-der Smagt, and Justin Bayer · 2019
Later among the works it cites.
On variational bounds of mutual information, 2019
Ben Poole, Sherjil Ozair, Aaron van den Oord, Alexander A. Alemi, and George Tucker · 2019
Later among the works it cites.
Explore, discover and learn: Unsupervised discovery of state-covering skills
Víctor Campos, Alexander Trott, Caiming Xiong, Richard Socher, Xavier Giró-i-Nieto, and Jordi Torres · 2020
Closest in time.
Dynamics-aware unsupervised discovery of skills
Archit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar, and Karol Hausman · 2020
Closest in time.
On mutual information maximization for representation learning
Michael Tschannen, Josip Djolonga, Paul K. Rubenstein, Sylvain Gelly, and Mario Lucic · 2020
Closest in time.
Ave: Assistance via empowerment
Yuqing Du, Stas Tiomkin, Emre Kiciman, Daniel Polani, Pieter Abbeel, and Anca Dragan · 2021
Closest in time.