Fetching the paper…
Reading the bibliography…
Artificial intelligence is commonly defined as the ability to achieve goals in the world.
Finding Structure in Reinforcement Learning
Thrun, Sebastian and Schwartz, Anton · 1994
Earlier work this paper cites.
Temporal Difference Learning and TD-Gammon
Tesauro, Gerald · 1995
Earlier work this paper cites.
The MAXQ Method for Hierarchical Reinforcement Learning
Dietterich, Thomas G · 1998
Earlier work this paper cites.
Roles of Macro-actions in Accelerating Reinforcement Learning
McGovern, Amy and Sutton, Richard S · 1998
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Sutton, Richard S. and Barto, Andrew G · 1998
Earlier work this paper cites.
Between MDPs and semi-MDPs: A Framework for Temporal Abstraction in Reinforcement Learning
Sutton, Richard S., Precup, Doina, and Singh, Satinder · 1999
Earlier work this paper cites.
Automatic Discovery of Subgoals in Reinforcement Learning using Diverse Density
McGovern, Amy and Barto, Andrew G · 2001
Earlier work this paper cites.
Discovering Hierarchy in Reinforcement Learning with HEXQ
Hengst, Bernhard · 2002
Earlier work this paper cites.
Q-Cut - Dynamic Discovery of Sub-goals in Reinforcement Learning
Menache, Ishai, Mannor, Shie, and Shimkin, Nahum · 2002
Earlier work this paper cites.
PolicyBlocks: An Algorithm for Creating Useful Macro-Actions in Reinforcement Learning
Pickett, Marc and Barto, Andrew G · 2002
Cited alongside, same era.
Using Relative Novelty to Identify Useful Temporal Abstractions in Reinforcement Learning
Şimşek, Özgür and Barto, Andrew G · 2004
Cited alongside, same era.
Dynamic Abstraction in Reinforcement Learning via Clustering
Mannor, Shie, Menache, Ishai, Hoze, Amit, and Klein, Uri · 2004
Cited alongside, same era.
Autonomous Inverted Helicopter Flight via Reinforcement Learning
Ng, Andrew Y., Coates, Adam, Diel, Mark, Ganapathi, Varun, Schulte, Jamie, Tse, Ben, Berger, Eric, and Liang, Eric · 2004
Cited alongside, same era.
Intrinsically Motivated Reinforcement Learning
Singh, Satinder P., Barto, Andrew G., and Chentanez, Nuttapong · 2004
Cited alongside, same era.
Identifying Useful Subgoals in Reinforcement Learning by Local Graph Partitioning
Skill Discovery in Continuous Reinforcement Learning Domains using Skill Chaining
Konidaris, George and Barreto, Andre S · 2009
Later among the works it cites.
Formal Theory of Creativity, Fun, and Intrinsic Motivation (1990-2010)
Schmidhuber, Jürgen · 2010
Later among the works it cites.
Algorithms for Reinforcement Learning
Szepesvári, Csaba · 2010
Later among the works it cites.
Abandoning Objectives: Evolution Through the Search for Novelty Alone
Lehman, Joel and Stanley, Kenneth O · 2011
Later among the works it cites.
Intrinsic Motivation and Reinforcement Learning
Barto, Andrew G · 2013
Later among the works it cites.
The Arcade Learning Environment: An Evaluation Platform for General Agents
Bellemare, M. G., Naddaf, Y., Veness, J., and Bowling, M · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Şimşek, Özgür, Wolfe, Alicia P., and Barto, Andrew G · 2005
Cited alongside, same era.
Intrinsic Motivation Systems for Autonomous Mental Development
Oudeyer, P.Y., Kaplan, F., and Hafner, VV · 2007
Cited alongside, same era.
Skill Characterization Based on Betweenness
Şimşek, Özgür and Barto, Andrew G · 2008
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Rusu, Andrei A., Veness, Joel, Bellemare, Marc G., Graves, Alex, Riedmiller, Martin, Fidjeland, Andreas K., Ostrovski, Georg, Petersen, Stig, Beattie, Charles, Sadik, Amir, Antonoglou, Ioannis, King, Helen, Kumaran, Dharshan, Wierstra, Daan, Legg, Shane, and Hassabis, Demis · 2015
Later among the works it cites.
Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation
Kulkarni, Tejas D., Narasimhan, Karthik R., Saeedi, Ardavan, and Tenenbaum, Joshua B · 2016
Closest in time.