Fetching the paper…
Reading the bibliography…
This paper shows how a single mechanism allows knowledge to be constructed layer by layer directly from an agent's raw sensorimotor stream.
Pragmatism, a New Name for Some Old Ways of Thinking: Popular Lectures on Philosophy
W. James · 1907
Earlier work this paper cites.
The Organization of Behavior: A Neuropsychological Theory
Donald O. Hebb · 1949
Earlier work this paper cites.
The Principles of Psychology
William James · 1950
Earlier work this paper cites.
The Construction of Reality in the Child
Jean Piaget · 1954
Earlier work this paper cites.
Intelligence: its organization and development
Michael Cunningham · 1972
Earlier work this paper cites.
A model for the encoding of experiential information
Joseph D. Becker · 1973
Earlier work this paper cites.
Adaptation in Natural and Artificial Systems
J. H. Holland · 1975
Earlier work this paper cites.
Computer science as empirical inquiry: Symbols and search
Allen Newell and Herbert A. Simon · 1976
Earlier work this paper cites.
The map-learning critter
Benjamin J. Kuipers · 1985
Earlier work this paper cites.
A robust, qualitative method for robot spatial learning
Benjamin J. Kuipers and Yung-Tai Byun · 1988
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Richard S. Sutton · 1988
Earlier work this paper cites.
Made-Up Minds: A Constructivist Approach to Artificial Intelligence
Gary L. Drescher · 1991
Earlier work this paper cites.
Incremental development of complex behaviors through automatic construction of sensory-motor hierarchies
Mark B. Ring · 1991
Earlier work this paper cites.
Two methods for hierarchy learning in reinforcement environments
Mark B. Ring · 1993
Cited alongside, same era.
Continual Learning in Reinforcement Environments
Mark B. Ring · 1994
Cited alongside, same era.
On learning how to learn learning strategies
Jürgen Schmidhuber · 1995
Cited alongside, same era.
Map learning with uninterpreted sensors and effectors
David Pierce and Benjamin Kuipers · 1997
Cited alongside, same era.
CHILD: A first step towards continual learning
Mark B. Ring · 1997
Cited alongside, same era.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto · 1998
Cited alongside, same era.
Temporal abstraction in temporal-difference networks
Eddie J. Rafols · 2006
Later among the works it cites.
Temporal abstraction in temporal-difference networks
Richard S. Sutton, Eddie J. Rafols, and Anna Koop · 2006
Later among the works it cites.
Grounding abstractions in predictive state representations
Brian Tanner, Vadim Bulitko, Anna Koop, and Cosmin Paduraru · 2007
Later among the works it cites.
The im-clever project: Intrinsically motivated cumulative learning versatile robots
Gianluca Baldassarre, Marco Mirolli, Francesco Mannella, Daniele Caligiore, Elisabetta Visalberghi, Francesco Natale, Valentina Truppa, Gloria Sabbatini, Eugenio Guglielmelli, Flavio Keller, Domenico Campolo, Peter Redgrave, Kevin Gurney, Tom Stafford, Jochen Triesch, Cornelius Weber, Constantin Rothkopf, Ulrich Nehmzow, Mia Siddique, Mark Lee, Martin Huelse, Juergen Schmidhuber, Faustino Gomez, Alexander Foester, Julian Togelius, and Andrew Barto · 2009
Later among the works it cites.
GQ( λ \lambda ): A general gradient algorithm for temporal-difference prediction learning with eligibility traces
Hamid R. Maei and Richard S. Sutton · 2010
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Richard S. Sutton, Doina Precup, and Satinder P. Singh · 1999
Cited alongside, same era.
Predictive representations of state
Michael L. Littman, Richard S. Sutton, and Satinder Singh · 2002
Cited alongside, same era.
Universal Artificial Intelligence: Sequential Decisions based on Algorithmic Probability
M. Hutter · 2004
Cited alongside, same era.
Bayesian integration in sensorimotor learning
Konrad P. Kording and Daniel M. Wolpert · 2004
Cited alongside, same era.
Temporal-difference networks
Richard S. Sutton and Brian Tanner · 2005
Cited alongside, same era.
Bayesian decision theory in sensorimotor control
Konrad P. Körding and Daniel M. Wolpert · 2006
Cited alongside, same era.
Motor learning
Daniel M. Wolpert and J. Randall Flanagan · 2010
Later among the works it cites.
Horde: a scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
Richard S. Sutton, Joseph Modayil, Michael Delp, Thomas Degris, Patrick M. Pilarski, Adam White, and Doina Precup · 2011
Later among the works it cites.
How to set the switches on this thing
Peter Dayan · 2012
Later among the works it cites.
Acquiring diverse predictive knowledge in real time by temporal-difference learning
Joseph Modayil, Adam White, Patrick M. Pilarski, and Richard S. Sutton · 2012
Later among the works it cites.
Better Generalization with Forecasts
Tom Schaul and Mark B Ring · 2013
Later among the works it cites.
Probabilistic machine learning and artificial intelligence
Zoubin Ghahramani · 2015
Later among the works it cites.
Bayesian reinforcement learning: A survey
Mohammad Ghavamzadeh, Shie Mannor, Joelle Pineau, and Aviv Tamar · 2016
Later among the works it cites.