Fetching the paper…
Reading the bibliography…
We consider the problem of how a teacher algorithm can enable an unknown Deep Reinforcement Learning (DRL) student to become good at a skill over a wide range of diverse environments.
Multidimensional binary search trees used for associative searching
J. L. Bentley · 1975
Earlier work this paper cites.
Maximum likelihood from incomplete data via the em algorithm
A. P. Dempster, N. M. Laird, and D. B. Rubin · 1977
Earlier work this paper cites.
Model selection and akaike’s information criterion (aic): The general theory and its analytical extensions
H. Bozdogan · 1987
Earlier work this paper cites.
Learning and development in neural networks: the importance of starting small
J. L. Elman · 1993
Earlier work this paper cites.
The infinite gaussian mixture model
C. E. Rasmussen · 2000
Earlier work this paper cites.
The nonstochastic multiarmed bandit problem
P. Auer, N. Cesa-Bianchi, Y. Freund, and R. E. Schapire · 2002
Earlier work this paper cites.
Bringing up robot: Fundamental mechanisms for creating a self-motivated, self-organizing architecture
D. Blank, D. Kumar, L. Meeden, and J. Marshall · 2003
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
P.-Y. Oudeyer, F. Kaplan, and V. V. Hafner · 2006
Earlier work this paper cites.
In search of the neural circuits of intrinsic motivation
F. Kaplan and P.-Y. Oudeyer · 2007
Earlier work this paper cites.
Flexible shaping: How learning in small steps helps
K. A. Krueger and P. Dayan · 2008
Earlier work this paper cites.
R-IAC: robust intrinsically motivated exploration and active learning
A. Baranes and P. Oudeyer · 2009
Cited alongside, same era.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Cited alongside, same era.
Transfer learning for reinforcement learning domains: A survey
M. E. Taylor and P. Stone · 2009
Cited alongside, same era.
The Strategic Student Approach for Life-Long Exploration and Learning
M. Lopes and P.-Y. Oudeyer · 2012
Cited alongside, same era.
Active learning of inverse models with intrinsically motivated goal exploration in robots
A. Baranes and P.-Y. Oudeyer · 2012
Cited alongside, same era.
Self-organization of early vocal development in infants and machines: The role of intrinsic motivation
C. Moulin-Frier, S. M. Nguyen, and P.-Y. Oudeyer · 2013
Automated curriculum learning for neural networks
A. Graves, M. G. Bellemare, J. Menick, R. Munos, and K. Kavukcuoglu · 2017
Later among the works it cites.
Intrinsically motivated goal exploration processes with automatic curriculum learning
S. Forestier, Y. Mollard, and P. Oudeyer · 2017
Later among the works it cites.
Reward-guided curriculum for robust reinforcement learning
P. R. S. K. Mysore, S · 2018
Later among the works it cites.
Automatic goal generation for reinforcement learning agents
C. Florensa, D. Held, X. Geng, and P. Abbeel · 2018
Later among the works it cites.
Reinforcement learning for improving agent design
D. Ha · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multi-Armed Bandits for Intelligent Tutoring Systems
B. Clément, D. Roy, P.-Y. Oudeyer, and M. Lopes · 2015
Cited alongside, same era.
Modular active curiosity-driven discovery of tool use
S. Forestier and P. Oudeyer · 2016
Cited alongside, same era.
The malmo platform for artificial intelligence experimentation
M. Johnson, K. Hofmann, T. Hutton, and D. Bignell · 2016
Cited alongside, same era.
Openai gym, 2016
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Curiosity driven exploration of learned disentangled goal spaces
A. Laversanne-Finot, A. Pere, and P.-Y. Oudeyer · 2018
Later among the works it cites.
CURIOUS: intrinsically motivated modular multi-goal reinforcement learning
C. Colas, P. Oudeyer, O. Sigaud, P. Fournier, and M. Chetouani · 2019
Closest in time.
Teacher-student curriculum learning
T. Matiisen, A. Oliver, T. Cohen, and J. Schulman · 2019
Closest in time.
R. Wang, J. Lehman, J. Clune, and K. O. Stanley · 2019
Closest in time.