Samples are not all useful: Denoising policy gradient updates using variance
Original
Yannis Flet-Berliac and Philippe Preux · 1904
Earlier work this paper cites.
Alan M Turing
1950
Earlier work this paper cites.
Training and tracking in robotics
Oliver G Selfridge et al · 1985
Earlier work this paper cites.
Curious model-building control systems
Jürgen Schmidhuber · 1991
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Long-Ji Lin · 1992
Earlier work this paper cites.
Learning and development in neural networks: the importance of starting small
Jeffrey L. Elman · 1993
Earlier work this paper cites.
Prioritized sweeping: Reinforcement learning with less data and less time
Andrew W Moore and Christopher G Atkeson · 1993
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
Pierre-Yves Oudeyer et al · 2007
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio et al · 2009
Earlier work this paper cites.
Flexible shaping: How learning in small steps helps
Kai A. Krueger and Peter Dayan · 2009
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Matthew E. Taylor and Peter Stone · 2009
Earlier work this paper cites.
The interaction of maturational constraints and intrinsic motivations in active motor development
A. Baranes and P. Oudeyer · 2011
Earlier work this paper cites.
The strategic student approach for life-long exploration and learning
Manuel Lopes and Pierre-Yves Oudeyer · 2012
Earlier work this paper cites.
Active learning of inverse models with intrinsically motivated goal exploration in robots
Adrien Baranes and Pierre-Yves Oudeyer · 2013
Earlier work this paper cites.
How to discount deep reinforcement learning: Towards new dynamic strategies
Vincent François-Lavet et al · 2015
Earlier work this paper cites.
Language understanding for text-based games using deep reinforcement learning
Karthik Narasimhan et al · 2015
Earlier work this paper cites.
Unifying count-based exploration and intrinsic motivation
Marc Bellemare et al · 2016
Earlier work this paper cites.
Reinforcement learning with unsupervised auxiliary tasks
Max Jaderberg et al · 2016
Earlier work this paper cites.
Hindsight experience replay
Marcin Andrychowicz et al · 2017
Earlier work this paper cites.
Emergent complexity via multi-agent competition
Trapit Bansal et al · 2017
Earlier work this paper cites.
Reverse curriculum generation for reinforcement learning
Carlos Florensa et al · 2017
Earlier work this paper cites.
Intrinsically motivated goal exploration processes with automatic curriculum learning
Sébastien Forestier et al · 2017
Earlier work this paper cites.