Fetching the paper…
Reading the bibliography…
While Deep Reinforcement Learning (DRL) has emerged as a promising approach to many complex tasks, it remains challenging to train a single DRL agent that is capable of undertaking multiple different continuous control tasks.
V. Mnih, A. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, Asynchronous methods for deep reinforcement learning, ICML’16 , pp. 1928–1937
1937
Earlier work this paper cites.
2001
Earlier work this paper cites.
2003
Earlier work this paper cites.
B. Goertzel and C. Pennachin, Artificial general intelligence, Springer , Vol. 2, 2007
2007
Earlier work this paper cites.
2010
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, Mujoco: A physics engine for model-based control, IEEE/RSJ International Conference on Intelligent Robots and Systems , 2012, pp. 5026–5033
2012
Cited alongside, same era.
2015
Cited alongside, same era.
V. Mnih, et al
2015
Cited alongside, same era.
L. Liu, U. Dogan, and K. Hofmann, Decoding multitask dqn in the world of minecraft, European Workshop on Reinforcement Learning (EWRL) , 2016
2016
Cited alongside, same era.
D. E. Carlo, T. Davide, B. Andrea and R. Marcello, Sharing knowledge in multi-task deep reinforcement learning, ICLR’20
Cited in the paper.
2017
Later among the works it cites.
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, and et al
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Czarnecki, R. Pascanu, S. Osindero, S. Jayakumar, G. Swirszcz, and M. Jaderberg, Distilling policy distillation, AISTATS’19 , pp. 1331–1340
Cited in the paper.
S. Fujimoto, V. H. Herke, and D. Meger, Addressing function approximation error in actor-critic methods, ICML’18
Cited in the paper.
T. Haarnoja, A. Zhou, P. Abbeel and S. Levine, Soft actor-critic: off-policy maximum entropy deep reinforcement learning with a stochastic actor, ICML’18
Cited in the paper.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver and D. Wierstra, Continuous control with deep reinforcement learning, ICLR’16
Cited in the paper.
E. Parisotto, J. L. Ba, and R. Salakhutdinov, Actor-mimic: Deep multitask and transfer reinforcement learning, ICLR’16
Cited in the paper.
A. A. Rusu, S. G. Colmenarejo, C. Gulcehre, G. Desjardins, J. Kirkpatrick, and et al
Cited in the paper.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, Trust region policy optimization, ICML’15 , pp. 1889–1897
Cited in the paper.