Fetching the paper…
Reading the bibliography…
The ability to exploit prior experience to solve novel problems rapidly is a hallmark of biological learning systems and of great practical importance for artificial ones.
Multitask learning
R. Caruana · 1997
Earlier work this paper cites.
A model of inductive bias learning
J. Baxter · 2000
Earlier work this paper cites.
Relative entropy policy search
J. Peters, K. Mülling, and Y. Altün · 2010
Earlier work this paper cites.
Playing atari with deep reinforcement learning, 2013
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Pac-inspired option discovery in lifelong reinforcement learning
E. Brunskill and L. Li · 2014
Earlier work this paper cites.
Taming the noise in reinforcement learning via soft updates, 2015
R. Fox, A. Pakman, and N. Tishby · 2015
Earlier work this paper cites.
Learning continuous control policies by stochastic value gradients
N. Heess, G. Wayne, D. Silver, T. Lillicrap, T. Erez, and Y. Tassa · 2015
Earlier work this paper cites.
Rl 2 : Fast reinforcement learning via slow reinforcement learning, 2016
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel · 2016
Earlier work this paper cites.
Learning to reinforcement learn
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Earlier work this paper cites.
Learning and transfer of modulated locomotor controllers
N. Heess, G. Wayne, Y. Tassa, T. Lillicrap, M. Riedmiller, and D. Silver · 2016
Earlier work this paper cites.
Safe and efficient off-policy reinforcement learning
R. Munos, T. Stepleton, A. Harutyunyan, and M. Bellemare · 2016
Earlier work this paper cites.
Building machines that learn and think like people
B. M. Lake, T. D. Ullman, J. B. Tenenbaum, and S. J. Gershman · 2017
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
C. Finn, P. Abbeel, and S. Levine · 2017
Earlier work this paper cites.
Distral: Robust multitask reinforcement learning
Y. Teh, V. Bapst, W. M. Czarnecki, J. Quan, J. Kirkpatrick, R. Hadsell, N. Heess, and R. Pascanu · 2017
Cited alongside, same era.
Reinforcement learning with deep energy-based policies, 2017
T. Haarnoja, H. Tang, P. Abbeel, and S. Levine · 2017
Cited alongside, same era.
Successor features for transfer in reinforcement learning
A. Barreto, W. Dabney, R. Munos, J. J. Hunt, T. Schaul, H. P. van Hasselt, and D. Silver · 2017
Cited alongside, same era.
Relative entropy regularized policy iteration
A. Abdolmaleki, J. T. Springenberg, J. Degrave, S. Bohez, Y. Tassa, D. Belov, N. Heess, and M. Riedmiller · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
IMPALA: Scalable distributed deep-RL with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, S. Legg, and K. Kavukcuoglu · 2018
Later among the works it cites.
Challenges of real-world reinforcement learning, 2019
G. Dulac-Arnold, D. Mankowitz, and T. Hester · 2019
Later among the works it cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
K. Rakelly, A. Zhou, D. Quillen, C. Finn, and S. Levine · 2019
Later among the works it cites.
Varibad: A very good method for bayes-adaptive deep rl via meta-learning, 2019
L. Zintgraf, K. Shiarlis, M. Igl, S. Schulze, Y. Gal, K. Hofmann, and S. Whiteson · 2019
Later among the works it cites.
Meta-learning of sequential strategies
P. A. Ortega, J. X. Wang, M. Rowland, T. Genewein, Z. Kurth-Nelson, R. Pascanu, N. Heess, J. Veness, A. Pritzel, P. Sprechmann, et al · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Fujimoto, H. Van Hoof, and D. Meger · 2018
Cited alongside, same era.
Learning an embedding space for transferable robot skills
K. Hausman, J. T. Springenberg, Z. Wang, N. Heess, and M. Riedmiller · 2018
Cited alongside, same era.
Composing entropic policies using divergence correction, 2018
J. J. Hunt, A. Barreto, T. P. Lillicrap, and N. Heess · 2018
Cited alongside, same era.
Probabilistic model-agnostic meta-learning
C. Finn, K. Xu, and S. Levine · 2018
Cited alongside, same era.
Meta-reinforcement learning of structured exploration strategies
A. Gupta, R. Mendonca, Y. Liu, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Reptile: a scalable metalearning algorithm
A. Nichol and J. Schulman · 2018
Cited alongside, same era.
A simple neural attentive meta-learner
N. Mishra, M. Rohaninejad, X. Chen, and P. Abbeel · 2018
Cited alongside, same era.
Later among the works it cites.
Meta reinforcement learning as task inference
J. Humplik, A. Galashov, L. Hasenclever, P. A. Ortega, Y. W. Teh, and N. Heess · 2019
Later among the works it cites.
Rapid learning or feature reuse? towards understanding the effectiveness of maml
A. Raghu, M. Raghu, S. Bengio, and O. Vinyals · 2019
Later among the works it cites.
Information asymmetry in KL-regularized RL
A. Galashov, S. Jayakumar, L. Hasenclever, D. Tirumala, J. Schwarz, G. Desjardins, W. M. Czarnecki, Y. W. Teh, R. Pascanu, and N. Heess · 2019
Later among the works it cites.
Exploiting hierarchy for learning and transfer in kl-regularized rl
D. Tirumala, H. Noh, A. Galashov, L. Hasenclever, A. Ahuja, G. Wayne, R. Pascanu, Y. W. Teh, and N. Heess · 2019
Later among the works it cites.
Regularized hierarchical policies for compositional transfer in robotics
M. Wulfmeier, A. Abdolmaleki, R. Hafner, J. T. Springenberg, M. Neunert, T. Hertweck, T. Lampe, N. Siegel, N. Heess, and M. Riedmiller · 2019
Later among the works it cites.
Meta-q-learning
R. Fakoor, P. Chaudhari, S. Soatto, and A. J. Smola · 2020
Closest in time.
Q-learning in enormous action spaces via amortized approximate maximization, 2020
T. V. de Wiele, D. Warde-Farley, A. Mnih, and V. Mnih · 2020
Closest in time.