Fetching the paper…
Reading the bibliography…
Meta-reinforcement learning algorithms can enable autonomous agents, such as robots, to quickly acquire new behaviors by leveraging prior experience in a set of related training tasks.
Evolutionary principles in self-referential learning, or on learning how to learn: the meta-meta-… hook
J. Schmidhuber · 1987
Earlier work this paper cites.
Acquiring robot skills via reinforcement learning
V. Gullapalli, J. A. Franklin, and H. Benbrahim · 1994
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra · 1998
Earlier work this paper cites.
A Bayesian framework for concept learning
J. B. Tenenbaum · 1999
Earlier work this paper cites.
Robotic grasping and contact: A review
A. Bicchi and V. Kumar · 2000
Earlier work this paper cites.
Interpretation of force and moment signals for compliant peg-in-hole assembly
W. S. Newman, Y. Zhao, and Y.-H. Pao · 2001
Earlier work this paper cites.
A bayesian approach to unsupervised one-shot learning of object categories
L. Fei-Fei et al · 2003
Earlier work this paper cites.
Point-based value iteration: An anytime algorithm for pomdps
J. Pineau, G. Gordon, S. Thrun, et al · 2003
Earlier work this paper cites.
Decentralized algorithms for multi-robot manipulation via caging
G. A. Pereira, M. F. Campos, and V. Kumar · 2004
Earlier work this paper cites.
Probabilistic robotics, vol. 1, 2005
S. Thrun, W. Burgard, D. Fox, et al · 2005
Earlier work this paper cites.
Neural network based state estimation of dynamical systems
N. Yadaiah and G. Sowmya · 2006
Earlier work this paper cites.
Graphical models, exponential families, and variational inference
M. J. Wainwright and M. I. Jordan · 2008
Earlier work this paper cites.
A bayesian approach for learning and planning in partially observable markov decision processes
S. Ross, J. Pineau, B. Chaib-draa, and P. Kreitmann · 2011
Earlier work this paper cites.
Robot manipulation of deformable objects
D. Henrich and H. Wörn · 2012
Earlier work this paper cites.
Solving nonlinear continuous state-action-observation pomdps for mechanical systems with gaussian noise
M. P. Deisenroth and J. Peters · 2012
Earlier work this paper cites.
Autonomous reinforcement learning on raw visual input data in a real world application
S. Lange, M. Riedmiller, and A. Voigtländer · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Reinforcement learning in robotics: A survey
J. Kober, J. A. Bagnell, and J. Peters · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Task transfer via collaborative manipulation for insertion assembly
K. Kronander, E. Burdet, and A. Billard · 2014
Earlier work this paper cites.
Memory-based control with recurrent neural networks
N. Heess, J. J. Hunt, T. P. Lillicrap, and D. Silver · 2015
Earlier work this paper cites.
Deep recurrent q-learning for partially observable mdps
M. Hausknecht and P. Stone · 2015
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Cited alongside, same era.
Learning to reinforcement learn
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Cited alongside, same era.
Rl2: Fast reinforcement learning via slow reinforcement learning
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel · 2016
Cited alongside, same era.
Deep spatial autoencoders for visuomotor learning
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel · 2016
Cited alongside, same era.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Deep object pose estimation for semantic robotic grasping of household objects
J. Tremblay, T. To, B. Sundaralingam, Y. Xiang, D. Fox, and S. Birchfield · 2018
Later among the works it cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Meta-reinforcement learning of structured exploration strategies
A. Gupta, R. Mendonca, Y. Liu, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
K. Rakelly, A. Zhou, D. Quillen, C. Finn, and S. Levine · 2019
Later among the works it cites.
Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model
A. X. Lee, A. Nagabandi, P. Abbeel, and S. Levine · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Self-supervised visual descriptor learning for dense correspondence
T. Schmidt, R. Newcombe, and D. Fox · 2016
Cited alongside, same era.
Deep variational bayes filters: Unsupervised learning of state space models from raw data
M. Karl, M. Soelch, J. Bayer, and P. van der Smagt · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
C. Finn, P. Abbeel, and S. Levine · 2017
Cited alongside, same era.
Deep predictive policy training using reinforcement learning
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman · 2017
Cited alongside, same era.
Meta-learning with temporal convolutions
N. Mishra, M. Rohaninejad, X. Chen, and P. Abbeel · 2017
Cited alongside, same era.
One-shot visual imitation learning via meta-learning
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine · 2017
Cited alongside, same era.
Qmdp-net: Deep learning for planning under partial observability
P. Karkus, D. Hsu, and W. S. Lee · 2017
Cited alongside, same era.
Solar: Deep structured representations for model-based reinforcement learning
M. Zhang, S. Vikram, L. Smith, P. Abbeel, M. Johnson, and S. Levine · 2019
Later among the works it cites.
Deep reinforcement learning for industrial insertion tasks with visual inputs and natural reward signals
G. Schoettler, A. Nair, J. Luo, S. Bahl, J. A. Ojea, E. Solowjow, and S. Levine · 2019
Later among the works it cites.
R. Mendonca, A. Gupta, R. Kralev, P. Abbeel, S. Levine, and C. Finn · 2019
Later among the works it cites.
Varibad: A very good method for bayes-adaptive deep rl via meta-learning
L. Zintgraf, K. Shiarlis, M. Igl, S. Schulze, Y. Gal, K. Hofmann, and S. Whiteson · 2019
Later among the works it cites.
Meta reinforcement learning as task inference
J. Humplik, A. Galashov, L. Hasenclever, P. A. Ortega, Y. W. Teh, and N. Heess · 2019
Later among the works it cites.
Learning one-shot imitation from humans without humans
A. Bonardi, S. James, and A. J. Davison · 2019
Later among the works it cites.
Meta reinforcement learning for sim-to-real domain adaptation
K. Arndt, M. Hazara, A. Ghadirzadeh, and V. Kyrki · 2019
Later among the works it cites.
Contextual reinforcement learning of visuo-tactile multi-fingered grasping policies
V. Kumar, T. Herman, D. Fox, S. Birchfield, and J. Tremblay · 2019
Later among the works it cites.
Self-supervised correspondence in visuomotor policy learning
P. Florence, L. Manuelli, and R. Tedrake · 2019
Later among the works it cites.
Improving sample efficiency in model-free reinforcement learning from images
D. Yarats, A. Zhang, I. Kostrikov, B. Amos, J. Pineau, and R. Fergus · 2019
Later among the works it cites.
Mid-level visual representations improve generalization and sample efficiency for learning visuomotor policies
A. Sax, B. Emi, A. R. Zamir, L. J. Guibas, S. Savarese, and J. Malik · 2019
Later among the works it cites.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Later among the works it cites.
Deepmdp: Learning continuous latent space models for representation learning
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare · 2019
Later among the works it cites.
Generalized hidden parameter mdps transferable model-based rl in a handful of trials
C. F. Perez, F. P. Such, and T. Karaletsos · 2020
Closest in time.
Rapidly adaptable legged robots via evolutionary meta-learning
X. Song, Y. Yang, K. Choromanski, K. Caluwaerts, W. Gao, C. Finn, and J. Tan · 2020
Closest in time.
End-to-end robotic reinforcement learning without reward engineering
A. Singh, L. Yang, K. Hartikainen, C. Finn, and S. Levine · 2020
Closest in time.