Fetching the paper…
Reading the bibliography…
Efficiently adapting to new environments and changes in dynamics is critical for agents to successfully operate in the real world.
Evolutionary principles in self-referential learning, or on learning how to learn: the meta-meta-… hook
Jürgen Schmidhuber. 1987 · 1987
Earlier work this paper cites.
Adaptive inverse control. In Applications of Artificial Neural Networks
Bernard Widrow. 1990 · 1990
Earlier work this paper cites.
On the optimization of a synaptic learning rule. In Preprints Conf. Optimality in Artificial and Biological Neural Networks
Samy Bengio, Yoshua Bengio, Jocelyn Cloutier, and Jan Gecsei. 1992 · 1992
Earlier work this paper cites.
System identification
Lennart Ljung. 1998 · 1998
Earlier work this paper cites.
Learning to learn
Sebastian Thrun and Lorien Pratt. 1998 · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation. In Advances in neural information processing systems
Richard S Sutton, David A McAllester, Satinder P Singh, and Yishay Mansour. 2000 · 2000
Earlier work this paper cites.
Learning to learn using gradient descent. In International Conference on Artificial Neural Networks
Sepp Hochreiter, A Steven Younger, and Peter R Conwell. 2001 · 2001
Earlier work this paper cites.
Resilient machines through continuous self-modeling
Josh Bongard, Victor Zykov, and Hod Lipson. 2006 · 2006
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Robots that can adapt like animals
Antoine Cully, Jeff Clune, Danesh Tarapore, and Jean-Baptiste Mouret. 2015 · 2015
Earlier work this paper cites.
Ensemble-CIO: Full-body dynamic motion planning that transfers to physical humanoids. In Intelligent Robots and Systems (IROS), 2015 IEEE/RSJ International Conference on
Igor Mordatch, Kendall Lowrey, and Emanuel Todorov. 2015 · 2015
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Cited alongside, same era.
RL2: Fast Reinforcement Learning via Slow Reinforcement Learning
Yan Duan, John Schulman, Xi Chen, Peter L Bartlett, Ilya Sutskever, and Pieter Abbeel. 2016b · 2016
Cited alongside, same era.
One-shot learning of manipulation skills with online dynamics adaptation and neural network priors. In International Conference on Intelligent Robots and Systems (IROS)
Justin Fu, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Identity mappings in deep residual networks. In European conference on computer vision
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks
Learning to reinforcement learn
Jane X Wang, Zeb Kurth-Nelson, Dhruva Tirumala, Hubert Soyer, Joel Z Leibo, Remi Munos, Charles Blundell, Dharshan Kumaran, and Matt Botvinick. 2017 · 2017
Later among the works it cites.
Preparing for the unknown: Learning a universal policy with online system identification
Wenhao Yu, Jie Tan, C Karen Liu, and Greg Turk. 2017 · 2017
Later among the works it cites.
Learning to Adapt: Meta-Learning for Model-Based Control
Ignasi Clavera, Anusha Nagabandi, Ronald S Fearing, Pieter Abbeel, Sergey Levine, and Chelsea Finn. 2018a · 2018
Later among the works it cites.
Learning to Learn with Gradients
Chelsea Finn. 2018 · 2018
Later among the works it cites.
Meta-Reinforcement Learning of Structured Exploration Strategies
Abhishek Gupta, Russell Mendonca, YuXuan Liu, Pieter Abbeel, and Sergey Levine. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Meta-SGD: Learning to Learn Quickly for Few Shot Learning
Zhenguo Li, Fengwei Zhou, Fei Chen, and Hang Li. 2017 · 2017
Cited alongside, same era.
Epopt: Learning robust neural network policies using model ensembles
Aravind Rajeswaran, Sarvjeet Ghotra, Balaraman Ravindran, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
CAD2RL: Real single-image flight without a single real image
Fereshteh Sadeghi and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Proximal Policy Optimization Algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Learning to learn: Meta-critic networks for sample efficient learning
Flood Sung, Li Zhang, Tao Xiang, Timothy Hospedales, and Yongxin Yang. 2017 · 2017
Cited alongside, same era.
Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World
Joshua Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Model-Based Reinforcement Learning via Meta-Policy Optimization. In Proceedings of The 2nd Conference on Robot Learning
Ignasi Clavera, Jonas Rothfuss, John Schulman, Yasuhiro Fujita, Tamim Asfour, and Pieter Abbeel. 2018b
Cited in the paper.
Later among the works it cites.
Rein Houthooft, Richard Y Chen, Phillip Isola, Bradly C Stadie, Filip Wolski, Jonathan Ho, and Pieter Abbeel. 2018 · 2018
Later among the works it cites.
A simple neural attentive meta-learner
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen, and Pieter Abbeel. 2018 · 2018
Later among the works it cites.
Sim-to-real transfer of robotic control with dynamics randomization. In International Conference on Robotics and Automation (ICRA)
Xue Bin Peng, Marcin Andrychowicz, Wojciech Zaremba, and Pieter Abbeel. 2018 · 2018
Later among the works it cites.
Meta Reinforcement Learning with Latent Variable Gaussian Processes
Steindór Sæmundsson, Katja Hofmann, and Marc Peter Deisenroth. 2018 · 2018
Later among the works it cites.
Some considerations on learning to explore via meta-reinforcement learning
Bradly C Stadie, Ge Yang, Rein Houthooft, Xi Chen, Yan Duan, Yuhuai Wu, Pieter Abbeel, and Ilya Sutskever. 2018 · 2018
Later among the works it cites.
Sim-to-Real: Learning Agile Locomotion For Quadruped Robots
Jie Tan, Tingnan Zhang, Erwin Coumans, Atil Iscen, Yunfei Bai, Danijar Hafner, Steven Bohez, and Vincent Vanhoucke. 2018 · 2018
Later among the works it cites.