Fetching the paper…
Reading the bibliography…
Traditional model-based RL relies on hand-specified or learned models of transition dynamics of the environment.
Mosaic model for sensorimotor learning and control
Masahiko Haruno, Daniel M Wolpert, and Mitsuo Kawato · 2001
Earlier work this paper cites.
A tutorial on the cross-entropy method
Pieter-Tjerk De Boer, Dirk P Kroese, Shie Mannor, and Reuven Y Rubinstein · 2005
Earlier work this paper cites.
PILCO: A Model-Based and Data-Efficient Approach to Policy Search
Marc P Deisenroth and Carl E. Rasmussen · 2011
Earlier work this paper cites.
A Survey on Policy Search for Robotics
Marc Peter Deisenroth · 2011
Earlier work this paper cites.
Model Predictive Control
E. F. Camacho and C. Bordons · 2013
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations
Finale Doshi-Velez and George Konidaris · 2016
Cited alongside, same era.
Improving PILCO with Bayesian Neural Network Dynamics Models
Yarin Gal, Rowan Thomas Mcallister, and Carl Edward Rasmussen · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Robust and efficient transfer learning with hidden parameter markov decision processes
Taylor W Killian, Samuel Daulton, George Konidaris, and Finale Doshi-Velez · 2017
Cited alongside, same era.
Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles
Balaji Lakshminarayanan, Alexander Pritzel, and Charles Blundell · 2017
Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine · 2018
Closest in time.
Learning to Adapt: Meta-Learning for Model-Based Control
Ignasi Clavera, Anusha Nagabandi, Ronald S. Fearing, Pieter Abbeel, Sergey Levine, and Chelsea Finn · 2018
Closest in time.
Learning an embedding space for transferable robot skills
Karol Hausman, Jost Tobias Springenberg, Ziyu Wang, Nicolas Heess, and Martin Riedmiller · 2018
Closest in time.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine · 2018
Closest in time.
Direct policy transfer via hidden parameter markov decision processes
Jiayu Yao, Taylor Killian, George Konidaris, and Finale Doshi-Velez · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.