Value-aware loss function for model-based reinforcement learning
Amir-massoud Farahmand, Andre Barreto, and Daniel Nikovski · 2017
Cited alongside, same era.
Path integral networks: End-to-end differentiable optimal control
Original
Masashi Okada, Luca Rigazio, and Takenobu Aoshima · 2017
Cited alongside, same era.
Self-correcting models for model-based reinforcement learning
Erik Talvitie · 2017
Cited alongside, same era.
Information theoretic MPC for model-based reinforcement learning
Grady Williams, Nolan Wagener, Brian Goldfain, Paul Drews, James M Rehg, Byron Boots, and Evangelos A Theodorou · 2017
Cited alongside, same era.
Differentiable MPC for end-to-end planning and control
Brandon Amos, Ivan Jimenez, Jacob Sacks, Byron Boots, and J Zico Kolter · 2018
Cited alongside, same era.
Applied optimal control: optimization, estimation and control
Arthur Earl Bryson · 2018
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine · 2018
Cited alongside, same era.
Iterative value-aware model learning
Amir-massoud Farahmand · 2018
Cited alongside, same era.
Learning actionable representations with goal-conditioned policies
Original
Dibya Ghosh, Abhishek Gupta, and Sergey Levine · 2018
Cited alongside, same era.
Learning latent dynamics for planning from pixels
Original
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson · 2018
Cited alongside, same era.