Fetching the paper…

On the model-based stochastic value gradient for continuous reinforcement learning · Around