Gradient estimation using stochastic computation graphs
Schulman, J., Heess, N., Weber, T., and Abbeel, P · 2015
Later among the works it cites.
Learning and policy search in stochastic dynamical systems with Bayesian neural networks
Original
Depeweg, S., Hernández-Lobato, J. M., Doshi-Velez, F., and Udluft, S · 2016
Later among the works it cites.
Improving PILCO with Bayesian neural network dynamics models
Gal, Y., McAllister, R., and Rasmussen, C. E · 2016
Later among the works it cites.
Variational inference for Monte Carlo objectives
Mnih, A. and Rezende, D · 2016
Later among the works it cites.
Differentiation of the Cholesky decomposition
Original
Murray, I · 2016
Later among the works it cites.
Exponential expressivity in deep neural networks through transient chaos
Poole, B., Lahiri, S., Raghu, M., Sohl-Dickstein, J., and Ganguli, S · 2016
Later among the works it cites.
The generalized reparameterization gradient
Ruiz, F. R., AUEB, M. T. R., and Blei, D · 2016
Later among the works it cites.
Stability of controllers for Gaussian process forward models
Vinogradska, J., Bischoff, B., Nguyen-Tuong, D., Romer, A., Schmidt, H., and Peters, J · 2016
Later among the works it cites.
Black-box data-efficient policy search for robotics
Chatzilygeroudis, K., Rama, R., Kaushik, R., Goepp, D., Vassiliades, V., and Mouret, J.-B · 2017
Later among the works it cites.