Fetching the paper…
Reading the bibliography…
We wish to compute the gradient of an expectation over a finite or countably infinite sample space having $K \leq \infty$ categories.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J · 1992
Earlier work this paper cites.
Rao-Blackwellisation of sampling schemes
Casella, G. and Robert, C. P · 1996
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Lecun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Introduction to Stochastic Search and Optimization
Spall, J. C · 2003
Earlier work this paper cites.
N-mixture models for estimating population size from spatially replicated counts
Royle, J. A · 2004
Earlier work this paper cites.
Hierarchical probabilistic neural network language model
Morin, F. and Bengio, Y · 2005
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Gutmann, M. and Hyvärinen, A · 2012
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Leonard, N., and Courville, A · 2013
Earlier work this paper cites.
Deep autoregressive networks
Gregor, K., Mnih, A., and Wierstra, D · 2014
Earlier work this paper cites.
Auto-encoding variational Bayes
Kingma, D. and Welling, M · 2014
Cited alongside, same era.
Semi-supervised learning with deep generative models
Kingma, D. P., Rezende, D. J., Mohamed, S., and Welling, M · 2014
Cited alongside, same era.
Neural variational inference and learning in belief networks
Mnih, A. and Gregor, K · 2014
Cited alongside, same era.
Recurrent models of visual attention
Mnih, V., Heess, N., Graves, A., et al · 2014
Cited alongside, same era.
Black box variational inference
Ranganath, R., Gerrish, S., and Blei, D. M · 2014
Cited alongside, same era.
Combine Monte Carlo with exhaustive search: Effective variational inference and policy gradient reinforcement learning
Titsias K, M · 2014
Cited alongside, same era.
MuProp: Unbiased backpropagation for stochastic neural networks
Gu, S., Levine, S., Sutskever, I., and Mnih, A · 2016
Later among the works it cites.
Variational inference for Monte Carlo objectives
Mnih, A. and Rezende, D. J · 2016
Later among the works it cites.
Variational inference: A review for statisticians
Blei, D. M., Kucukelbir, A., and McAuliffe, J. D · 2017
Later among the works it cites.
Categorical reparameterization with Gumbel-softmax
Jang, E., Gu, S., and Poole, B · 2017
Later among the works it cites.
The concrete distribution: A continuous relaxation of discrete random variables
Maddison, C. J., Mnih, A., and Teh, Y. W · 2017
Later among the works it cites.
REBAR: Low-variance, unbiased gradient estimates for discrete latent variable models
Tucker, G., Mnih, A., Maddison, C. J., Lawson, J., and Sohl-Dickstein, J · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
DRAW: a recurrent neural network for image generation
Gregor, K., Danihelka, I., Graves, A., Rezende, D., and Wierstra, D · 2015
Cited alongside, same era.
Adam: a method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Cited alongside, same era.
Local expectation gradients for black box variational inference
Titsias K, M. and Lázaro-Gredilla, M · 2015
Cited alongside, same era.
Later among the works it cites.
Backpropagation through the void: Optimizing control variates for black-box gradient estimation
Grathwohl, W., Choi, D., Wu, Y., Roeder, G., and Duvenaud, D · 2018
Closest in time.
Memory augmented policy optimization for program synthesis with generalization
Liang, C., Norouzi, M., Berant, J., Le, Q., and Lao, N · 2018
Closest in time.