Fetching the paper…
Reading the bibliography…
Training models with discrete latent variables is challenging due to the difficulty of estimating the gradients accurately.
Likelihood ratio gradient estimation for stochastic systems
Glynn, P. W. (1990) · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J. (1992) · 1992
Earlier work this paper cites.
An introduction to variational methods for graphical models
Jordan, M. I., Ghahramani, Z., Jaakkola, T. S., and Saul, L. K. (1999) · 1999
Earlier work this paper cites.
Gradient estimation
Fu, M. C. (2006) · 2006
Earlier work this paper cites.
Variational bayesian inference with stochastic search
Paisley, J., Blei, D. M., and Jordan, M. I. (2012) · 2012
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Léonard, N., and Courville, A. (2013) · 2013
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Maas, A. L., Hannun, A. Y., and Ng, A. Y. (2013) · 2013
Earlier work this paper cites.
Monte Carlo theory, methods and examples
Owen, A. B. (2013) · 2013
Earlier work this paper cites.
Automated variational inference in probabilistic programming
Wingate, D. and Weber, T. (2013) · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M. (2014) · 2014
Earlier work this paper cites.
Neural variational inference and learning in belief networks
Mnih, A. and Gregor, K. (2014) · 2014
Cited alongside, same era.
Black box variational inference
Ranganath, R., Gerrish, S., and Blei, D. M. (2014) · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D. (2014) · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. and Ba, J. (2015) · 2015
Cited alongside, same era.
Local expectation gradients for black box variational inference
Titsias, M. K. and Lázaro-Gredilla, M. (2015) · 2015
Cited alongside, same era.
Stochastic gradient estimation with finite differences
Buesing, L., Weber, T., and Mohamed, S. (2016) · 2016
Cited alongside, same era.
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
Maddison, C. J., Mnih, A., and Teh, Y. W. (2017) · 2017
Later among the works it cites.
REBAR: Low-variance, unbiased gradient estimates for discrete latent variable models
Tucker, G., Mnih, A., Maddison, C. J., Lawson, J., and Sohl-Dickstein, J. (2017) · 2017
Later among the works it cites.
Backpropagation through the void: Optimizing control variates for black-box gradient estimation
Grathwohl, W., Choi, D., Wu, Y., Roeder, G., and Duvenaud, D. (2018) · 2018
Later among the works it cites.
Buy 4 reinforce samples, get a baseline for free!
Kool, W., van Hoof, H., and Welling, M. (2019) · 2019
Later among the works it cites.
Adaptive antithetic sampling for variance reduction
Ren, H., Zhao, S., and Ermon, S. (2019) · 2019
Later among the works it cites.
Differentiable antithetic sampling for variance reduction in stochastic variational inference
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Importance weighted autoencoders
Burda, Y., Grosse, R., and Salakhutdinov, R. (2016) · 2016
Cited alongside, same era.
MuProp: Unbiased backpropagation for stochastic neural networks
Gu, S., Levine, S., Sutskever, I., and Mnih, A. (2016) · 2016
Cited alongside, same era.
Variational inference for monte carlo objectives
Mnih, A. and Rezende, D. (2016) · 2016
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Jang, E., Gu, S., and Poole, B. (2017) · 2017
Cited alongside, same era.
Wu, M., Goodman, N., and Ermon, S. (2019) · 2019
Later among the works it cites.
ARSM: Augment-REINFORCE-swap-merge estimator for gradient backpropagation through categorical variables
Yin, M., Yue, Y., and Zhou, M. (2019) · 2019
Later among the works it cites.
ARM: Augment-REINFORCE-merge gradient for stochastic binary networks
Yin, M. and Zhou, M. (2019) · 2019
Later among the works it cites.
Probabilistic Best Subset Selection by Gradient-Based Optimization
Yin, M., Ho, N., Yan, B., Qian, X., and Zhou, M. (2020) · 2020
Closest in time.