Fetching the paper…
Reading the bibliography…
Gradient estimation in models with discrete latent variables is a challenging problem, because the simplest unbiased estimators tend to have high variance.
Conditional expectation and unbiased sequential estimation
David Blackwell · 1947
Earlier work this paper cites.
Likelihood ratio gradient estimation for stochastic systems
Peter W Glynn · 1990
Earlier work this paper cites.
Information and the accuracy attainable in the estimation of statistical parameters
C. Radhakrishna Rao · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
Ruslan Salakhutdinov and Iain Murray · 2008
Earlier work this paper cites.
MNIST handwritten digit database
Yann LeCun and Corinna Cortes · 2010
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio · 2012
Earlier work this paper cites.
Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Yoshua Bengio, Nicholas Léonard, and Aaron Courville · 2013
Earlier work this paper cites.
Low-rank approximations for conditional feedforward computation in deep neural networks
Andrew Davis and Itamar Arel · 2013
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Diederik P Kingma and Max Welling · 2013
Earlier work this paper cites.
Handbook of monte carlo methods , volume 706
Dirk P Kroese, Thomas Taimre, and Zdravko I Botev · 2013
Earlier work this paper cites.
A* Sampling
Chris J. Maddison, Daniel Tarlow, and Tom Minka · 2014
Earlier work this paper cites.
Neural variational inference and learning in belief networks
Andriy Mnih and Karol Gregor · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Cited alongside, same era.
Importance weighted autoencoders
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov · 2015
Cited alongside, same era.
Gradient estimation using stochastic computation graphs
John Schulman, Nicolas Heess, Theophane Weber, and Pieter Abbeel · 2015
Cited alongside, same era.
Muprop: Unbiased backpropagation for stochastic neural networks
Shixiang Gu, Sergey Levine, Ilya Sutskever, and Andriy Mnih · 2016
Cited alongside, same era.
A Poisson process model for Monte Carlo
Chris J. Maddison · 2016
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
Chris J. Maddison, Andriy Mnih, and Yee Whye Teh · 2017
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Igor Mordatch and Pieter Abbeel · 2017
Later among the works it cites.
REBAR : Low-variance, unbiased gradient estimates for discrete latent variable models
George Tucker, Andriy Mnih, Chris J. Maddison, and Jascha Sohl-Dickstein · 2017
Later among the works it cites.
Estimating means in a finite universe, 2017
Tim Vieira · 2017
Later among the works it cites.
Improved variational autoencoders for text modeling using dilated convolutions
Zichao Yang, Zhiting Hu, Ruslan Salakhutdinov, and Taylor Berg-Kirkpatrick · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Variational inference for monte carlo objectives
Andriy Mnih and Danilo J Rezende · 2016
Cited alongside, same era.
Unsupervised learning of task-specific tree structures with tree-lstms
Jihun Choi, Kang Min Yoo, and Sang-goo Lee · 2017
Cited alongside, same era.
Hierarchical multiscale recurrent neural networks
Junyoung Chung, Sungjin Ahn, and Yoshua Bengio · 2017
Cited alongside, same era.
Categorical Reparametrization with Gumble-Softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Backpropagation through the void: Optimizing control variates for black-box gradient estimation
Will Grathwohl, Dami Choi, Yuhuai Wu, Geoffrey Roeder, and David Duvenaud · 2018
Later among the works it cites.
Rao-blackwellized stochastic gradients for discrete distributions
Runjing Liu, Jeffrey Regier, Nilesh Tripuraneni, Michael I Jordan, and Jon McAuliffe · 2018
Later among the works it cites.
Listops: A diagnostic dataset for latent tree learning, 2018
Nikita Nangia and Samuel R. Bowman · 2018
Later among the works it cites.
Latent structure models for natural language processing
André F. T. Martins, Tsvetomila Mihaylova, Nikita Nangia, and Vlad Niculae · 2019
Later among the works it cites.
Credit assignment techniques in stochastic computation graphs
Théophane Weber, Nicolas Heess, Lars Buesing, and David Silver · 2019
Later among the works it cites.
Estimating gradients for discrete random variables by sampling without replacement
Wouter Kool, Herke van Hoof, and Max Welling · 2020
Closest in time.