Fetching the paper…
Reading the bibliography…
The reparameterization trick enables optimizing large scale stochastic computation graphs via gradient descent.
Statistical theory of extreme values and some practical applications: a series of lectures
Emil Julius Gumbel · 1954
Earlier work this paper cites.
Individual Choice Behavior: A Theoretical Analysis
R. Duncan Luce · 1959
Earlier work this paper cites.
Concepts of independence for proportions with a generalization of the dirichlet distribution
Robert J Connor and James E Mosimann · 1969
Earlier work this paper cites.
The relationship between luce’s choice axiom, thurstone’s theory of comparative judgment, and the double exponential distribution
John I Yellott · 1977
Earlier work this paper cites.
Logistic-normal distributions: Some properties and uses
J Atchison and Sheng M Shen · 1980
Earlier work this paper cites.
A general class of distributions on the simplex
J Aitchison · 1985
Earlier work this paper cites.
Likelihood ratio gradient estimation for stochastic systems
Peter W Glynn · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Dependence properties of generalized liouville distributions on the simplex
William S Rayens and Cidambi Srinivasan · 1994
Earlier work this paper cites.
Continuous sigmoidal belief networks trained using slice sampling
Brendan Frey · 1997
Earlier work this paper cites.
Variance reduction techniques for gradient estimates in reinforcement learning
Evan Greensmith, Peter L. Bartlett, and Jonathan Baxter · 2004
Earlier work this paper cites.
Correlated topic models
David Blei and John Lafferty · 2006
Earlier work this paper cites.
Gradient estimation
Michael C Fu · 2006
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
Ruslan Salakhutdinov and Iain Murray · 2008
Earlier work this paper cites.
Semantic hashing
Ruslan Salakhutdinov and Geoffrey Hinton · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
On a class of distributions on the simplex
Stefano Favaro, Georgia Hadjicharalambous, and Igor Prünster · 2011
Earlier work this paper cites.
Perturb-and-map random fields: Using discrete optimization to learn and sample from energy models
George Papandreou and Alan L Yuille · 2011
Cited alongside, same era.
On the partition function and random maximum a-posteriori perturbations
Tamir Hazan and Tommi Jaakkola · 2012
Cited alongside, same era.
Variational bayesian inference with stochastic search
John William Paisley, David M. Blei, and Michael I. Jordan · 2012
Cited alongside, same era.
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio, Nicholas Léonard, and Aaron Courville · 2013
Cited alongside, same era.
Karol Gregor, Ivo Danihelka, Andriy Mnih, Charles Blundell, and Daan Wierstra · 2013
Cited alongside, same era.
Learning to transduce with unbounded memory
Edward Grefenstette, Karl Moritz Hermann, Mustafa Suleyman, and Phil Blunsom · 2015
Later among the works it cites.
Draw: A recurrent neural network for image generation
Karol Gregor, Ivo Danihelka, Alex Graves, Danilo Jimenez Rezende, and Daan Wierstra · 2015
Later among the works it cites.
Gradient estimation using stochastic computation graphs
John Schulman, Nicolas Heess, Theophane Weber, and Pieter Abbeel · 2015
Later among the works it cites.
Local expectation gradients for black box variational inference
Michalis Titsias and Miguel Lázaro-Gredilla · 2015
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stochastic variational inference
Matthew D Hoffman, David M Blei, Chong Wang, and John William Paisley · 2013
Cited alongside, same era.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2014
Cited alongside, same era.
A ∗ Sampling
Chris J Maddison, Daniel Tarlow, and Tom Minka · 2014
Cited alongside, same era.
Neural variational inference and learning in belief networks
Andriy Mnih and Karol Gregor · 2014
Cited alongside, same era.
Recurrent Models of Visual Attention
Volodymyr Mnih, Nicolas Heess, Alex Graves, and koray kavukcuoglu · 2014
Cited alongside, same era.
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov · 2016
Closest in time.
Hybrid computing using a neural network with dynamic external memory
Alex Graves, Greg Wayne, Malcolm Reynolds, Tim Harley, Ivo Danihelka, Agnieszka Grabska-Barwińska, Sergio Gómez Colmenarejo, Edward Grefenstette, Tiago Ramalho, John Agapiou, et al · 2016
Closest in time.
MuProp: Unbiased backpropagation for stochastic neural networks
Shixiang Gu, Sergey Levine, Ilya Sutskever, and Andriy Mnih · 2016
Closest in time.
Perturbation, Optimization, and Statistics
Tamir Hazan, George Papandreou, and Daniel Tarlow · 2016
Closest in time.
Categorical Reparameterization with Gumbel-Softmax
E. Jang, S. Gu, and B. Poole · 2016
Closest in time.
Semantic parsing with semi-supervised sequential autoencoders
Tomáš Kočiský, Gábor Melis, Edward Grefenstette, Chris Dyer, Wang Ling, Phil Blunsom, and Karl Moritz Hermann · 2016
Closest in time.
A Poisson process model for Monte Carlo
Chris J Maddison · 2016
Closest in time.
Variational inference for monte carlo objectives
Andriy Mnih and Danilo Jimenez Rezende · 2016
Closest in time.
Rejection sampling variational inference
Christian A Naesseth, Francisco JR Ruiz, Scott W Linderman, and David M Blei · 2016
Closest in time.
The generalized reparameterization gradient
Francisco JR Ruiz, Michalis K Titsias, and David M Blei · 2016
Closest in time.
Theano: A Python framework for fast computation of mathematical expressions
Theano Development Team · 2016
Closest in time.