Fetching the paper…
Reading the bibliography…
By providing a simple and efficient way of computing low-variance gradients of continuous random variables, the reparameterization trick has become the technique of choice for training a variety of latent variable models.
“Über die “Ganzzahligkeit” der Atomgewicht und verwandte Fragen.”
Richard von Mises · 1918
Earlier work this paper cites.
“Über die “Ganzzahligkeit” der Atomgewicht und verwandte Fragen.”
Richard von Mises · 1918
Earlier work this paper cites.
“Algorithm AS 32: The Incomplete Gamma Integral”
G.. Bhattacharjee · 1970
Earlier work this paper cites.
“Algorithm AS 32: The Incomplete Gamma Integral”
G.. Bhattacharjee · 1970
Earlier work this paper cites.
“Algorithm AS 32: The Incomplete Gamma Integral”
G.. Bhattacharjee · 1970
Earlier work this paper cites.
“Algorithm AS 32: The Incomplete Gamma Integral”
G.. Bhattacharjee · 1970
Earlier work this paper cites.
“Algorithm 518: Incomplete Bessel Function I 0 I_{0} . The Von Mises Distribution”
Geoffrey Hill · 1977
Earlier work this paper cites.
“Algorithm 518: Incomplete Bessel Function I 0 I_{0} . The Von Mises Distribution”
Geoffrey Hill · 1977
Earlier work this paper cites.
“Algorithm 518: Incomplete Bessel Function I 0 I_{0} . The Von Mises Distribution”
Geoffrey Hill · 1977
Earlier work this paper cites.
“Algorithm 518: Incomplete Bessel Function I 0 I_{0} . The Von Mises Distribution”
Geoffrey Hill · 1977
Earlier work this paper cites.
“Doubly stochastic variational Bayes for non-conjugate inference”
Michalis Titsias and Miguel Lázaro-Gredilla · 1979
Earlier work this paper cites.
“Efficient simulation of the von Mises distribution”
DJ Best and Nicholas Fisher · 1979
Earlier work this paper cites.
“Doubly stochastic variational Bayes for non-conjugate inference”
Michalis Titsias and Miguel Lázaro-Gredilla · 1979
Earlier work this paper cites.
“Efficient simulation of the von Mises distribution”
DJ Best and Nicholas Fisher · 1979
Earlier work this paper cites.
“Rethinking LDA: Why priors matter”
Hanna Wallach, David Mimno and Andrew McCallum · 1981
Earlier work this paper cites.
“Rethinking LDA: Why priors matter”
Hanna Wallach, David Mimno and Andrew McCallum · 1981
Earlier work this paper cites.
“Algorithm AS 187: Derivatives of the incomplete gamma integral”
RJ Moore · 1982
Earlier work this paper cites.
“Algorithm AS 187: Derivatives of the incomplete gamma integral”
RJ Moore · 1982
Earlier work this paper cites.
“Non-Uniform Random Variate Generation”
Luc Devroye · 1986
Earlier work this paper cites.
“Non-Uniform Random Variate Generation”
Luc Devroye · 1986
Earlier work this paper cites.
“Perturbation analysis gives strongly consistent sensitivity estimates for the M/G/1 queue”
Rajan Suri and Michael Zazanis · 1988
Earlier work this paper cites.
“Perturbation analysis gives strongly consistent sensitivity estimates for the M/G/1 queue”
Rajan Suri and Michael Zazanis · 1988
Earlier work this paper cites.
“Likelihood ratio gradient estimation for stochastic systems”
Peter Glynn · 1990
Earlier work this paper cites.
“Likelihood ratio gradient estimation for stochastic systems”
Peter Glynn · 1990
Earlier work this paper cites.
“Simple statistical gradient-following algorithms for connectionist reinforcement learning”
Ronald Williams · 1992
Earlier work this paper cites.
“Simple statistical gradient-following algorithms for connectionist reinforcement learning”
Ronald Williams · 1992
Earlier work this paper cites.
“Bayesian parameter estimation via variational methods”
Tommi Jaakkola and Michael Jordan · 2000
Earlier work this paper cites.
“A simple method for generating gamma variables”
George Marsaglia and Wai Tsang · 2000
Earlier work this paper cites.
“Cephes math library”, 2000
SL Moshier · 2000
Earlier work this paper cites.
“Bayesian parameter estimation via variational methods”
Tommi Jaakkola and Michael Jordan · 2000
Earlier work this paper cites.
“A simple method for generating gamma variables”
George Marsaglia and Wai Tsang · 2000
Earlier work this paper cites.
“Cephes math library”, 2000
SL Moshier · 2000
Earlier work this paper cites.
“SciPy: Open source scientific tools for Python”, 2001
Eric Jones, Travis Oliphant and Pearu Peterson · 2001
Earlier work this paper cites.
“Derivative of GammaRegularized with respect to a a ”, 2001
The Wolfram Functions Site · 2001
Earlier work this paper cites.
“SciPy: Open source scientific tools for Python”, 2001
Eric Jones, Travis Oliphant and Pearu Peterson · 2001
Earlier work this paper cites.
“Derivative of GammaRegularized with respect to a a ”, 2001
The Wolfram Functions Site · 2001
Earlier work this paper cites.
“Latent dirichlet allocation”
David Blei, Andrew Ng and Michael Jordan · 2003
Earlier work this paper cites.
“Latent dirichlet allocation”
David Blei, Andrew Ng and Michael Jordan · 2003
Earlier work this paper cites.
“Rcv1: A new benchmark collection for text categorization research”
David Lewis, Yiming Yang, Tony Rose and Fan Li · 2004
Earlier work this paper cites.
“Rcv1: A new benchmark collection for text categorization research”
David Lewis, Yiming Yang, Tony Rose and Fan Li · 2004
Earlier work this paper cites.
“Correlated topic models”
David Blei and John Lafferty · 2005
Earlier work this paper cites.
“Correlated topic models”
David Blei and John Lafferty · 2005
Earlier work this paper cites.
“Dynamic topic models”
David Blei and John Lafferty · 2006
Earlier work this paper cites.
“Gradient estimation”
Michael Fu · 2006
Earlier work this paper cites.
“Dynamic topic models”
David Blei and John Lafferty · 2006
Earlier work this paper cites.
“Gradient estimation”
Michael Fu · 2006
Earlier work this paper cites.
“A collapsed variational Bayesian inference algorithm for latent Dirichlet allocation”
Yee Teh, David Newman and Max Welling · 2007
Earlier work this paper cites.
“Numerical Recipes 3rd Edition: The Art of Scientific Computing”
William Press, Saul Teukolsky, William Vetterling and Brian Flannery · 2007
Cited alongside, same era.
“A collapsed variational Bayesian inference algorithm for latent Dirichlet allocation”
Yee Teh, David Newman and Max Welling · 2007
Cited alongside, same era.
“Numerical Recipes 3rd Edition: The Art of Scientific Computing”
William Press, Saul Teukolsky, William Vetterling and Brian Flannery · 2007
Cited alongside, same era.
“Directional statistics”
Kanti Mardia and Peter Jupp · 2009
Cited alongside, same era.
“Directional statistics”
Kanti Mardia and Peter Jupp · 2009
Cited alongside, same era.
“Online learning for latent dirichlet allocation”
Matthew Hoffman, Francis Bach and David Blei · 2010
Cited alongside, same era.
“The Generalized Reparameterization Gradient”
Francisco Ruiz, Michalis Titsias and David Blei · 2016
Later among the works it cites.
“TensorFlow: A System for Large-Scale Machine Learning.”
Martı́n Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving and Michael Isard · 2016
Later among the works it cites.
“Importance weighted autoencoders”
Yuri Burda, Roger Grosse and Ruslan Salakhutdinov · 2016
Later among the works it cites.
“A theoretically grounded application of dropout in recurrent neural networks”
Yarin Gal and Zoubin Ghahramani · 2016
Later among the works it cites.
“Dropout as a Bayesian approximation: Representing model uncertainty in deep learning”
Yarin Gal and Zoubin Ghahramani · 2016
Later among the works it cites.
“Stochastic backpropagation through mixture density distributions”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Understanding the difficulty of training deep feedforward neural networks”
Xavier Glorot and Yoshua Bengio · 2010
Cited alongside, same era.
“Eigen v3”, 2010
Gaël Guennebaud and Benoît Jacob · 2010
Cited alongside, same era.
“Software Framework for Topic Modelling with Large Corpora”
Radim Řehůřek and Petr Sojka · 2010
Cited alongside, same era.
“Online learning for latent dirichlet allocation”
Matthew Hoffman, Francis Bach and David Blei · 2010
Cited alongside, same era.
“Understanding the difficulty of training deep feedforward neural networks”
Xavier Glorot and Yoshua Bengio · 2010
Cited alongside, same era.
“Eigen v3”, 2010
Gaël Guennebaud and Benoît Jacob · 2010
Cited alongside, same era.
Alex Graves · 2016
Later among the works it cites.
“Approximate inference for deep latent gaussian mixtures”
Eric Nalisnick, Lars Hertel and Padhraic Smyth · 2016
Later among the works it cites.
“The Generalized Reparameterization Gradient”
Francisco Ruiz, Michalis Titsias and David Blei · 2016
Later among the works it cites.
“TensorFlow Distributions”
Joshua Dillon, Ian Langmore, Dustin Tran, Eugene Brevdo, Srinivas Vasudevan, Dave Moore, Brian Patton, Alex Alemi, Matt Hoffman and Rif Saurous · 2017
Later among the works it cites.
“Categorical reparameterization with gumbel-softmax”
Eric Jang, Shixiang Gu and Ben Poole · 2017
Later among the works it cites.
“Automatic differentiation variational inference”
Alp Kucukelbir, Dustin Tran, Rajesh Ranganath, Andrew Gelman and David Blei · 2017
Later among the works it cites.
“The concrete distribution: A continuous relaxation of discrete random variables”
Chris Maddison, Andriy Mnih and Yee Teh · 2017
Later among the works it cites.
“Variational dropout sparsifies deep neural networks”
Dmitry Molchanov, Arsenii Ashukha and Dmitry Vetrov · 2017
Later among the works it cites.
“Reparameterization gradients through acceptance-rejection sampling algorithms”
Christian Naesseth, Francisco Ruiz, Scott Linderman and David Blei · 2017
Later among the works it cites.
“Stick-breaking variational autoencoders”
Eric Nalisnick and Padhraic Smyth · 2017
Later among the works it cites.
“Automatic differentiation in PyTorch”
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga and Adam Lerer · 2017
Later among the works it cites.
“Sticking the landing: An asymptotically zero-variance gradient estimator for variational inference”
Geoffrey Roeder, Yuhuai Wu and David Duvenaud · 2017
Later among the works it cites.
“Autoencoding variational inference for topic models”
Akash Srivastava and Charles Sutton · 2017
Later among the works it cites.
“mpmath: a Python library for arbitrary-precision floating-point arithmetic (version 1.0.0)” http://mpmath.org/
Fredrik Johansson · 2017
Later among the works it cites.
“Reparameterization gradients through acceptance-rejection sampling algorithms”
Christian Naesseth, Francisco Ruiz, Scott Linderman and David Blei · 2017
Later among the works it cites.
“Autoencoding variational inference for topic models”
Akash Srivastava and Charles Sutton · 2017
Later among the works it cites.
“TensorFlow Distributions”
Joshua Dillon, Ian Langmore, Dustin Tran, Eugene Brevdo, Srinivas Vasudevan, Dave Moore, Brian Patton, Alex Alemi, Matt Hoffman and Rif Saurous · 2017
Later among the works it cites.
“Categorical reparameterization with gumbel-softmax”
Eric Jang, Shixiang Gu and Ben Poole · 2017
Later among the works it cites.
“Automatic differentiation variational inference”
Alp Kucukelbir, Dustin Tran, Rajesh Ranganath, Andrew Gelman and David Blei · 2017
Later among the works it cites.
“The concrete distribution: A continuous relaxation of discrete random variables”
Chris Maddison, Andriy Mnih and Yee Teh · 2017
Later among the works it cites.
“Variational dropout sparsifies deep neural networks”
Dmitry Molchanov, Arsenii Ashukha and Dmitry Vetrov · 2017
Later among the works it cites.
“Reparameterization gradients through acceptance-rejection sampling algorithms”
Christian Naesseth, Francisco Ruiz, Scott Linderman and David Blei · 2017
Later among the works it cites.
“Stick-breaking variational autoencoders”
Eric Nalisnick and Padhraic Smyth · 2017
Later among the works it cites.
“Automatic differentiation in PyTorch”
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga and Adam Lerer · 2017
Later among the works it cites.
“Sticking the landing: An asymptotically zero-variance gradient estimator for variational inference”
Geoffrey Roeder, Yuhuai Wu and David Duvenaud · 2017
Later among the works it cites.
“Autoencoding variational inference for topic models”
Akash Srivastava and Charles Sutton · 2017
Later among the works it cites.
“mpmath: a Python library for arbitrary-precision floating-point arithmetic (version 1.0.0)” http://mpmath.org/
Fredrik Johansson · 2017
Later among the works it cites.
“Reparameterization gradients through acceptance-rejection sampling algorithms”
Christian Naesseth, Francisco Ruiz, Scott Linderman and David Blei · 2017
Later among the works it cites.
“Autoencoding variational inference for topic models”
Akash Srivastava and Charles Sutton · 2017
Later among the works it cites.
“Hyperspherical Variational Auto-Encoders”
Tim Davidson, Luca Falorsi, Nicola De, Thomas Kipf and Jakub Tomczak · 2018
Closest in time.
“Pathwise Derivatives for Multivariate Distributions”
Martin Jankowiak and Theofanis Karaletsos · 2018
Closest in time.
“Pathwise Derivatives Beyond the Reparameterization Trick”
Martin Jankowiak and Fritz Obermeyer · 2018
Closest in time.
“Variational Inference In Pachinko Allocation Machines”
Akash Srivastava and Charles Sutton · 2018
Closest in time.
“WHAI: Weibull Hybrid Autoencoding Inference for Deep Topic Modeling”
Hao Zhang, Bo Chen, Dandan Guo and Mingyuan Zhou · 2018
Closest in time.
“Hyperspherical Variational Auto-Encoders”
Tim Davidson, Luca Falorsi, Nicola De, Thomas Kipf and Jakub Tomczak · 2018
Closest in time.
“Hyperspherical Variational Auto-Encoders”
Tim Davidson, Luca Falorsi, Nicola De, Thomas Kipf and Jakub Tomczak · 2018
Closest in time.
“Pathwise Derivatives for Multivariate Distributions”
Martin Jankowiak and Theofanis Karaletsos · 2018
Closest in time.
“Pathwise Derivatives Beyond the Reparameterization Trick”
Martin Jankowiak and Fritz Obermeyer · 2018
Closest in time.
“Variational Inference In Pachinko Allocation Machines”
Akash Srivastava and Charles Sutton · 2018
Closest in time.
“WHAI: Weibull Hybrid Autoencoding Inference for Deep Topic Modeling”
Hao Zhang, Bo Chen, Dandan Guo and Mingyuan Zhou · 2018
Closest in time.
“Hyperspherical Variational Auto-Encoders”
Tim Davidson, Luca Falorsi, Nicola De, Thomas Kipf and Jakub Tomczak · 2018
Closest in time.