Fetching the paper…
Reading the bibliography…
Probabilistic models with discrete latent variables naturally capture datasets composed of discrete classes.
Efficient estimation of free energy differences from Monte Carlo data
Charles H. Bennett · 1976
Earlier work this paper cites.
Information processing in dynamical systems: Foundations of harmony theory
Paul Smolensky · 1986
Earlier work this paper cites.
Replica Monte Carlo simulation of spin-glasses
Robert H. Swendsen and Jian-Sheng Wang · 1986
Earlier work this paper cites.
Probabilistic Reasoning in Intelligent Systems: Networks of Plausible Inference
Judea Pearl · 1988
Earlier work this paper cites.
Sequential updating of conditional probabilities on directed graphical structures
David J. Spiegelhalter and Steffen L. Lauritzen · 1990
Earlier work this paper cites.
Connectionist learning of belief networks
Radford M. Neal · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Approximating probabilistic inference in Bayesian belief networks is NP-hard
Paul Dagum and Michael Luby · 1993
Earlier work this paper cites.
Autoencoders, minimum description length, and Helmholtz free energy
Geoffrey E. Hinton and R. S. Zemel · 1994
Earlier work this paper cites.
Emergence of simple-cell receptive field properties by learning a sparse code for natural images
Bruno A. Olshausen and David J. Field · 1996
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
Michael I. Jordan, Zoubin Ghahramani, Tommi S. Jaakkola, and Lawrence K. Saul · 1999
Earlier work this paper cites.
Differentiation under the integral sign with weak derivatives
Steve Cheng · 2006
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Geoffrey Hinton, Simon Osindero, and Yee-Whye Teh · 2006
Earlier work this paper cites.
Building blocks for variational Bayesian learning of latent variable models
Tapani Raiko, Harri Valpola, Markus Harva, and Juha Karhunen · 2007
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
Ruslan Salakhutdinov and Iain Murray · 2008
Earlier work this paper cites.
Statistically optimal analysis of samples from multiple equilibrium states
Michael R. Shirts and John D. Chodera · 2008
Earlier work this paper cites.
Training restricted Boltzmann machines using approximations to the likelihood gradient
Tijmen Tieleman · 2008
Earlier work this paper cites.
Evaluating probabilities under high-dimensional latent variable models
Iain Murray and Ruslan R. Salakhutdinov · 2009
Earlier work this paper cites.
Deep Boltzmann machines
Ruslan Salakhutdinov and Geoffrey E. Hinton · 2009
Earlier work this paper cites.
Restricted Boltzmann machines are hard to approximately evaluate or simulate
Philip M. Long and Rocco Servedio · 2010
Cited alongside, same era.
Inductive principles for restricted Boltzmann machine learning
Benjamin M Marlin, Kevin Swersky, Bo Chen, and Nando de Freitas · 2010
Cited alongside, same era.
Unsupervised models of images by spike-and-slab rbms
Aaron C. Courville, James S. Bergstra, and Yoshua Bengio · 2011
Cited alongside, same era.
The neural autoregressive distribution estimator
Hugo Larochelle and Iain Murray · 2011
Cited alongside, same era.
Variational Baysian inference with stochastic search
John Paisley, David M. Blei, and Michael I. Jordan · 2012
Cited alongside, same era.
Adaptive dropout for training deep neural networks
Jimmy Ba and Brendan Frey · 2013
Cited alongside, same era.
Learning deep generative models with doubly stochastic MCMC
Chao Du, Jun Zhu, and Bo Zhang · 2015
Later among the works it cites.
DRAW: A recurrent neural network for image generation
Karol Gregor, Ivo Danihelka, Alex Graves, and Daan Wierstra · 2015
Later among the works it cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Later among the works it cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Later among the works it cites.
Alireza Makhzani, Jonathon Shlens, Navdeep Jaitly, and Ian Goodfellow · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio, Nicholas Léonard, and Aaron Courville · 2013
Cited alongside, same era.
Enhanced gradient for training restricted Boltzmann machines
KyungHyun Cho, Tapani Raiko, and Alexander Ilin · 2013
Cited alongside, same era.
One-shot learning by inverting a compositional causal process
Brenden M. Lake, Ruslan R. Salakhutdinov, and Josh Tenenbaum · 2013
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Deep autoregressive networks
Karol Gregor, Ivo Danihelka, Andriy Mnih, Charles Blundell, and Daan Wierstra · 2014
Cited alongside, same era.
Semi-supervised learning with deep generative models
Diederik P Kingma, Shakir Mohamed, Danilo Jimenez Rezende, and Max Welling · 2014
Cited alongside, same era.
Techniques for learning binary stochastic feedforward neural networks
Tapani Raiko, Mathias Berglund, Guillaume Alain, and Laurent Dinh · 2015
Later among the works it cites.
Semi-supervised learning with ladder networks
Antti Rasmus, Mathias Berglund, Mikko Honkala, Harri Valpola, and Tapani Raiko · 2015
Later among the works it cites.
Variational inference with normalizing flows
Danilo Rezende and Shakir Mohamed · 2015
Later among the works it cites.
Markov chain Monte Carlo and variational inference: Bridging the gap
Tim Salimans, Diederik P. Kingma, Max Welling, et al · 2015
Later among the works it cites.
Bidirectional Helmholtz machines
Jörg Bornschein, Samira Shabanian, Asja Fischer, and Yoshua Bengio · 2016
Closest in time.
Generating sentences from a continuous space
Samuel R. Bowman, Luke Vilnis, Oriol Vinyals, Andrew M. Dai, Rafal Jozefowicz, and Samy Bengio · 2016
Closest in time.
Importance weighted autoencoders
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov · 2016
Closest in time.
Stochastic backpropagation through mixture density distributions
Alex Graves · 2016
Closest in time.
Composing graphical models with neural networks for structured representations and fast inference
Matthew Johnson, David K Duvenaud, Alexander B Wiltschko, Sandeep R Datta, and Ryan P Adams · 2016
Closest in time.
Variational inference with Rényi divergence
Yingzhen Li and Richard E. Turner · 2016
Closest in time.
Variational inference for Monte Carlo objectives
Andriy Mnih and Danilo J. Rezende · 2016
Closest in time.
A structured variational auto-encoder for learning deep hierarchies of sparse features
Tim Salimans · 2016
Closest in time.
Ladder variational autoencoders
Casper Kaae Sønderby, Tapani Raiko, Lars Maaløe, Søren Kaae Sønderby, and Ole Winther · 2016
Closest in time.
The variational Gaussian process
Dustin Tran, Rajesh Ranganath, and David M. Blei · 2016
Closest in time.