Fetching the paper…
Reading the bibliography…
We present doubly stochastic gradient MCMC, a simple and generic method for (approximate) Bayesian inference of deep generative models (DGMs) in a collapsed continuous parameter space.
A stochastic approximation method
Robbins, H. and Monro, S · 1951
Earlier work this paper cites.
A practical bayesian framework for backpropagation networks
MacKay, D · 1992
Earlier work this paper cites.
Connectionist learning of belief networks
Neal, R. M · 1992
Earlier work this paper cites.
Bayesian learning for neural networks
Neal, R. M · 1995
Earlier work this paper cites.
Mean field theory for sigmoid belief networks
Saul, L., Jaakkola, T., and Jordan, M · 1996
Earlier work this paper cites.
Online Algorithms and Stochastic Approximations
Bottou, L · 1998
Earlier work this paper cites.
Graphical models for machine learning and digital communication
Frey, B. J · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Lecun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
Jordan, M., Ghahramani, Z., Jaakkola, T., and Saul, L · 1999
Earlier work this paper cites.
Modeling high-dimensional discrete data with multi-layer neural networks
Bengio, Y. and Bengio, S · 2000
Earlier work this paper cites.
Monte Carlo Statistical Methods
Robert, C. and Casella, G · 2005
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Hinton, G. E., Osindero, S., and Teh, Y · 2006
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
Salakhutdinov, R. and Murray, I · 2008
Earlier work this paper cites.
Evaluating probabilities under high-dimensional latent variable models
Murray, I. and Salakhutdinov, R · 2009
Earlier work this paper cites.
Deep Boltzmann machines
Salakhutdinov, R. and Hinton, G. E · 2009
Earlier work this paper cites.
Learning the structure of deep sparse graphical models
Adams, R., Wallach, H., and Ghahramani, Z · 2010
Cited alongside, same era.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X. and Bengio, Y · 2010
Cited alongside, same era.
Inductive principles for restricted boltzmann machine learning
Marlin, B., Swersky, K., Chen, B., and Freitas, N · 2010
Cited alongside, same era.
The neural autoregressive distribution estimator
Larochelle, H. and Murray, I · 2011
Cited alongside, same era.
Bayesian learning via stochastic gradient Langevin dynamics
Welling, M. and Teh, Y. W · 2011
Cited alongside, same era.
Bayesian posterior sampling via stochastic gradient fisher scoring
Ahn, S., Korattikara, A., and Welling, M · 2012
Auto-encoding variational Bayes
Kingma, D. P. and Welling, M · 2014
Later among the works it cites.
Neural variational inference and learning in belief networks
Mnih, A. and Gregor, K · 2014
Later among the works it cites.
Iterative neural autoregressive distribution estimator nade-k
Raiko, T., Li, Y., Cho, K., and Bengio, Y · 2014
Later among the works it cites.
Black box variational inference
Ranganath, R., Gerrish, S., and Blei, D. M · 2014
Later among the works it cites.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
Later among the works it cites.
Doubly stochastic variational bayes for non-conjugate inference
Titsias, M. K. and Lázaro-Gredilla, M · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Modeling temporal dependencies in high-dimensional sequences: Application to polyphonic music generation and transcription
Boulanger-Lewandowski, N., Bengio, Y., and Vincent, P · 2012
Cited alongside, same era.
Enhanced gradient for training restricted boltzmann machines
Cho, K., Raiko, T., and Ilin, A · 2013
Cited alongside, same era.
One-shot learning by inverting a compositional causal process
Lake, B. M., Salakhutdinov, R., and Tenenbaum, J · 2013
Cited alongside, same era.
Monte Carlo theory, methods and examples
Owen, A. B · 2013
Cited alongside, same era.
Deep generative stochastic networks trainable by backprop
Bengio, Y., Laufer, E., Alain, G., and Yosinski, J · 2014
Cited alongside, same era.
Stochastic gradient hamiltonian monte carlo
Chen, T., Fox, E., and Guestrin, C · 2014
Cited alongside, same era.
A deep and tractable density estimator
Uria, B., Murray, I., and Larochelle, H · 2014
Later among the works it cites.
Reweighted wake-sleep
Bornschein, J. and Bengio, Y · 2015
Closest in time.
Importance weighted autoencoders
Burda, Y., Grosse, R. B., and Salakhutdinov, R · 2015
Closest in time.
Sparse autoregressive networks
Goessling, M. and Amit, Y · 2015
Closest in time.
Draw: A recurrent neural network for image generation
Gregor, K., Danihelka, I., Graves, A., Rezende, D., and Wierstra, D · 2015
Closest in time.
Neural adaptive sequential monte carlo
Gu, S., Ghahramani, Z., and Turner, R. E · 2015
Closest in time.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. L · 2015
Closest in time.
High-order stochastic gradient thermostats for bayesian learning of deep models
Li, C., Chen, C., Fan, K., and Carin, L · 2016
Closest in time.