Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
M. I. Jordan, Z. Ghahramani, T. S. Jaakkola, and L. K. Saul · 1999
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
R. Salakhutdinov and I. Murray · 2008
Earlier work this paper cites.
The neural autoregressive distribution estimator
H. Larochelle and I. Murray · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng · 2011
Earlier work this paper cites.
Variational bayesian inference with stochastic search
J. Paisley, D. M. Blei, and M. I. Jordan · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Y. Bengio, A. Courville, and P. Vincent · 2013
Earlier work this paper cites.
Stochastic variational inference
M. D. Hoffman, D. M. Blei, C. Wang, and J. Paisley · 2013
Earlier work this paper cites.
Auto-Encoding Variational Bayes
Original
M. Kingma, Diederik P; Welling · 2013
Earlier work this paper cites.
One-shot learning by inverting a compositional causal process
B. M. Lake, R. R. Salakhutdinov, and J. Tenenbaum · 2013
Earlier work this paper cites.
Nice: Non-linear independent components estimation
Original
L. Dinh, D. Krueger, and Y. Bengio · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Original
D. Kingma and J. Ba · 2014
Earlier work this paper cites.
Semi-Supervised Learning with Deep Generative Models
D. P. Kingma, D. J. Rezende, S. Mohamed, and M. Welling · 2014
Earlier work this paper cites.
Neural variational inference and learning in belief networks
A. Mnih and K. Gregor · 2014
Earlier work this paper cites.
Stochastic Backpropagation and Approximate Inference in Deep Generative Models
Original
D. J. Rezende, S. Mohamed, and D. Wierstra · 2014
Earlier work this paper cites.
Factoring variations in natural images with deep gaussian mixture models
A. van den Oord and B. Schrauwen · 2014
Earlier work this paper cites.
Generating sentences from a continuous space
Original
S. Bowman, L. Vilnis, O. Vinyals, A. Dai, R. Jozefowicz, and S. Bengio · 2015
Earlier work this paper cites.
Accurate and conservative estimates of mrf log-likelihood using reverse annealing
Y. Burda, R. Grosse, and R. Salakhutdinov · 2015
Earlier work this paper cites.