Fetching the paper…
Reading the bibliography…
Neural Autoregressive Distribution Estimators (NADEs) have recently been shown as successful alternatives for modeling high dimensional multimodal distributions.
Practical markov chain monte carlo
C. J. Geyer · 1992
Earlier work this paper cites.
Mixture density networks
C. M. Bishop · 1994
Earlier work this paper cites.
Sampling from multimodal distributions using tempered transitions
R. M. Neal · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Modeling high-dimensional discrete data with multi-layer neural networks
Y. Bengio and S. Bengio · 2000
Earlier work this paper cites.
Annealed importance sampling
R. M. Neal · 2001
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
G. E. Hinton, S. Osindero, and Y. Teh · 2006
Earlier work this paper cites.
Reducing the dimensionality of data with neural networks
G. E. Hinton and R. Salakhutdinov · 2006
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
P. Vincent, H. Larochelle, Y. Bengio, and P.-A. Manzagol · 2008
Earlier work this paper cites.
Deep Boltzmann machines
R. Salakhutdinov and G. Hinton · 2009
Cited alongside, same era.
Deep Boltzmann machines
R. Salakhutdinov and G. Hinton · 2009
Cited alongside, same era.
Theano: a CPU and GPU math expression compiler
J. Bergstra, O. Breuleux, F. Bastien, P. Lamblin, R. Pascanu, G. Desjardins, J. Turian, D. Warde-Farley, and Y. Bengio · 2010
Cited alongside, same era.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P.-A. Manzagol · 2010
Cited alongside, same era.
The Neural Autoregressive Distribution Estimator
H. Larochelle and I. Murray · 2011
Cited alongside, same era.
Theano: new features and speed improvements
F. Bastien, P. Lamblin, R. Pascanu, J. Bergstra, I. J. Goodfellow, A. Bergeron, N. Bouchard, and Y. Bengio · 2012
Cited alongside, same era.
What regularized auto-encoders learn from the data generating distribution
G. Alain and Y. Bengio · 2013
Later among the works it cites.
Deep learning of representations: looking forward
Y. Bengio · 2013
Later among the works it cites.
Unsupervised feature learning and deep learning: A review and new perspectives
Y. Bengio, A. Courville, and P. Vincent · 2013
Later among the works it cites.
Generalized denoising auto-encoders as generative models
Y. Bengio, L. Yao, G. Alain, and P. Vincent · 2013
Later among the works it cites.
Multi-prediction deep Boltzmann machines
I. Goodfellow, M. Miraz, A. Courville, and Y. Bengio · 2013
Later among the works it cites.
A deep and tractable density estimator
B. Uria, I. Murray, and H. Larochelle · 2013
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improving neural networks by preventing co-adaptation of feature detectors
G. E. Hinton, N. Srivastava, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2012
Cited alongside, same era.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Cited alongside, same era.
Later among the works it cites.
Rnade: The real-valued neural autoregressive density-estimator
B. Uria, I. Murray, and H. Larochelle · 2013
Later among the works it cites.
Deep generative stochastic networks trainable by backprop
Y. Bengio, E. Thibodeau-Laufer, G. Alain, and J. Yosinski · 2014
Closest in time.