Equation of state calculations by fast computing machines
N. Metropolis, A. W. Rosenbluth, M. N. Rosenbluth, A. H. Teller, and E. Teller · 1953
Earlier work this paper cites.
Monte carlo sampling methods using markov chains and their applications
W. K. Hastings · 1970
Earlier work this paper cites.
Exponential convergence of langevin distributions and their discrete approximations
G. O. Roberts and R. L. Tweedie · 1996
Earlier work this paper cites.
Optimal scaling of discrete approximations to langevin diffusions
G. O. Roberts and J. S. Rosenthal · 1998
Earlier work this paper cites.
Products of experts
G. E. Hinton · 1999
Earlier work this paper cites.
A kernel method for the two-sample-problem
A. Gretton, K. M. Borgwardt, M. Rasch, B. Schölkopf, and A. J. Smola · 2006
Earlier work this paper cites.
A tutorial on energy-based learning
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang · 2006
Earlier work this paper cites.
Visualizing data using t-sne
L. Van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
P. Vincent, H. Larochelle, Y. Bengio, and P.-A. Manzagol · 2008
Earlier work this paper cites.
Patchmatch: a randomized correspondence algorithm for structural image editing
C. Barnes, E. Shechtman, A. Finkelstein, and D. Goldman · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Visualizing higher-layer features of a deep network
D. Erhan, Y. Bengio, A. Courville, and P. Vincent · 2009
Earlier work this paper cites.
Probabilistic graphical models: principles and techniques
D. Koller and N. Friedman · 2009
Earlier work this paper cites.
The neural autoregressive distribution estimator
H. Larochelle and I. Murray · 2011
Earlier work this paper cites.
Contractive auto-encoders: Explicit invariance during feature extraction
S. Rifai, P. Vincent, X. Muller, X. Glorot, and Y. Bengio · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient langevin dynamics
M. Welling and Y. W. Teh · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Earlier work this paper cites.
Better mixing via deep representations
Y. Bengio, G. Mesnil, Y. Dauphin, and S. Rifai · 2013
Earlier work this paper cites.
Devise: A deep visual-semantic embedding model
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, M. A. Ranzato, and T. Mikolov · 2013
Earlier work this paper cites.
Caffe: An open source convolutional architecture for fast feature embedding
Y. Jia · 2013
Earlier work this paper cites.
Texture modeling with convolutional spike-and-slab rbms and deep extensions
H. Luo, P. L. Carrier, A. C. Courville, and Y. Bengio · 2013
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Original
K. Simonyan, A. Vedaldi, and A. Zisserman · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Original
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, and R. Fergus · 2013
Earlier work this paper cites.