Fetching the paper…
Reading the bibliography…
We introduce a novel training principle for probabilistic models that is an alternative to maximum likelihood.
Perturbation theory and finite markov chains
Paul J Schweitzer · 1968
Earlier work this paper cites.
Simple statistical gradient-following algorithms connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Learning continuous attractors in recurrent networks
Sebastian H. Seung · 1998
Earlier work this paper cites.
Products of experts
Geoffrey E. Hinton · 1999
Earlier work this paper cites.
Comparison of perturbation bounds for the stationary distribution of a markov chain
Grace E. Cho, Carl D. Meyer, Carl, and D. Meyer · 2000
Earlier work this paper cites.
Dependency networks for inference, collaborative filtering, and data visualization
David Heckerman, David Maxwell Chickering, Christopher Meek, Robert Rounthwaite, and Carl Kadie · 2000
Earlier work this paper cites.
Learning iterative image reconstruction in the neural abstraction pyramid
Sven Behnke · 2001
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Geoffrey E. Hinton, Simon Osindero, and Yee Whye Teh · 2006
Earlier work this paper cites.
Consistency of pseudolikelihood estimation of fully visible boltzmann machines
Aapo Hyvärinen · 2006
Earlier work this paper cites.
Greedy layer-wise training of deep networks
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle · 2007
Earlier work this paper cites.
Efficient sparse coding algorithms
Honglak Lee, Alexis Battle, Rajat Raina, and Andrew Ng · 2007
Earlier work this paper cites.
Efficient learning of sparse representations with an energy-based model
M. Ranzato, C. Poultney, S. Chopra, and Y. LeCun · 2007
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
R. Collobert and J. Weston · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Pascal Vincent, Hugo Larochelle, Yoshua Bengio, and Pierre-Antoine Manzagol · 2008
Earlier work this paper cites.
Learning deep architectures for AI
Yoshua Bengio · 2009
Earlier work this paper cites.
Deep Boltzmann machines
Ruslan Salakhutdinov and Geoffrey E. Hinton · 2009
Earlier work this paper cites.
Theano: a CPU and GPU math expression compiler
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio · 2010
Earlier work this paper cites.
Phone recognition with the mean-covariance restricted Boltzmann machine
George E. Dahl, Marc’Aurelio Ranzato, Abdel-rahman Mohamed, and Geoffrey E. Hinton · 2010
Cited alongside, same era.
Binary coding of speech spectrograms using a deep auto-encoder
L. Deng, M. Seltzer, D. Yu, A. Acero, A. Mohamed, and G. Hinton · 2010
Cited alongside, same era.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
M. Gutmann and A. Hyvarinen · 2010
Cited alongside, same era.
Quickly generating representative samples from an RBM-derived process
Olivier Breuleux, Yoshua Bengio, and Pascal Vincent · 2011
Cited alongside, same era.
The Neural Autoregressive Distribution Estimator
H. Larochelle and I. Murray · 2011
Cited alongside, same era.
Sum-product networks: A new deep architecture
Hoifung Poon and Pedro Domingos · 2011
Fast gradient-based inference with continuous latent variable models in auxiliary form
Diederik P. Kingma · 2013
Later among the works it cites.
Texture modeling with convolutional spike-and-slab RBMs and deep extensions
Heng Luo, Pierre Luc Carrier, Aaron Courville, and Yoshua Bengio · 2013
Later among the works it cites.
Generalized denoising auto-encoders as generative models
Bengio, Yoshua, Yao, Li, Alain, Guillaume, and Vincent, Pascal · 2013
Later among the works it cites.
Deep generative stochastic networks trainable by backprop
Yoshua Bengio, Eric Thibodeau-Laufer, Guillaume Alain, and Jason Yosinski · 2014
Later among the works it cites.
Jörg Bornschein and Yoshua Bengio · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Réseaux de neurones à relaxation entraînés par critère d’autoencodeur débruitant
François Savard · 2011
Cited alongside, same era.
Conversational speech transcription using context-dependent deep neural networks
Frank Seide, Gang Li, and Dong Yu · 2011
Cited alongside, same era.
Improving neural networks by preventing co-adaptation of feature detectors
Geoffrey E. Hinton, Nitish Srivastava, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2012
Cited alongside, same era.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Cited alongside, same era.
Deep Boltzmann machines and the centering trick
Gregoire Montavon and Klaus-Robert Muller · 2012
Cited alongside, same era.
A generative process for sampling contractive auto-encoders
Salah Rifai, Yoshua Bengio, Yann Dauphin, and Pascal Vincent · 2012
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Later among the works it cites.
Deep autoregressive networks
Karol Gregor, Ivo Danihelka, Andriy Mnih, Charles Blundell, and Daan Wierstra · 2014
Later among the works it cites.
Auto-encoding variational bayes
Durk P. Kingma and Max Welling · 2014
Later among the works it cites.
Neural variational inference and learning in belief networks
Andriy Mnih and Karol Gregor · 2014
Later among the works it cites.
Deep directed generative autoencoders
Sherjil Ozair and Yoshua Bengio · 2014
Later among the works it cites.
Multimodal transitions for generative stochastic networks
Sherjil Ozair, Li Yao, and Yoshua Bengio · 2014
Later among the works it cites.
Stochastic backpropagation and approximate inference in deep generative models
Danilo J. Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Later among the works it cites.
On the equivalence between deep nade and generative stochastic networks
Li Yao, Sherjil Ozair, Kyunghyun Cho, and Yoshua Bengio · 2014
Later among the works it cites.
Deep supervised and convolutional generative stochastic network for protein secondary structure prediction
Jian Zhou and Olga G. Troyanskaya · 2014
Later among the works it cites.
General stochastic networks for classification
Matthias Zöhrer and Franz Pernkopf · 2014
Later among the works it cites.
Nice: Non-linear independent components estimation
Laurent Dinh, David Krueger, and Yoshua Bengio · 2015
Closest in time.