Massively parallel architectures for AI: NETL, thistle, and Boltzmann machines
Fahlman, S. E., Hinton, G. E., and Sejnowski, T. J. (1983) · 1983
Earlier work this paper cites.
Boltzmann machines: Constraint satisfaction networks that learn
Hinton, G. E., Sejnowski, T. J., and Ackley, D. H. (1984) · 1984
Earlier work this paper cites.
A learning algorithm for Boltzmann machines
Ackley, D. H., Hinton, G. E., and Sejnowski, T. J. (1985) · 1985
Earlier work this paper cites.
Learning and relearning in Boltzmann machines
Hinton, G. E. and Sejnowski, T. J. (1986) · 1986
Earlier work this paper cites.
Simple statistical gradient-following algorithms connectionist reinforcement learning
Williams, R. J. (1992) · 1992
Earlier work this paper cites.
Higher order statistical decorrelation without information loss
Deco, G. and Brauer, W. (1995) · 1995
Earlier work this paper cites.
Does the wake-sleep algorithm learn good density estimators?
Frey, B. J., Hinton, G. E., and Dayan, P. (1996) · 1996
Earlier work this paper cites.
Graphical models for machine learning and digital communication
Frey, B. J. (1998) · 1998
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Hinton, G. E., Osindero, S., and Teh, Y. (2006) · 2006
Earlier work this paper cites.
Learning multiple layers of representation
Hinton, G. E. (2007) · 2007
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L. (2009) · 2009
Earlier work this paper cites.
Deep Boltzmann machines
Salakhutdinov, R. and Hinton, G. (2009) · 2009
Earlier work this paper cites.
What does classifying more than 10,000 image categories tell us?
Deng, J., Berg, A. C., Li, K., and Fei-Fei, L. (2010) · 2010
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Gutmann, M. and Hyvarinen, A. (2010) · 2010
Earlier work this paper cites.
Fast gradient-based inference with continuous latent variable models in auxiliary form
Original
Kingma, D. P. (2013) · 2013
Earlier work this paper cites.
Characterization and computation of local nash equilibria in continuous games
Ratliff, L. J., Burden, S. A., and Sastry, S. S. (2013) · 2013
Earlier work this paper cites.
Deep generative stochastic networks trainable by backprop
Bengio, Y., Thibodeau-Laufer, E., Alain, G., and Yosinski, J. (2014) · 2014
Earlier work this paper cites.
NICE: Non-linear independent components estimation
Original
Dinh, L., Krueger, D., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
On distinguishability criteria for estimating generative models
Goodfellow, I. J. (2014) · 2014
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, I. J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2014b) · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
Kingma, D. and Ba, J. (2014) · 2014
Earlier work this paper cites.