Fetching the paper…
Reading the bibliography…
Categorical variables are a natural choice for representing discrete structure in the world.
Statistical theory of extreme values and some practical applications: a series of lectures
E. J. Gumbel · 1954
Earlier work this paper cites.
Likelihood ratio gradient estimation for stochastic systems
P. W Glynn · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
On the quantitative analysis of deep belief networks
R. Salakhutdinov and I. Murray · 2008
Earlier work this paper cites.
The neural autoregressive distribution estimator
H. Larochelle and I. Murray · 2011
Earlier work this paper cites.
Variational Bayesian Inference with Stochastic Search
J. Paisley, D. Blei, and M. Jordan · 2012
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Y. Bengio, N. Léonard, and A. Courville · 2013
Earlier work this paper cites.
K. Gregor, I. Danihelka, A. Mnih, C. Blundell, and D. Wierstra · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Alex Graves, Greg Wayne, and Ivo Danihelka · 2014
Cited alongside, same era.
Semi-supervised learning with deep generative models
D. P. Kingma, S. Mohamed, D. J. Rezende, and M. Welling · 2014
Cited alongside, same era.
A* sampling
C. J. Maddison, D. Tarlow, and T. Minka · 2014
Cited alongside, same era.
Neural variational inference and learning in belief networks
A. Mnih and K. Gregor · 2014
Cited alongside, same era.
Techniques for learning binary stochastic feedforward neural networks
T. Raiko, M. Berglund, G. Alain, and L. Dinh · 2014
Cited alongside, same era.
Gradient estimation using stochastic computation graphs
J. Schulman, N. Heess, T. Weber, and P. Abbeel · 2015
Hierarchical multiscale recurrent neural networks
J. Chung, S. Ahn, and Y. Bengio · 2016
Closest in time.
Hybrid computing using a neural network with dynamic external memory
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-Barwińska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, et al · 2016
Closest in time.
MuProp: Unbiased Backpropagation for Stochastic Neural Networks
S. Gu, S. Levine, I. Sutskever, and A Mnih · 2016
Closest in time.
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
C. J. Maddison, A. Mnih, and Y. Whye Teh · 2016
Closest in time.
Variational inference for monte carlo objectives
A. Mnih and D. J. Rezende · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Rethinking the inception architecture for computer vision
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. C. Courville, R. Salakhutdinov, R. S. Zemel, and Y. Bengio · 2015
Cited alongside, same era.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel · 2016
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra
Cited in the paper.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra
Cited in the paper.
Regularizing neural networks by penalizing confident output distributions
Gabriel Pereyra, Geoffrey Hinton, George Tucker, and Lukasz Kaiser · 2016
Closest in time.
Scaling Memory-Augmented Neural Networks with Sparse Reads and Writes
J. W Rae, J. J Hunt, T. Harley, I. Danihelka, A. Senior, G. Wayne, A. Graves, and T. P Lillicrap · 2016
Closest in time.
Discrete Variational Autoencoders
J. T. Rolfe · 2016
Closest in time.