Fetching the paper…
Reading the bibliography…
We present an efficient algorithm for maximum likelihood estimation (MLE) of exponential family models, with a general parametrization of the energy function that includes neural networks.
Convex Analysis , volume 28 of Princeton Mathematics Series
R. T. Rockafellar · 1970
Earlier work this paper cites.
Statistical analysis of non-lattice data
J. Besag · 1975
Earlier work this paper cites.
Markov Random Fields and their applications
R. Kinderman and J. L. Snell · 1980
Earlier work this paper cites.
Fundamentals of Statistical Exponential Families , volume 9 of Lecture notes-monograph series
Lawrence D. Brown · 1986
Earlier work this paper cites.
Composite likelihood methods
B. G. Lindsay · 1988
Earlier work this paper cites.
Nonlinear Programming
D. P. Bertsekas · 1995
Earlier work this paper cites.
Grade: Gibbs reaction and diffusion equations
Song Chun Zhu and David Mumford · 1998
Earlier work this paper cites.
Conditional random fields: Probabilistic modeling for segmenting and labeling sequence data
J. D. Lafferty, A. McCallum, and F. Pereira · 2001
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
Geoffrey E. Hinton · 2002
Earlier work this paper cites.
Estimation of non-normalized statistical models using score matching
A. Hyvärinen · 2005
Earlier work this paper cites.
A tutorial on energy-based learning
Yann LeCun, Sumit Chopra, Raia Hadsell, M Ranzato, and F Huang · 2006
Earlier work this paper cites.
Connections between score matching, contrastive divergence, and pseudolikelihood for continuous-valued variables
Aapo Hyvärinen · 2007
Earlier work this paper cites.
Training restricted Boltzmann machines using approximations to the likelihood gradient
Tijmen Tieleman · 2008
Earlier work this paper cites.
Graphical models, exponential families, and variational inference
M. J. Wainwright and M. I. Jordan · 2008
Earlier work this paper cites.
Using fast weights to improve persistent contrastive divergence
Tijmen Tieleman and Geoffrey Hinton · 2009
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Michael Gutmann and Aapo Hyvärinen · 2010
Earlier work this paper cites.
Regularized estimation of image statistics by score matching
Diederik P Kingma and Yann LeCun · 2010
Earlier work this paper cites.
Non-local contrastive objectives
David Vickrey, Cliff Chiung-Yu Lin, and Daphne Koller · 2010
Earlier work this paper cites.
MCMC using Hamiltonian dynamics
Radford M Neal et al · 2011
Cited alongside, same era.
Minimum probability flow learning
Jascha Sohl-Dickstein, Peter Battaglino, and Michael R DeWeese · 2011
Cited alongside, same era.
A kernel two-sample test
A. Gretton, K. Borgwardt, M. Rasch, B. Schoelkopf, and A. Smola · 2012
Cited alongside, same era.
A fast and simple algorithm for training neural probabilistic language models
Andriy Mnih and Yee Whye Teh · 2012
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Auto-encoding variational Bayes
Diederik P Kingma and Max Welling · 2014
Cited alongside, same era.
Improved training of wasserstein gans
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron C Courville · 2017
Later among the works it cites.
Learning deep energy models: Contrastive divergence vs. amortized MLE
Qiang Liu and Dilin Wang · 2017
Later among the works it cites.
A-nice-MC: Adversarial training for mcmc
Jiaming Song, Shengjia Zhao, and Stefano Ermon · 2017
Later among the works it cites.
Density estimation in infinite dimensional exponential families
Bharath Sriperumbudur, Kenji Fukumizu, Arthur Gretton, Aapo Hyvärinen, and Revant Kumar · 2017
Later among the works it cites.
Hamiltonian variational auto-encoder
Anthony L Caterini, Arnaud Doucet, and Dino Sejdinovic · 2018
Later among the works it cites.
Coupled variational bayes via optimization embedding
Bo Dai, Hanjun Dai, Niao He, Weiyang Liu, Zhen Liu, Jianshu Chen, Lin Xiao, and Le Song · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stochastic backpropagation and approximate inference in deep generative models
Danilo J Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Cited alongside, same era.
Large-scale log-determinant computation through stochastic chebyshev expansions
Insu Han, Dmitry Malioutov, and Jinwoo Shin · 2015
Cited alongside, same era.
A complete recipe for stochastic gradient MCMC
Yi-An Ma, Tianqi Chen, and Emily Fox · 2015
Cited alongside, same era.
Variational inference with normalizing flows
Danilo Jimenez Rezende and Shakir Mohamed · 2015
Cited alongside, same era.
Provable bayesian inference via particle mirror descent
Bo Dai, Niao He, Hanjun Dai, and Le Song · 2016
Cited alongside, same era.
Deep directed generative models with energy-based probability estimation
Taesup Kim and Yoshua Bengio · 2016
Cited alongside, same era.
Later among the works it cites.
Glow: Generative flow with invertible 1 × 1 1\times 1 convolutions
Diederik P Kingma and Prafulla Dhariwal · 2018
Later among the works it cites.
Generalizing hamiltonian monte carlo with neural networks
Daniel Levy, Matthew D Hoffman, and Jascha Sohl-Dickstein · 2018
Later among the works it cites.
Spectral normalization for generative adversarial networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida · 2018
Later among the works it cites.
Sparse and deep generalizations of the frame model
Ying Nian Wu, Jianwen Xie, Yang Lu, and Song-Chun Zhu · 2018
Later among the works it cites.
Monge-Ampère flow for generative modeling
Linfeng Zhang, Weinan E, and Lei Wang · 2018
Later among the works it cites.
Minimum Stein Discrepancy Estimators
Alessandro Barp, Francois-Xavier Briol, Andrew B. Duncan, Mark Girolami, and Lester Mackey · 2019
Closest in time.
Kernel exponential family estimation via doubly dual embedding
Bo Dai, Hanjun Dai, Arthur Gretton, Le Song, Dale Schuurmans, and Niao He · 2019
Closest in time.
Implicit generation and generalization in energy-based models
Yilun Du and Igor Mordatch · 2019
Closest in time.
Meta-learning for stochastic gradient MCMC
Wenbo Gong, Yingzhen Li, and José Miguel Hernández-Lobato · 2019
Closest in time.
FFJORD: Free-form continuous dynamics for scalable reversible generative models
Will Grathwohl, Ricky TQ Chen, Jesse Betterncourt, Ilya Sutskever, and David Duvenaud · 2019
Closest in time.
Learning deep kernels for exponential family densities
Li Wenliang, Dougal Sutherland, Heiko Strathmann, and Arthur Gretton · 2019
Closest in time.