Fetching the paper…
Reading the bibliography…
Energy-based models (EBMs) are generative models that are usually trained via maximum likelihood estimation.
Zur theorie der gesellschaftsspiele
J. v. Neumann · 1928
Earlier work this paper cites.
Sur un theoreme fondamentale de la theorie des jeux
H. Kneser · 1952
Earlier work this paper cites.
Information theory and statistical mechanics
E. T. Jaynes · 1957
Earlier work this paper cites.
Linear operators. Part I, General theory
N. Dunford and J. T. Schwartz · 1958
Earlier work this paper cites.
On general minimax theorems
M. Sion · 1958
Earlier work this paper cites.
A class of markov processes associated with nonlinear parabolic equations
H. McKean · 1967
Earlier work this paper cites.
Statistical mechanics: Rigorous results
D. Ruelle · 1969
Earlier work this paper cites.
Random coding strategies for minimum entropy
E. C. Posner · 1975
Earlier work this paper cites.
Topics in propagation of chaos
A.-S. Sznitman · 1991
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
A. Barron · 1993
Earlier work this paper cites.
Convex analysis and variational problems
I. Ekeland and R. Temam · 1999
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
G. E. Hinton · 2002
Earlier work this paper cites.
Reproducing Kernel Hilbert Space in Probability and Statistics
A. Berlinet and C. Thomas-Agnan · 2004
Earlier work this paper cites.
Techniques of Variational Analysis
J. Borwein and Q. Zhu · 2005
Earlier work this paper cites.
Estimation of non-normalized statistical models by score matching
A. Hyvärinen · 2005
Earlier work this paper cites.
Unifying divergence minimization and statistical inference via convex duality
Y. Altun and A. Smola · 2006
Earlier work this paper cites.
A tutorial on energy-based learning
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang · 2006
Earlier work this paper cites.
Maximum entropy density estimation with generalized regularization and an application to species distribution modeling
M. Dudík, S. J. Phillips, and R. E. Schapire · 2007
Earlier work this paper cites.
A kernel method for the two-sample-problem
A. Gretton, K. M. Borgwardt, M. Rasch, B. Schölkopf, and A. J. Smola · 2007
Earlier work this paper cites.
Efficient learning of sparse representations with an energy-based model
M. Ranzato, C. Poultney, S. Chopra, et al · 2007
Earlier work this paper cites.
Continuous neural networks
N. L. Roux and Y. Bengio · 2007
Earlier work this paper cites.
Random features for large-scale kernel machines
A. Rahimi and B. Recht · 2008
Cited alongside, same era.
Training restricted boltzmann machines using approximations to the likelihood gradient
T. Tieleman · 2008
Cited alongside, same era.
Graphical models, exponential families, and variational inference
M. Wainwright and M. Jordan · 2008
Cited alongside, same era.
Kernel methods for deep learning
Y. Cho and L. K. Saul · 2009
Cited alongside, same era.
Using fast weights to improve persistent contrastive divergence
T. Tieleman and G. Hinton · 2009
Cited alongside, same era.
Elementary Principles in Statistical Mechanics: Developed with Especial Reference to the Rational Foundation of Thermodynamics
J. W. Gibbs · 2010
Cited alongside, same era.
G. M. Rotskoff and E. Vanden-Eijnden · 2018
Later among the works it cites.
Efficient and principled score estimation with nyström kernel exponential families
D. J. Sutherland, H. Strathmann, M. Arbel, and A. Gretton · 2018
Later among the works it cites.
Maximum mean discrepancy gradient flow
M. Arbel, A. Korba, A. Salim, and A. Gretton · 2019
Later among the works it cites.
Implicit generation and generalization in energy-based models
Y. Du and I. Mordatch · 2019
Later among the works it cites.
A priori estimates of the population risk for two-layer neural networks
W. E, C. Ma, and L. Wu · 2019
Later among the works it cites.
Global convergence of neuron birth-death dynamics
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A connection between score matching and denoising autoencoders
P. Vincent · 2011
Cited alongside, same era.
A kernel two-sample test
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. Smola · 2012
Cited alongside, same era.
Foundations of Machine Learning
M. Mohri, A. Rostamizadeh, and A. Talwalkar · 2012
Cited alongside, same era.
Training generative neural networks via maximum mean discrepancy optimization
G. K. Dziugaite, D. M. Roy, and Z. Ghahramani · 2015
Cited alongside, same era.
Generative moment matching networks
Y. Li, K. Swersky, and R. Zemel · 2015
Cited alongside, same era.
A theory of generative convnet
J. Xie, Y. Lu, S.-C. Zhu, and Y. Wu · 2016
Cited alongside, same era.
G. M. Rotskoff, S. Jelassi, J. Bruna, and E. Vanden-Eijnden · 2019
Later among the works it cites.
Mean field analysis of neural networks: A central limit theorem
J. Sirignano and K. Spiliopoulos · 2019
Later among the works it cites.
Generative modeling by estimating gradients of the data distribution
Y. Song and S. Ermon · 2019
Later among the works it cites.
A dynamical central limit theorem for shallow neural networks, 2020
Z. Chen, G. M. Rotskoff, J. Bruna, and E. Vanden-Eijnden · 2020
Later among the works it cites.
A mean-field analysis of two-player zero-sum games
C. Domingo-Enrich, S. Jelassi, A. Mensch, G. Rotskoff, and J. Bruna · 2020
Later among the works it cites.
On the banach spaces associated with multi-layer relu networks: Function representation, approximation theory and gradient descent dynamics, 2020
W. E and S. Wojtowytsch · 2020
Later among the works it cites.
Denoising diffusion probabilistic models
J. Ho, A. Jain, and P. Abbeel · 2020
Later among the works it cites.
Adversarial score matching and improved sampling for image generation
A. Jolicoeur-Martineau, R. Piché-Taillefer, R. T. d. Combes, and I. Mitliagkas · 2020
Later among the works it cites.
Solving linear inverse problems using the prior implicit in a denoiser
Z. Kadkhodaie and E. P. Simoncelli · 2020
Later among the works it cites.
On gradient descent ascent for nonconvex-concave minimax problems
T. Lin, C. Jin, and M. I. Jordan · 2020
Later among the works it cites.
Improved techniques for training score-based generative models
Y. Song and S. Ermon · 2020
Later among the works it cites.
Diffusion models beat gans on image synthesis
P. Dhariwal and A. Nichol · 2021
Closest in time.
On energy-based models with overparametrized shallow neural networks
C. Domingo-Enrich, A. Bietti, E. Vanden-Eijnden, and J. Bruna · 2021
Closest in time.
How to train your energy-based models, 2021
Y. Song and D. P. Kingma · 2021
Closest in time.
Score-based generative modeling through stochastic differential equations
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole · 2021
Closest in time.