Fetching the paper…
Reading the bibliography…
Score matching is a popular method for estimating unnormalized statistical models.
The sizes of compact subsets of Hilbert space and continuity of Gaussian processes
Dudley, R. M · 1967
Earlier work this paper cites.
Estimation of the mean of a multivariate normal distribution
Stein, C. M · 1981
Earlier work this paper cites.
A stochastic estimator of the trace of the influence matrix for Laplacian smoothing splines
Hutchinson, M. F · 1990
Earlier work this paper cites.
Asymptotic Statistics
van der Vaart, A. W · 1998
Earlier work this paper cites.
Annealed importance sampling
Neal, R. M · 2001
Earlier work this paper cites.
Estimation of non-normalized statistical models by score matching
Hyvärinen, A · 2005
Earlier work this paper cites.
Kernel methods and the exponential family
Canu, S. and Smola, A · 2006
Earlier work this paper cites.
A tutorial on energy-based learning
LeCun, Y., Chopra, S., Hadsell, R., Ranzato, M. A., and Huang, F. J · 2006
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Gutmann, M. and Hyvärinen, A · 2010
Earlier work this paper cites.
Regularized estimation of image statistics by score matching
Kingma, D. P. and LeCun, Y · 2010
Earlier work this paper cites.
Bregman divergence as general framework to estimate unnormalized statistical models
Gutmann, M. U. and Hirayama, J.-i · 2011
Earlier work this paper cites.
A connection between score matching and denoising autoencoders
Vincent, P · 2011
Earlier work this paper cites.
Estimating the Hessian by back-propagating curvature
Martens, J., Sutskever, I., and Swersky, K · 2012
Earlier work this paper cites.
Wasserstein barycenter and its application to texture mixing
Rabin, J., Peyré, G., Delon, J., and Bernot, M · 2012
Cited alongside, same era.
Monte Carlo theory, methods and examples
Owen, A. B · 2013
Cited alongside, same era.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Cited alongside, same era.
Auto-encoding variational Bayes
Kingma, D. P. and Welling, M · 2014
Cited alongside, same era.
Clustering via mode seeking by direct estimation of the gradient of a log-density
Sasaki, H., Hyvärinen, A., and Sugiyama, M · 2014
Cited alongside, same era.
NICE: Non-linear independent components estimation
Dinh, L., Krueger, D., and Bengio, Y · 2015
Cited alongside, same era.
Measuring sample quality with kernels
Gorham, J. and Mackey, L · 2017
Later among the works it cites.
GANs trained by a two time-scale update rule converge to a local Nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Later among the works it cites.
Variational inference using implicit distributions
Huszár, F · 2017
Later among the works it cites.
Density estimation in infinite dimensional exponential families
Sriperumbudur, B., Fukumizu, K., Gretton, A., Hyvärinen, A., and Kumar, R · 2017
Later among the works it cites.
Hierarchical implicit models and likelihood-free variational inference
Tran, D., Ranganath, R., and Blei, D · 2017
Later among the works it cites.
Gradient estimators for implicit models
Li, Y. and Turner, R. E · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep learning face attributes in the wild
Liu, Z., Luo, P., Wang, X., and Tang, X · 2015
Cited alongside, same era.
Gradient-free Hamiltonian monte carlo with efficient kernel exponential families
Strathmann, H., Sejdinovic, D., Livingstone, S., Szabo, Z., and Gretton, A · 2015
Cited alongside, same era.
Tensorflow: A system for large-scale machine learning
Abadi, M., Barham, P., Chen, J., Chen, Z., Davis, A., Dean, J., Devin, M., Ghemawat, S., Irving, G., Isard, M., et al · 2016
Cited alongside, same era.
A kernelized Stein discrepancy for goodness-of-fit tests
Liu, Q., Lee, J., and Jordan, M · 2016
Cited alongside, same era.
Automatic differentiation in pytorch
Adam, P., Soumith, C., Gregory, C., Edward, Y., Zachary, D., Zeming, L., Alban, D., Luca, A., and Adam, L · 2017
Cited alongside, same era.
UCI machine learning repository, 2017
Dua, D. and Graff, C · 2017
Cited alongside, same era.
Later among the works it cites.
Deep energy estimator networks
Saremi, S., Mehrjou, A., Schölkopf, B., and Hyvärinen, A · 2018
Later among the works it cites.
A spectral approach to gradient estimation for implicit distributions
Shi, J., Sun, S., and Zhu, J · 2018
Later among the works it cites.
Efficient and principled score estimation with Nyström kernel exponential families
Sutherland, D., Strathmann, H., Arbel, M., and Gretton, A · 2018
Later among the works it cites.
Wasserstein auto-encoders
Tolstikhin, I., Bousquet, O., Gelly, S., and Schoelkopf, B · 2018
Later among the works it cites.
Functional variational Bayesian neural networks
Sun, S., Zhang, G., Shi, J., and Grosse, R · 2019
Closest in time.
Learning deep kernels for exponential family densities
Wenliang, L., Sutherland, D., Strathmann, H., and Gretton, A · 2019
Closest in time.