Fetching the paper…
Reading the bibliography…
Maximum Likelihood Estimators (MLE) has many good properties.
Information and the accuracy attainable in the estimation of statistical parameters
C. R. Rao · 1945
Earlier work this paper cites.
Mathematical methods of statistics
H. Cramér · 1946
Earlier work this paper cites.
A bound for the error in the normal approximation to the distribution of a sum of dependent random variables
C. Stein · 1972
Earlier work this paper cites.
A new look at the statistical model identification
H. Akaike · 1974
Earlier work this paper cites.
Asymptotic Statistics
A. W. van der Vaart · 1998
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
G. E. Hinton · 2002
Earlier work this paper cites.
Inequalities for spreads of matrix sums and products
J. K. Merikoski and R. Kumar · 2004
Earlier work this paper cites.
Estimation of non-normalized statistical models by score matching
A. Hyvärinen · 2005
Earlier work this paper cites.
Monte Carlo Statistical Methods
C. P. Robert and G. Casella · 2005
Earlier work this paper cites.
Numerical Optimization
J. Nocedal and S.J. Wright · 2006
Earlier work this paper cites.
Eigenvalues of rank-one updated matrices with some applications
J. Ding and A. Zhou · 2007
Earlier work this paper cites.
Some extensions of score matching
A. Hyvärinen · 2007
Earlier work this paper cites.
Direct importance estimation with model selection and its application to covariate shift adaptation
M. Sugiyama, S. Nakajima, H. Kashima, P. von Bünau, and M. Kawanabe · 2008
Cited alongside, same era.
Interpretation and generalization of score matching
S. Lyu · 2009
Cited alongside, same era.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
M. Gutmann and A. Hyvärinen · 2010
Cited alongside, same era.
Projection Matrices, Generalized Inverse Matrices, and Singular Value Decomposition
H. Yanai, K. Takeuchi, and Y. Takane · 2011
Cited alongside, same era.
A kernel two-sample test
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. Smola · 2012
Cited alongside, same era.
Jensen divergence based on fisher’s information
P. Sánchez-Moreno, A. Zarzo, and J. S. Dehesa · 2012
Stein variational gradient descent: A general purpose bayesian inference algorithm
Q. Liu and D. Wang · 2016
Later among the works it cites.
A kernelized stein discrepancy for goodness-of-fit tests
Q Liu, J. D. Lee, and M. Jordan · 2016
Later among the works it cites.
Learning to draw samples: With application to amortized MLE for generative adversarial learning
D. Wang and Q. Liu · 2016
Later among the works it cites.
A linear-time kernel goodness-of-fit test
W. Jitkrittum, W. Xu, Z. Szabó, K. Fukumizu, and A. Gretton · 2017
Later among the works it cites.
Control functionals for monte carlo integration
C. J. Oates, M. Girolami, and N. Chopin · 2017
Later among the works it cites.
Density estimation in infinite dimensional exponential families
B. Sriperumbudur, K. Fukumizu, A. Gretton, A. Hyvärinen, and R. Kumar · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Density Ratio Estimation in Machine Learning
M. Sugiyama, T. Suzuki, and T. Kanamori · 2012
Cited alongside, same era.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Cited alongside, same era.
Measuring sample quality with stein’s method
J. Gorham and L. Mackey · 2015
Cited alongside, same era.
A kernel test of goodness of fit
K. Chwialkowski, H. Strathmann, and Arthur Gretton · 2016
Cited alongside, same era.
Estimation of exponential-polynomial distribution by holonomic gradient descent
J. Hayakawa and A. Takemura · 2016
Cited alongside, same era.
Estimation of high-dimensional graphical models using regularized score matching
L. Lin, M. Drton, and A. Shojaie · 2016
Cited alongside, same era.
Later among the works it cites.
Stein points
W. Y. Chen, L. Mackey, J. Gorham, F. X. Briol, and C. Oates · 2018
Closest in time.
Gradient estimators for implicit models
Y. Li and R. E. Turner · 2018
Closest in time.
A spectral approach to gradient estimation for implicit distributions
J. Shi, S. Sun, and J. Zhu · 2018
Closest in time.
Generalized score matching for non-negative data
S. Yu, M. Drton, and A. Shojaie · 2018
Closest in time.
Minimum stein discrepancy estimators
A. Barp, F-X. Briol, A. Duncan, M. Girolami, and L. Mackey · 2019
Closest in time.