Fetching the paper…
Reading the bibliography…
We unify $\textit{kernel density estimation}$ and $\textit{empirical Bayes}$ and address a set of problems in unsupervised learning with a geometric interpretation of those methods, rooted in the $\textit{concentration of measure}$ phenomenon.
A logical calculus of the ideas immanent in nervous activity
Warren S McCulloch and Walter Pitts · 1943
Earlier work this paper cites.
A stochastic approximation method
Herbert Robbins and Sutton Monro · 1951
Earlier work this paper cites.
An empirical Bayes approach to statistics
Herbert Robbins · 1956
Earlier work this paper cites.
An empirical Bayes estimator of the mean of a normal population
Koichi Miyasawa · 1961
Earlier work this paper cites.
On estimation of a probability density function and mode
Emanuel Parzen · 1962
Earlier work this paper cites.
More is different
Philip W Anderson · 1972
Earlier work this paper cites.
Estimation of the mean of a multivariate normal distribution
Charles M Stein · 1981
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computational abilities
John J Hopfield · 1982
Earlier work this paper cites.
Structure and interpretation of computer programs
Harold Abelson and Gerald Jay Sussman · 1985
Earlier work this paper cites.
Learning and relearning in Boltzmann machines
Geoffrey E Hinton and Terrence J Sejnowski · 1986
Earlier work this paper cites.
Stochastic processes in physics and chemistry , volume 1
Nicolaas Godfried van Kampen · 1992
Earlier work this paper cites.
Mean shift, mode seeking, and clustering
Yizong Cheng · 1995
Earlier work this paper cites.
The nature of statistical learning theory
Vladimir Vapnik · 1995
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Asymptotic statistics
Aad van der Vaart · 2000
Earlier work this paper cites.
The concentration of measure phenomenon
Michel Ledoux · 2001
Cited alongside, same era.
Mean shift: A robust approach toward feature space analysis
Dorin Comaniciu and Peter Meer · 2002
Cited alongside, same era.
On the number of modes of a gaussian mixture
Miguel Á Carreira-Perpiñán and Christopher KI Williams · 2003
Cited alongside, same era.
Information theory, inference and learning algorithms
David MacKay · 2003
Cited alongside, same era.
Think globally, fit locally: unsupervised learning of low dimensional manifolds
Lawrence K Saul and Sam T Roweis · 2003
Cited alongside, same era.
Estimation of non-normalized statistical models by score matching
Aapo Hyvärinen · 2005
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Later among the works it cites.
A review of mean-shift algorithms for clustering
Miguel Á Carreira-Perpiñán · 2015
Later among the works it cites.
GSNs: generative stochastic networks
Guillaume Alain, Yoshua Bengio, Li Yao, Jason Yosinski, Eric Thibodeau-Laufer, Saizheng Zhang, and Pascal Vincent · 2016
Later among the works it cites.
Dense associative memory for pattern recognition
Dmitry Krotov and John J Hopfield · 2016
Later among the works it cites.
Associative content-addressable networks with exponentially many robust stable states
Rishidev Chaudhuri and Ila Fiete · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gaussian mean-shift is an EM algorithm
Miguel Á Carreira-Perpiñán · 2007
Cited alongside, same era.
Graphical models, exponential families, and variational inference
Martin J Wainwright and Michael I Jordan · 2008
Cited alongside, same era.
Learning deep architectures for AI
Yoshua Bengio · 2009
Cited alongside, same era.
Probabilistic graphical models: principles and techniques
Daphne Koller and Nir Friedman · 2009
Cited alongside, same era.
Least squares estimation without priors or supervision
Martin Raphan and Eero P Simoncelli · 2011
Cited alongside, same era.
A connection between score matching and denoising autoencoders
Pascal Vincent · 2011
Cited alongside, same era.
Automatic differentiation in PyTorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Later among the works it cites.
Swish: a self-gated activation function
Prajit Ramachandran, Barret Zoph, and Quoc V Le · 2017
Later among the works it cites.
On the optimization of deep networks: Implicit acceleration by overparameterization
Sanjeev Arora, Nadav Cohen, and Elad Hazan · 2018
Later among the works it cites.
Automatic differentiation in machine learning: a survey
Atilim Gunes Baydin, Barak A Pearlmutter, Alexey Andreyevich Radul, and Jeffrey Mark Siskind · 2018
Later among the works it cites.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
Stefan Elfwing, Eiji Uchibe, and Kenji Doya · 2018
Later among the works it cites.
Deep energy estimator networks
Saeed Saremi, Arash Mehrjou, Bernhard Schölkopf, and Aapo Hyvärinen · 2018
Later among the works it cites.
High-dimensional probability: An introduction with applications in data science
Roman Vershynin · 2018
Later among the works it cites.
On approximating ∇ f \nabla f with neural networks
Saeed Saremi · 2019
Closest in time.
High-dimensional statistics: A non-asymptotic viewpoint , volume 48
Martin J Wainwright · 2019
Closest in time.