Fetching the paper…
Reading the bibliography…
Neural networks, a central tool in machine learning, have demonstrated remarkable, high fidelity performance on image recognition and classification tasks.
A class of markov processes associated with nonlinear parabolic equations
H P McKean · 1966
Earlier work this paper cites.
Large deviations from the McKean-Vlasov limit for weakly interacting diffusions
Donald Dawson and Jürgen Gärtner · 1987
Earlier work this paper cites.
On the McKean-Vlasov Limit for Interacting Diffusions
Jürgen Gärtner · 1988
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
G Cybenko · 1989
Earlier work this paper cites.
Universal Approximation Using Radial-Basis-Function Networks
J Park and I W Sandberg · 1991
Earlier work this paper cites.
Topics in propagation of chaos
Alain-Sol Sznitman · 1991
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
A R Barron · 1993
Earlier work this paper cites.
The Variational Formulation of the Fokker–Planck Equation
Richard Jordan, David Kinderlehrer, and Felix Otto · 1998
Earlier work this paper cites.
Convergence of probability measures
Patrick Billingsley · 1999
Earlier work this paper cites.
Langevin equation for the density of a system of interacting Langevin processes
David S Dean · 1999
Earlier work this paper cites.
Large scale online learning
Léon Bottou and Yann LeCun · 2004
Earlier work this paper cites.
Gradient Flows in Metric Spaces and in the Space of Probability Measures
Luigi Ambrosio, Nicola Gigli, and Giuseppe Savaré · 2005
Earlier work this paper cites.
Generalized Neural-Network Representation of High-Dimensional Potential-Energy Surfaces
Jörg Behler and Michele Parrinello · 2007
Cited alongside, same era.
Optimal Transport: Old and New
Cédric Villani · 2009
Cited alongside, same era.
Random Matrices and Complexity of Spin Glasses
Antonio Auffinger, Gérard Ben Arous, and Jiří Černý · 2012
Cited alongside, same era.
Complexity of random smooth functions on the high-dimensional sphere
Antonio Auffinger and Gérard Ben Arous · 2013
Cited alongside, same era.
Scaling limits of interacting particle systems
Claude Kipnis and Claudio Landim · 2013
Cited alongside, same era.
The Loss Surfaces of Multilayer Networks
Anna Choromanska, Mikael Henaff, Michael Mathieu, Gérard Ben Arous, and Yann LeCun · 2014
Jens Berg and Kaj Nyström · 2017
Later among the works it cites.
W. E, J. Han, and A. Jentzen · 2017
Later among the works it cites.
On the diffusion approximation of nonconvex stochastic gradient descent
Wenqing Hu, Chris Junchi Li, Lei Li, and Jian-Guo Liu · 2017
Later among the works it cites.
Large deviation principle for empirical fields of log and riesz gases
Thomas Leblé and Sylvia Serfaty · 2017
Later among the works it cites.
Stochastic Neural Network Approach for Learning High-Dimensional Free Energy Surfaces
Elia Schneider, Luke Dai, Robert Q Topper, Christof Drechsel-Grau, and Mark E Tuckerman · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Explorations on high dimensional landscapes
Levent Sagun, V. Ugur Guney, Gérard Ben Arous, and Yann LeCun · 2014
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Cited alongside, same era.
Stochastic modified equations and adaptive stochastic gradient algorithms
Qianxiao Li, Cheng Tai, and Weinan E · 2015
Cited alongside, same era.
Coulomb gases and Ginzburg-Landau vortices
Sylvia Serfaty · 2015
Cited alongside, same era.
Breaking the curse of dimensionality with convex neural networks
Francis Bach · 2017
Cited alongside, same era.
C. Beck, W. E, and A. Jentzen · 2017
Cited alongside, same era.
Later among the works it cites.
Systems of Points with Coulomb Interactions
Sylvia Serfaty · 2017
Later among the works it cites.
Comparing Dynamics: Deep Neural Networks versus Glassy Systems
M. Baity-Jesi, L. Sagun, M. Geiger, S. Spigler, G. Ben Arous, C. Cammarota, Y. LeCun, M. Wyart, and G. Biroli · 2018
Closest in time.
On the global convergence of gradient descent for over-parameterized models using optimal transport
Lénaïc Chizat and Francis Bach · 2018
Closest in time.
Solving for high dimensional committor functions using artificial neural networks
Yuehaw Khoo, Jianfeng Lu, and Lexing Ying · 2018
Closest in time.
A mean field view of the landscape of two-layer neural networks
Song Mei, Andrea Montanari, and Phan-Minh Nguyen · 2018
Closest in time.
Mean Field Analysis of Neural Networks
Justin Sirignano and Konstantinos Spiliopoulos · 2018
Closest in time.
DeePCG: constructing coarse-grained models via deep neural networks
L. Zhang, J. Han, H. Wang, R. Car, and W. E · 2018
Closest in time.