Fetching the paper…
Reading the bibliography…
This paper proposes a new mean-field framework for over-parameterized deep neural networks (DNNs), which can be used to analyze neural network training.
Ordinary differential equations
Philip Hartman · 1964
Earlier work this paper cites.
Wandering random measures in the fleming-viot model
Donald A Dawson, Kenneth J Hochberg, et al · 1982
Earlier work this paper cites.
Abstract comparison principles and multivariable gronwall-bellman inequalities
Mihai Turinici · 1986
Earlier work this paper cites.
Topics in propagation of chaos
Alain-Sol Sznitman · 1991
Earlier work this paper cites.
Statistical mechanics of learning
Andreas Engel and Christian Van den Broeck · 2001
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2010
Earlier work this paper cites.
Matrix analysis
Rajendra Bhatia · 2013
Earlier work this paper cites.
Fixed point theory
Andrzej Granas and James Dugundji · 2013
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Optimal transport for applied mathematicians
Filippo Santambrogio · 2015
Earlier work this paper cites.
Identity matters in deep learning
Moritz Hardt and Tengyu Ma · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Learning and generalization in overparameterized neural networks, going beyond two layers
Zeyuan Allen-Zhu, Yuanzhi Li, and Yingyu Liang · 2018
Cited alongside, same era.
On the global convergence of gradient descent for over-parameterized models using optimal transport
Lenaic Chizat and Francis Bach · 2018
Cited alongside, same era.
Multilevel mutation-selection systems and set-valued duals
Donald A Dawson · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Cited alongside, same era.
Learning overparameterized neural networks via stochastic gradient descent on structured data
Yuanzhi Li and Yingyu Liang · 2018
Cited alongside, same era.
Xialiang Dou and Tengyuan Liang · 2019
Later among the works it cites.
Gradient descent finds global minima of deep neural networks
Simon S Du, Jason D Lee, Haochuan Li, Liwei Wang, and Xiyu Zhai · 2019
Later among the works it cites.
Gradient descent provably optimizes over-parameterized neural networks
Simon S Du, Xiyu Zhai, Barnabas Poczos, and Aarti Singh · 2019
Later among the works it cites.
Over parameterized two-level neural networks can learn nearoptimal feature representations
Cong Fang, Hanze Dong, and Tong Zhang · 2019
Later among the works it cites.
Convex formulation of overparameterized deep neural networks
Cong Fang, Yihong Gu, Weizhong Zhang, and Tong Zhang · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Song Mei, Andrea Montanari, and Phan-Minh Nguyen · 2018
Cited alongside, same era.
Grant M Rotskoff and Eric Vanden-Eijnden · 2018
Cited alongside, same era.
Regularization matters: Generalization and optimization of neural nets v.s. their induced kernel
Colin Wei, Jason D Lee, Qiang Liu, and Tengyu Ma · 2018
Cited alongside, same era.
Stochastic gradient descent optimizes over-parameterized deep relu networks
Difan Zou, Yuan Cao, Dongruo Zhou, and Quanquan Gu · 2018
Cited alongside, same era.
Fine-grained analysis of optimization and generalization for overparameterized two-layer neural networks
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, and Ruosong Wang · 2019
Cited alongside, same era.
A mean-field limit for certain deep neural networks
Dyego Araújo, Roberto I Oliveira, and Daniel Yukimura · 2019
Cited alongside, same era.
Can sgd learn recurrent neural networks with provable generalization?
Zeyuan Allen-Zhu and Yuanzhi Li · 2019
Cited alongside, same era.
Later among the works it cites.
Mean-field theory of two-layers neural networks: dimension-free bounds and kernel limit
Song Mei, Theodor Misiakiewicz, and Andrea Montanari · 2019
Later among the works it cites.
Mean field analysis of deep neural networks
Justin Sirignano and Konstantinos Spiliopoulos · 2019
Later among the works it cites.
Mean field analysis of neural networks: A central limit theorem
Justin Sirignano and Konstantinos Spiliopoulos · 2019
Later among the works it cites.
On the power and limitations of random features for understanding neural networks
Gilad Yehudai and Ohad Shamir · 2019
Later among the works it cites.
Fixup initialization: Residual learning without normalization
Hongyi Zhang, Yann N Dauphin, and Tengyu Ma · 2019
Later among the works it cites.
Mean-field analysis of two-layer neural networks: Non-asymptotic rates and generalization bounds
Zixiang Chen, Yuan Cao, Quanquan Gu, and Tong Zhang · 2020
Closest in time.
A rigorous framework for the mean field limit of multilayer neural networks
Phan-Minh Nguyen and Huy Tuan Pham · 2020
Closest in time.