On the global convergence of gradient descent for over-parameterized models using optimal transport
Lenaic Chizat and Francis Bach · 2018
Cited alongside, same era.
Gradient descent provably optimizes over-parameterized neural networks
Original
Simon S Du, Xiyu Zhai, Barnabas Poczos, and Aarti Singh · 2018
Cited alongside, same era.
A priori estimates of the population risk for two-layer neural networks
Original
Weinan E, Chao Ma, and Lei Wu · 2018
Cited alongside, same era.
A mean field view of the landscape of two-layer neural networks
Song Mei, Andrea Montanari, and Phan-Minh Nguyen · 2018
Cited alongside, same era.
Fast learning requires good memory: A time-space lower bound for parity learning
Ran Raz · 2018
Cited alongside, same era.
Neural networks as interacting particle systems: Asymptotic convexity of the loss landscape and universal scaling of the approximation error
Original
Grant M Rotskoff and Eric Vanden-Eijnden · 2018
Cited alongside, same era.
Distribution-specific hardness of learning neural networks
Ohad Shamir · 2018
Cited alongside, same era.
Maximum mean discrepancy gradient flow
Michael Arbel, Anna Korba, Adil Salim, and Arthur Gretton · 2019
Cited alongside, same era.
A mean-field limit for certain deep neural networks
Original
Dyego Araújo, Roberto I Oliveira, and Daniel Yukimura · 2019
Cited alongside, same era.
Barron spaces and the compositional function spaces for neural network models
Original
Weinan E, Chao Ma, and Lei Wu · 2019
Cited alongside, same era.
Machine learning from a continuous viewpoint
Original
Weinan E, Chao Ma, and Lei Wu · 2019
Cited alongside, same era.