Fetching the paper…
Reading the bibliography…
Obtaining theoretical guarantees for neural networks training appears to be a hard problem in a general case.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Earlier work this paper cites.
Automatic differentiation in pytorch
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A · 2017
Earlier work this paper cites.
On the global convergence of gradient descent for over-parameterized models using optimal transport
Chizat, L. and Bach, F · 2018
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A., Gabriel, F., and Hongler, C · 2018
Earlier work this paper cites.
A mean field view of the landscape of two-layer neural networks
Mei, S., Montanari, A., and Nguyen, P.-M · 2018
Earlier work this paper cites.
Collective evolution of weights in wide neural networks
Yarotsky, D · 2018
Earlier work this paper cites.
On exact computation with an infinitely wide neural net
Arora, S., Du, S. S., Hu, W., Li, Z., Salakhutdinov, R. R., and Wang, R · 2019
Cited alongside, same era.
Beyond linearization: On quadratic and higher-order approximation of wide neural networks
Bai, Y. and Lee, J. D · 2019
Cited alongside, same era.
On lazy training in differentiable programming
Chizat, L., Oyallon, E., and Bach, F · 2019
Cited alongside, same era.
Convex formulation of overparameterized deep neural networks
Fang, C., Gu, Y., Zhang, W., and Zhang, T · 2019
Cited alongside, same era.
Wide neural networks of any depth evolve as linear models under gradient descent
Lee, J., Xiao, L., Schoenholz, S., Bahri, Y., Novak, R., Sohl-Dickstein, J., and Pennington, J · 2019
Cited alongside, same era.
Mean-field theory of two-layers neural networks: dimension-free bounds and kernel limit
Mei, S., Misiakiewicz, T., and Montanari, A · 2019
Later among the works it cites.
Mean field limit of the learning dynamics of multilayer neural networks
Nguyen, P.-M · 2019
Later among the works it cites.
Trainability and accuracy of neural networks: an interacting particle system approach
Rotskoff, G. M. and Vanden-Eijnden, E · 2019
Later among the works it cites.
Mean field analysis of deep neural networks
Sirignano, J. and Spiliopoulos, K · 2019
Later among the works it cites.
Mean field analysis of neural networks: A law of large numbers
Sirignano, J. and Spiliopoulos, K · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…