J. Pennington and P. Worah, “Nonlinear random matrix theory for deep learning,” in Advances in Neural Information Processing Systems (2017) pp. 2637–2646
2017
Later among the works it cites.
M.S. Advani and A.M. Saxe, “High-dimensional dynamics of generalization error in neural networks,” (2017), arXiv:1710.03667
Original
2017
Later among the works it cites.
R. Ge, J.D. Lee, and T. Ma, “Learning one-hidden-layer neural networks with landscape design,” in ICLR (2017) arXiv:1711.00501
Original
2017
Later among the works it cites.
Y. Li and Y. Y., “Convergence analysis of two-layer neural networks with relu activation,” in Advances in Neural Information Processing Systems (2017) pp. 597–607
2017
Later among the works it cites.
D. Arpit, S. Jastrzębski, M.S. Kanwal, T. Maharaj, A. Fischer, A. Courville, and Y. Bengio, “A Closer Look at Memorization in Deep Networks,” in Proceedings of the 34th International Conference on Machine Learning (2017)
2017
Later among the works it cites.
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals, “Understanding deep learning requires rethinking generalization,” in ICLR (2017)
2017
Later among the works it cites.
M. Gabrié, A. Manoel, C. Luneau, J. Barbier, N. Macris, F. Krzakala, and L. Zdeborová, “Entropy and mutual information in models of deep neural networks,” in Advances in Neural Information Processing Systems 31 (2018) pp. 1826–1836
2018
Later among the works it cites.
E. Mossel, “Deep learning and hierarchical generative models,” (2018), arXiv:1612.09057
Original
2018
Later among the works it cites.
C. Louart, Z. Liao, and Romain Couillet, “A random matrix approach to neural networks,” The Annals of Applied Probability 28
2018
Later among the works it cites.
B. Aubin, A. Maillard, J. Barbier, F. Krzakala, N. Macris, and L. Zdeborová, “The committee machine: Computational to statistical gaps in learning a two-layers neural network,” in Advances in Neural Information Processing Systems 31 (2018) pp. 3227–3238
2018
Later among the works it cites.
M. Soltanolkotabi, A. Javanmard, and J. D. Lee, “Theoretical insights into the optimization landscape of over-parameterized shallow neural networks,” IEEE Transactions on Information Theory 65
2018
Later among the works it cites.
S. Mei, A. Montanari, and P. Nguyen, “A mean field view of the landscape of two-layer neural networks,” Proceedings of the National Academy of Sciences 115
2018
Later among the works it cites.
G.M. Rotskoff and E. Vanden-Eijnden, “Parameters as interacting particles: long time convergence and asymptotic error scaling of neural networks,” in Advances in Neural Information Processing Systems 31 (2018) pp. 7146–7155
2018
Later among the works it cites.
L. Chizat and F. Bach, “On the global convergence of gradient descent for over-parameterized models using optimal transport,” in Advances in Neural Information Processing Systems 31 (2018) pp. 3040–3050
2018
Later among the works it cites.
F. Farnia, J. Zhang, and D. Tse, “A spectral approach to generalization and optimization in neural networks,” in ICLR (2018)
2018
Later among the works it cites.
Y. Yoshida and M. Okada, “Data-dependence of plateau phenomenon in learning with neural network — statistical mechanical analysis,” in Advances in Neural Information Processing Systems 32 (2019) pp. 1720–1728
2019
Closest in time.
Z. Fan and A. Montanari, “The spectral norm of random inner-product kernel matrices,” Probability Theory and Related Fields 173
2019
Closest in time.
M.E.A. Seddik, M. Tamaazousti, and R. Couillet, “Kernel random matrices of large concentrated data: the example of gan-generated images,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (IEEE, 2019) pp. 7480–7484
2019
Closest in time.
J. Barbier, F. Krzakala, N. Macris, L. Miolane, and L. Zdeborová, “Optimal errors and phase transitions in high-dimensional generalized linear models,” Proceedings of the National Academy of Sciences 116
2019
Closest in time.
S. Goldt, M.S. Advani, A.M. Saxe, F. Krzakala, and L. Zdeborová, “Dynamics of stochastic gradient descent for two-layer neural networks in the teacher-student setup,” in Advances in Neural Information Processing Systems 32 (2019)
2019
Closest in time.
Y. Yoshida, R. Karakida, M. Okada, and S.-I. Amari, “Statistical mechanical analysis of learning dynamics of two-layer perceptron with multiple output units,” Journal of Physics A: Mathematical and Theoretical 52
2019
Closest in time.
S. Arora, N. Cohen, W. Hu, and Y. Luo, “Implicit Regularization in Deep Matrix Factorization,” in Advances in Neural Information Processing Systems 33 (2019)
2019
Closest in time.
J. Sirignano and K. Spiliopoulos, “Mean field analysis of neural networks: A central limit theorem,” Stochastic Processes and their Applications (2019), 10.1016/j.spa.2019.06.003
2019
Closest in time.
D. Kalimeris, G. Kaplun, P. Nakkiran, B. Edelman, T. Yang, B. Barak, and H. Zhang, “Sgd on neural networks learns functions of increasing complexity,” in Advances in Neural Information Processing Systems 32 (2019) pp. 3496–3506
2019
Closest in time.
N. Rahaman, A. Baratin, D. Arpit, F. Draxler, M. Lin, F. A Hamprecht, Y. Bengio, and A. Courville, “On the spectral bias of neural networks,” in ICML (2019)
2019
Closest in time.
U. Cohen, SY Chung, D.D. Lee, and H. Sompolinsky, “Separability and geometry of object manifolds in deep neural networks,” Nature communications 11
2020
Closest in time.
F. Gerace, B. Loureiro, F. Krzakala, M. Mézard, and L. Zdeborová, “Generalisation error in learning with random features and the hidden manifold model,” in 37th International Conference on Machine Learning (2020)
2020
Closest in time.
F. Mignacco, F. Krzakala, Y. M. Lu, and L. Zdeborová, “The role of regularization in classification of high-dimensional noisy gaussian mixture,” in 37th International Conference on Machine Learning (2020)
2020
Closest in time.