On the power of over-parametrization in neural networks with quadratic activation
Simon Du and Jason Lee · 2018
Later among the works it cites.
The committee machine: Computational to statistical gaps in learning a two-layers neural network
Benjamin Aubin, Antoine Maillard, Florent Krzakala, Nicolas Macris, Lenka Zdeborová, et al · 2018
Later among the works it cites.
Bad global minima exist and sgd can reach them
Original
Shengchao Liu, Dimitris Papailiopoulos, and Dimitris Achlioptas · 2019
Later among the works it cites.
Spurious valleys in one-hidden-layer neural network optimization landscapes
Luca Venturi, Afonso S Bandeira, and Joan Bruna · 2019
Later among the works it cites.
Passed & spurious: Descent algorithms and local minima in spiked matrix-tensor models
Stefano Sarao Mannelli, Florent Krzakala, Pierfrancesco Urbani, and Lenka Zdeborova · 2019
Later among the works it cites.
Optimal errors and phase transitions in high-dimensional generalized linear models
Jean Barbier, Florent Krzakala, Nicolas Macris, Léo Miolane, and Lenka Zdeborová · 2019
Later among the works it cites.
Dynamics of stochastic gradient descent for two-layer neural networks in the teacher-student setup
Sebastian Goldt, Madhu Advani, Andrew M Saxe, Florent Krzakala, and Lenka Zdeborová · 2019
Later among the works it cites.
Gradient descent with random initialization: Fast global convergence for nonconvex phase retrieval
Yuxin Chen, Yuejie Chi, Jianqing Fan, and Cong Ma · 2019
Later among the works it cites.
Stationary points of shallow neural networks with quadratic activation function
Original
David Gamarnik, Eren C Kızıldağ, and Ilias Zadik · 2019
Later among the works it cites.
Complex dynamics in simple neural networks: Understanding gradient flow in phase retrieval
Original
Stefano Sarao Mannelli, Giulio Biroli, Chiara Cammarota, Florent Krzakala, Pierfrancesco Urbani, and Lenka Zdeborová · 2020
Closest in time.