Mathematical analysis
Apostol, T. M · 1974
Earlier work this paper cites.
The set of all mxn rectangular real matrices of rank-r is connected by analytic regular arcs
Evard, J. C. and Jafari, F · 1994
Earlier work this paper cites.
Learning polynomials with neural networks
Andoni, A., Panigrahy, R., Valiant, G., and Zhang, L · 2014
Earlier work this paper cites.
On the computational efficiency of training neural networks
Livni, R., Shalev-Shwartz, S., and Shamir, O · 2014
Earlier work this paper cites.
The loss surfaces of multilayer networks
Choromanska, A., Hena, M., Mathieu, M., Arous, G. B., and LeCun, Y · 2015
Earlier work this paper cites.
Provable methods for training neural networks with sparse connectivity
Sedghi, H. and Anandkumar, A · 2015
Earlier work this paper cites.
Fast and accurate deep network learning by exponential linear units (elus)
Clevert, D., Unterthiner, T., and Hochreiter, S · 2016
Earlier work this paper cites.
The power of depth for feedforward neural networks
Eldan, R. and Shamir, O · 2016
Earlier work this paper cites.
Globally optimal training of generalized polynomial neural networks with nonlinear spectral methods
Gautier, A., Nguyen, Q., and Hein, M · 2016
Earlier work this paper cites.
Beating the perils of non-convexity: Guaranteed training of neural networks using tensor methods
Original
Janzamin, M., Sedghi, H., and Anandkumar, A · 2016
Earlier work this paper cites.
On the quality of the initial basin in overspecified networks
Safran, I. and Shamir, O · 2016
Earlier work this paper cites.
Benefits of depth in neural networks
Telgarsky, M · 2016
Earlier work this paper cites.
Globally optimal gradient descent for a convnet with gaussian inputs
Brutzkus, A. and Globerson, A · 2017
Earlier work this paper cites.
Global optimality in neural network training
Haeffele, B. D. and Vidal, R · 2017
Earlier work this paper cites.