Provable ICA with unknown gaussian noise, with implications for gaussian mixtures and autoencoders
Arora, S., Ge, R., Moitra, A., and Sachdeva, S. (2012) · 2012
Later among the works it cites.
Making gradient descent optimal for strongly convex stochastic optimization
Rakhlin, A., Shamir, O., and Sridharan, K. (2012) · 2012
Later among the works it cites.
Learning linear transformations
Frieze, A., Jerrum, M., and Kannan, R. (1996) · 2013
Later among the works it cites.
Fast detection of overlapping communities via online tensor methods
Original
Huang, F., Niranjan, U., Hakeem, M. U., and Anandkumar, A. (2013) · 2013
Later among the works it cites.
Low-rank matrix completion using alternating minimization
Jain, P., Netrapalli, P., and Sanghavi, S. (2013) · 2013
Later among the works it cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Original
Saxe, A. M., McClelland, J. L., and Ganguli, S. (2013) · 2013
Later among the works it cites.
Contrastive learning using spectral methods
Zou, J. Y., Hsu, D., Parkes, D. C., and Adams, R. P. (2013) · 2013
Later among the works it cites.
Tensor decompositions for learning latent variable models
Anandkumar, A., Ge, R., Hsu, D., Kakade, S. M., and Telgarsky, M. (2014) · 2014
Later among the works it cites.
The loss surface of multilayer networks
Original
Choromanska, A., Henaff, M., Mathieu, M., Arous, G. B., and LeCun, Y. (2014) · 2014
Later among the works it cites.
Identifying and attacking the saddle point problem in high-dimensional non-convex optimization
Dauphin, Y. N., Pascanu, R., Gulcehre, C., Cho, K., Ganguli, S., and Bengio, Y. (2014) · 2014
Later among the works it cites.