Stochastic variance reduction for nonconvex optimization
Reddi, S. J., Hefny, A., Sra, S., Poczos, B., and Smola, A · 2016
Cited alongside, same era.
Finding approximate local minima faster than gradient descent
Agarwal, N., Allen-Zhu, Z., Bullins, B., Hazan, E., and Ma, T · 2017
Cited alongside, same era.
Lower bounds for finding stationary points I
Carmon, Y., Duchi, J. C., Hinder, O., and Sidford, A · 2017
Cited alongside, same era.
How to escape saddle points efficiently
Jin, C., Ge, R., Netrapalli, P., Kakade, S. M., and Jordan, M. I · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A · 2017
Cited alongside, same era.
How to make the gradients small stochastically: Even faster convex and nonconvex SGD
Allen-Zhu, Z · 2018
Cited alongside, same era.
First order methods beyond convexity and Lipschitz gradient continuity with applications to quadratic inverse problems
Bolte, J., Sabach, S., Teboulle, M., and Vaisbourd, Y · 2018
Cited alongside, same era.
Gradient sampling methods for nonsmooth optimization
Original
Burke, J. V., Curtis, F. E., Lewis, A. S., Overton, M. L., and Simões, L. E · 2018
Cited alongside, same era.
Accelerated methods for nonconvex optimization
Carmon, Y., Duchi, J. C., Hinder, O., and Sidford, A · 2018
Cited alongside, same era.
Escaping saddles with stochastic gradients
Daneshmand, H., Kohler, J., Lucchi, A., and Hofmann, T · 2018
Cited alongside, same era.
Stochastic subgradient method converges on tame functions
Davis, D., Drusvyatskiy, D., Kakade, S., and Lee, J. D · 2018
Cited alongside, same era.
Stochastic methods for composite and weakly convex optimization problems
Duchi, J. C. and Ruan, F · 2018
Cited alongside, same era.