Fetching the paper…
Reading the bibliography…
Mode connectivity is a surprising phenomenon in the loss landscape of deep nets.
On connected sublevel sets in deep learning
Nguyen, Q. (2019) · 1901
Earlier work this paper cites.
User-friendly tail bounds for sums of random matrices
Tropp, J. A. (2012) · 2012
Earlier work this paper cites.
Identifying and attacking the saddle point problem in high-dimensional non-convex optimization
Dauphin, Y. N., Pascanu, R., Gulcehre, C., Cho, K., Ganguli, S., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A. (2014) · 2014
Earlier work this paper cites.
The loss surfaces of multilayer networks
Choromanska, A., Henaff, M., Mathieu, M., Arous, G. B., and LeCun, Y. (2015) · 2015
Earlier work this paper cites.
Efficient object localization using convolutional networks
Tompson, J., Goroshin, R., Jain, A., LeCun, Y., and Bregler, C. (2015) · 2015
Earlier work this paper cites.
Topology and geometry of half-rectified network optimization
Freeman, C. D. and Bruna, J. (2016) · 2016
Cited alongside, same era.
Deep learning without poor local minima
Kawaguchi, K. (2016) · 2016
Cited alongside, same era.
Stronger generalization bounds for deep nets via a compression approach
Arora, S., Ge, R., Neyshabur, B., and Zhang, Y. (2018) · 2018
Cited alongside, same era.
Essentially no barriers in neural network energy landscape
Draxler, F., Veschgini, K., Salmhofer, M., and Hamprecht, F. A. (2018) · 2018
Cited alongside, same era.
Loss surfaces, mode connectivity, and fast ensembling of dnns
Garipov, T., Izmailov, P., Podoprikhin, D., Vetrov, D. P., and Wilson, A. G. (2018) · 2018
Cited alongside, same era.
Understanding the loss surface of neural networks for binary classification
Liang, S., Sun, R., Li, Y., and Srikant, R. (2018) · 2018
Later among the works it cites.
On the importance of single directions for generalization
Morcos, A. S., Barrett, D. G., Rabinowitz, N. C., and Botvinick, M. (2018) · 2018
Later among the works it cites.
On the loss landscape of a class of deep neural networks with no bad local valleys
Nguyen, Q., Mukkamala, M. C., and Hein, M. (2018) · 2018
Later among the works it cites.
Spurious local minima are common in two-layer relu neural networks
Safran, I. and Shamir, O. (2018) · 2018
Later among the works it cites.
Spurious valleys in two-layer neural network optimization landscapes
Venturi, L., Bandeira, A. S., and Bruna, J. (2018) · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Keshari, R., Singh, R., and Vatsa, M. (2018) · 2018
Cited alongside, same era.
Later among the works it cites.
A critical view of global optimality in deep learning
Yun, C., Sra, S., and Jadbabaie, A. (2018) · 2018
Later among the works it cites.