Fetching the paper…
Reading the bibliography…
There has been a lot of recent interest in trying to characterize the error surface of deep models.
Neural networks and principal component analysis: Learning from examples without local minima
Baldi, P. and Hornik, K · 1989
Earlier work this paper cites.
Statistics of critical points of gaussian fields on large-dimensional spaces
Bray, Alan J. and Dean, David S · 2007
Earlier work this paper cites.
Replica symmetry breaking condition exposed by random matrix calculation of landscape complexity
Fyodorov, Yan V. and Williams, Ian · 2007
Earlier work this paper cites.
Mean field theory of spin glasses: statistics and dynamics
Parisi, Giorgio · 2007
Earlier work this paper cites.
Identifying and attacking the saddle point problem in high dimensional non-convex optimization
Dauphin, Yann, Pascanu, Razvan, Gulcehre, Caglar, Cho, Kyunhyun, Ganguli, Surya, and Bengio, Yoshua · 2013
Earlier work this paper cites.
Learning hierarchical category structure in deep neural networks
Saxe, Andrew, McClelland, James, and Ganguli, Surya · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, Ian J., Shlens, Jonathon, and Szegedy, Christian · 2014
Earlier work this paper cites.
Explorations on high dimensional landscapes
Sagun, Levent, Guney, Ugur, Arous, Gerard Ben, and LeCun, Yann · 2014
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural network
Saxe, Andrew, McClelland, James, and Ganguli, Surya · 2014
Earlier work this paper cites.
Deep learning in neural networks: An overview
Schmidhuber, J · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Cited alongside, same era.
The loss surfaces of multilayer networks
Choromanska, Anna, Henaff, Mikael, Mathieu, Michaël, Arous, Gérard Ben, and LeCun, Yann · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2015
Cited alongside, same era.
Deep learning
LeCun, Yann, Bengio, Yoshua, and Hinton, Geoffrey · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Rusu, Andrei A, Veness, Joel, Bellemare, Marc G, Graves, Alex, Riedmiller, Martin, Fidjeland, Andreas K, Ostrovski, Georg, et al · 2015
Qualitatively characterizing neural network optimization problems
Goodfellow, Ian J, Vinyals, Oriol, and Saxe, Andrew M · 2016
Closest in time.
Deep learning without poor local minima
Kawaguchi, Kenji · 2016
Closest in time.
Why does deep and cheap learning work so well?, 2016
Lin, Henry W. and Tegmark, Max · 2016
Closest in time.
Asynchronous methods for deep reinforcement learning
Mnih, Volodymyr, Badia, Adria Puigdomenech, Mirza, Mehdi, Graves, Alex, Lillicrap, Timothy P, Harley, Tim, Silver, David, and Kavukcuoglu, Koray · 2016
Closest in time.
Distribution-specific hardness of learning neural networks
Shamir, Ohad · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
Nguyen, Anh Mai, Yosinski, Jason, and Clune, Jeff · 2015
Cited alongside, same era.
On the quality of the initial basin in overspecified neural networks
Safran, Itay and Shamir, Ohad · 2015
Cited alongside, same era.
Robustness of classifiers: from adversarial to random noise
Fawzi, Alhussein, Moosavi-Dezfooli, Seyed-Mohsen, and Frossard, Pascal · 2016
Cited alongside, same era.
Closest in time.
Mastering the game of Go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J., Guez, Arthur, Sifre, Laurent, van den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, Dieleman, Sander, Grewe, Dominik, Nham, John, Kalchbrenner, Nal, Sutskever, Ilya, Lillicrap, Timothy, Leach, Madeleine, Kavukcuoglu, Koray, Graepel, Thore, and Hassabis, Demis · 2016
Closest in time.
No bad local minima: Data independent training error guarantees for multilayer neural networks
Soudry, Daniel and Carmon, Yair · 2016
Closest in time.
Rethinking the inception architecture for computer vision
Szegedy, Christian, Vanhoucke, Vincent, Ioffe, Sergey, Shlens, Jonathon, and Wojna, Zbigniew · 2016
Closest in time.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Yonghui, Schuster, Mike, Chen, Zhifeng, Le, Quoc V., Norouzi, Mohammad, Macherey, Wolfgang, Krikun, Maxim, Cao, Yuan, Gao, Qin, Macherey, Klaus, Klingner, Jeff, Shah, Apurva, Johnson, Melvin, Liu, Xiaobing, Kaiser, Lukasz, Gouws, Stephan, Kato, Yoshikiyo, Kudo, Taku, Kazawa, Hideto, Stevens, Keith, Kurian, George, Patil, Nishant, Wang, Wei, Young, Cliff, Smith, Jason, Riesa, Jason, Rudnick, Alex, Vinyals, Oriol, Corrado, Greg, Hughes, Macduff, and Dean, Jeffrey · 2016
Closest in time.