Fetching the paper…
Reading the bibliography…
Despite remarkable performance on a variety of tasks, many properties of deep neural networks are not yet theoretically understood.
Kernel methods for deep learning
Youngmin Cho and Lawrence Saul · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Earlier work this paper cites.
The MNIST database of handwritten digit images for machine learning research
Li Deng · 2012
Earlier work this paper cites.
Generalized Bessel numbers and some combinatorial settings
Gi-Sang Cheon, Ji-Hwan Jung, and Louis W. Shapiro · 2013
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng · 2015
Earlier work this paper cites.
The power of depth for feedforward neural networks
Ronen Eldan and Ohad Shamir · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Integer sequences connected to the Laplace continued fraction and Ramanujan’s identity
Alexander Kreinin · 2016
Earlier work this paper cites.
Exponential expressivity in deep neural networks through transient chaos
Ben Poole, Subhaneil Lahiri, Maithra Raghu, Jascha Sohl-Dickstein, and Surya Ganguli · 2016
Cited alongside, same era.
Deep information propagation
Samuel S. Schoenholz, Justin Gilmer, Surya Ganguli, and Jascha Sohl-Dickstein · 2017
Cited alongside, same era.
Fashion-MNIST: a novel image dataset for benchmarking machine learning algorithms, 2017
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Cited alongside, same era.
Which neural net architectures give rise to exploding and vanishing gradients?
Boris Hanin · 2018
Cited alongside, same era.
High-Dimensional Probability: An Introduction with Applications in Data Science
Roman Vershynin · 2018
Cited alongside, same era.
Neural architecture search: a survey
James Martens, Andy Ballard, Guillaume Desjardins, Grzegorz Swirszcz, Valentin Dalibard, Jascha Sohl-Dickstein, and Samuel S. Schoenholz · 2021
Later among the works it cites.
Deep limits and a cut-off phenomenon for neural networks
Benny Avelin and Anders Karlsson · 2022
Later among the works it cites.
Why neural networks find simple solutions: The many regularizers of geometric complexity
Benoit Dherin, Michael Munn, Mihaela Rosca, and David GT Barrett · 2022
Later among the works it cites.
The neural covariance SDE: Shaped infinite depth-and-width networks at initialization
Mufan Bill Li, Mihai Nica, and Daniel M. Roy · 2022
Later among the works it cites.
A Johnson-Lindenstrauss framework for randomly initialized CNNs
Ido Nachum, Jan Hazla, Michael Gastpar, and Anatoly Khina · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter · 2019
Cited alongside, same era.
On the impact of the activation function on deep neural networks training
Soufiane Hayou, Arnaud Doucet, and Judith Rousseau · 2019
Cited alongside, same era.
Deep networks and the multiple manifold problem
Sam Buchanan, Dar Gilboa, and John Wright · 2021
Cited alongside, same era.
The Principles of Deep Learning Theory: An Effective Theory Approach to Understanding Neural Networks
Daniel A. Roberts, Sho Yaida, and Boris Hanin · 2022
Later among the works it cites.
Random fully connected neural networks as perturbatively solvable hierarchies, 2023
Boris Hanin · 2023
Closest in time.