Fetching the paper…
Reading the bibliography…
Recent work by Jacot et al.
Mean-field behaviour of neural tangent kernel for deep neural networks
S. Hayou, A. Doucet, and J. Rousseau · 1905
Earlier work this paper cites.
Mean-field behaviour of neural tangent kernel for deep neural networks
S. Hayou, A. Doucet, and J. Rousseau · 1905
Earlier work this paper cites.
Stochastic gradient descent optimizes over-parameterized deep ReLU networks
D. Zou, Y. Cao, D. Zhou, and Q. Gu · 1905
Earlier work this paper cites.
Spherical Harmonics: An Elementary Treatise on Harmonic Functions, with Applications
T.M. MacRobert · 1967
Earlier work this paper cites.
Spherical Harmonics: An Elementary Treatise on Harmonic Functions, with Applications
T.M. MacRobert · 1967
Earlier work this paper cites.
Bayesian learning for neural networks
R.M. Neal · 1995
Earlier work this paper cites.
Bayesian learning for neural networks
R.M. Neal · 1995
Earlier work this paper cites.
Random synaptic feedback weights support error backpropagation for deep learning
T. Lillicrap, D. Cownden, D. Tweed, and C. Akerman · 2016
Earlier work this paper cites.
Exponential expressivity in deep neural networks through transient chaos
B. Poole, S. Lahiri, M. Raghu, J. Sohl-Dickstein, and S. Ganguli · 2016
Earlier work this paper cites.
Random synaptic feedback weights support error backpropagation for deep learning
T. Lillicrap, D. Cownden, D. Tweed, and C. Akerman · 2016
Earlier work this paper cites.
Exponential expressivity in deep neural networks through transient chaos
B. Poole, S. Lahiri, M. Raghu, J. Sohl-Dickstein, and S. Ganguli · 2016
Earlier work this paper cites.
Deep information propagation
S.S. Schoenholz, J. Gilmer, S. Ganguli, and J. Sohl-Dickstein · 2017
Earlier work this paper cites.
Mean field residual networks: On the edge of chaos
G. Yang and S. Schoenholz · 2017
Earlier work this paper cites.
Understanding deep learning requires rethinking generalization
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals · 2017
Earlier work this paper cites.
Deep information propagation
S.S. Schoenholz, J. Gilmer, S. Ganguli, and J. Sohl-Dickstein · 2017
Earlier work this paper cites.
Mean field residual networks: On the edge of chaos
G. Yang and S. Schoenholz · 2017
Earlier work this paper cites.
Understanding deep learning requires rethinking generalization
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals · 2017
Earlier work this paper cites.
Gradient descent learns one-hidden-layer CNN: Don’t be afraid of spurious local minima
S.S. Du, J.D. Lee, Y. Tian, B. Poczos, and A Singh · 2018
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Earlier work this paper cites.
In ICLR , 2018
J. Lee, Y. Bahri, R. Novak, S.S. Schoenholz, J. Pennington, and J. Sohl-Dickstein · 2018
Earlier work this paper cites.
Gaussian process behaviour in wide deep neural networks
A.G. Matthews, J. Hron, M. Rowland, R.E. Turner, and Z. Ghahramani · 2018
Earlier work this paper cites.
Optimization landscape and expressivity of deep CNNs
Q. Nguyen and M. Hein · 2018
Earlier work this paper cites.
Dynamical isometry and a mean field theory of cnns: How to train 10,000-layer vanilla convolutional neural networks
L. Xiao, Y. Bahri, J. Sohl-Dickstein, S. S. Schoenholz, and P. Pennington · 2018
Earlier work this paper cites.
Stochastic gradient descent optimizes over-parameterized deep ReLU networks
D. Zou, Y. Cao, D. Zhou, and Q. Gu · 2018
Cited alongside, same era.
Gradient descent learns one-hidden-layer CNN: Don’t be afraid of spurious local minima
S.S. Du, J.D. Lee, Y. Tian, B. Poczos, and A Singh · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Cited alongside, same era.
In ICLR , 2018
J. Lee, Y. Bahri, R. Novak, S.S. Schoenholz, J. Pennington, and J. Sohl-Dickstein · 2018
Cited alongside, same era.
Gaussian process behaviour in wide deep neural networks
A.G. Matthews, J. Hron, M. Rowland, R.E. Turner, and Z. Ghahramani · 2018
Cited alongside, same era.
Optimization landscape and expressivity of deep CNNs
Neural tangents: Fast and easy infinite neural networks in python
Roman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee, Alexander A. Alemi, Jascha Sohl-Dickstein, and Samuel S. Schoenholz · 2020
Closest in time.
Disentangling trainability and generalization in deep neural networks
Lechao Xiao, Jeffrey Pennington, and Samuel Schoenholz · 2020
Closest in time.
Tensor programs iii: Neural matrix laws
G. Yang · 2020
Closest in time.
On the similarity between the laplace and neural tangent kernels
A. Geifman, A. Yadav, Y. Kasten, M. Galun, D. Jacobs, and R. Basri · 2020
Closest in time.
Finite depth and width corrections to the neural tangent kernel
B. Hanin and M. Nica · 2020
Closest in time.
Dynamics of deep neural networks and neural tangent hierarchy
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Q. Nguyen and M. Hein · 2018
Cited alongside, same era.
Dynamical isometry and a mean field theory of cnns: How to train 10,000-layer vanilla convolutional neural networks
L. Xiao, Y. Bahri, J. Sohl-Dickstein, S. S. Schoenholz, and P. Pennington · 2018
Cited alongside, same era.
On the inductive bias of neural tangent kernels
A. Bietti and J. Mairal · 2019
Cited alongside, same era.
Generalization bounds of stochastic gradient descent for wide and deep neural networks
Y. Cao and Q. Gu · 2019
Cited alongside, same era.
A note on lazy training in supervised differentiable programming
L. Chizat and F. Bach · 2019
Cited alongside, same era.
Universal statistics of Fisher information in deep neural networks: Mean field approach
R. Karakida, S. Akaho, and S. Amari · 2019
Cited alongside, same era.
Wide neural networks of any depth evolve as linear models under gradient descent
J. Lee, L. Xiao, S. Schoenholz, Y. Bahri, J. Sohl-Dickstein, and J. Pennington · 2019
Cited alongside, same era.
J. Huang and H.T Yau · 2020
Closest in time.
Why do deep residual networks generalize better than deep feedforward networks? – a neural tangent kernel perspective
K. Huang, Y. Wang, M. Tao, and T. Zhao · 2020
Closest in time.
Neural tangents: Fast and easy infinite neural networks in python
Roman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee, Alexander A. Alemi, Jascha Sohl-Dickstein, and Samuel S. Schoenholz · 2020
Closest in time.
Disentangling trainability and generalization in deep neural networks
Lechao Xiao, Jeffrey Pennington, and Samuel Schoenholz · 2020
Closest in time.
Tensor programs iii: Neural matrix laws
G. Yang · 2020
Closest in time.
Deep equals shallow for reLU networks in kernel regimes
A. Bietti and F. Bach · 2021
Closest in time.
Towards understanding the spectral bias of deep learning
Y. Cao, Z. Fang, Y. Wu, D. Zhou, and Q. Gu · 2021
Closest in time.
Deep neural tangent kernel and laplace kernel have the same rkhs
L. Chen and S. Xu · 2021
Closest in time.
Linearized two-layers neural networks in high dimension
B. Ghorbani, S. Mei, T. Misiakiewicz, and A. Montanari · 2021
Closest in time.
The spectrum of Fisher information of deep networks achieving dynamical isometry
T. Hayase and R. Karakida · 2021
Closest in time.
Stable resnet
S. Hayou, E. Clerico, B. He, G. Deligiannidis, A. Doucet, and J. Rousseau · 2021
Closest in time.
Deep equals shallow for reLU networks in kernel regimes
A. Bietti and F. Bach · 2021
Closest in time.
Towards understanding the spectral bias of deep learning
Y. Cao, Z. Fang, Y. Wu, D. Zhou, and Q. Gu · 2021
Closest in time.
Deep neural tangent kernel and laplace kernel have the same rkhs
L. Chen and S. Xu · 2021
Closest in time.
Linearized two-layers neural networks in high dimension
B. Ghorbani, S. Mei, T. Misiakiewicz, and A. Montanari · 2021
Closest in time.
The spectrum of Fisher information of deep networks achieving dynamical isometry
T. Hayase and R. Karakida · 2021
Closest in time.
Stable resnet
S. Hayou, E. Clerico, B. He, G. Deligiannidis, A. Doucet, and J. Rousseau · 2021
Closest in time.