“Wide neural networks of any depth evolve as linear models under gradient descent”, 2019
Original
Jaehoon Lee et al · 1902
Earlier work this paper cites.
“Scaling limits of wide neural networks with weight sharing: Gaussian process behavior, gradient independence, and neural tangent kernel derivation”, 2019
Original
Greg Yang · 1902
Earlier work this paper cites.
“Bayesian learning for neural networks”
Radford Neal · 1995
Earlier work this paper cites.
“Multidimensional diffusion processes”
Daniel Stroock and SR Varadhan · 1997
Earlier work this paper cites.
“Tensor programs ii: Neural tangent kernel for any architecture”, 2020
Original
Greg Yang · 2006
Earlier work this paper cites.
“Kernel methods for deep learning”
Youngmin Cho and Lawrence Saul · 2009
Earlier work this paper cites.
“Markov processes: characterization and convergence”
Stewart Ethier and Thomas Kurtz · 2009
Earlier work this paper cites.
“Feature Learning in Infinite-Width Neural Networks”
Original
Greg Yang and Edward. Hu · 2011
Earlier work this paper cites.
“Analyzing Finite Neural Networks: Can We Trust Neural Tangent Kernel Theory?”, 2020
Original
Mariia Seleznova and Gitta Kutyniok · 2012
Earlier work this paper cites.
“Brownian motion and stochastic calculus”
Ioannis Karatzas and Steven Shreve · 2012
Earlier work this paper cites.
“Delving deep into rectifiers: Surpassing human-level performance on imagenet classification”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2015
Earlier work this paper cites.
“Deep information propagation”
Original
Samuel Schoenholz, Justin Gilmer, Surya Ganguli and Jascha Sohl-Dickstein · 2016
Earlier work this paper cites.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Earlier work this paper cites.
“Mean field residual networks: on the edge of chaos”
Greg Yang and Samuel Schoenholz · 2017
Earlier work this paper cites.
“SymPy: symbolic computing in Python”
Aaron Meurer et al · 2017
Earlier work this paper cites.
“Deep Neural Networks as Gaussian Processes”
Jaehoon Lee et al · 2018
Earlier work this paper cites.
“Neural tangent kernel: Convergence and generalization in neural networks”
Original
Arthur Jacot, Franck Gabriel and Clément Hongler · 2018
Earlier work this paper cites.
“Trainability and Accuracy of Neural Networks: An Interacting Particle System Approach”, 2018
Original
Grant. Rotskoff and Eric Vanden-Eijnden · 2018
Earlier work this paper cites.