Fetching the paper…
Reading the bibliography…
We propose a theoretical understanding of neural networks in terms of Wilsonian effective field theory.
Tentativo di una teoria dell’emissione dei raggi "beta "
E Fermi · 1933
Earlier work this paper cites.
Cognitron: A self-organizing multilayered neural network
Kunihiko Fukushima · 1975
Earlier work this paper cites.
Neocognitron: A self-organizing neural network model for a mechanism of pattern recognition unaffected by shift in position
Kunihiko Fukushima · 1980
Earlier work this paper cites.
Learning internal representations by error propagation
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams · 1985
Earlier work this paper cites.
BAYESIAN LEARNING FOR NEURAL NETWORKS
Radford M Neal · 1995
Earlier work this paper cites.
Computing with infinite networks
Christopher KI Williams · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Object recognition with gradient-based learning
Yann LeCun, Patrick Haffner, Léon Bottou, and Yoshua Bengio · 1999
Earlier work this paper cites.
Quantum field theory in a nutshell
A. Zee · 2003
Earlier work this paper cites.
Chameleon cosmology
Justin Khoury and Amanda Weltman · 2004
Earlier work this paper cites.
Kernel methods for deep learning
Youngmin Cho and Lawrence K Saul · 2009
Earlier work this paper cites.
Spectral networks and locally connected networks on graphs
Joan Bruna, Wojciech Zaremba, Arthur Szlam, and Yann LeCun · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Quantum Field Theory and the Standard Model
Matthew D. Schwartz · 2014
Earlier work this paper cites.
Deep convolutional networks on graph-structured data
Mikael Henaff, Joan Bruna, and Yann LeCun · 2015
Cited alongside, same era.
Convolutional networks on graphs for learning molecular fingerprints
David K Duvenaud, Dougal Maclaurin, Jorge Iparraguirre, Rafael Bombarell, Timothy Hirzel, Alan Aspuru-Guzik, and Ryan P Adams · 2015
Cited alongside, same era.
Gated graph sequence neural networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard Zemel · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Deep convolutional networks as shallow gaussian processes
Adrià Garriga-Alonso, Laurence Aitchison, and Carl E. Rasmussen · 2019
Later among the works it cites.
Greg Yang · 2019
Later among the works it cites.
Greg Yang · 2019
Later among the works it cites.
Wide neural networks of any depth evolve as linear models under gradient descent
Jaehoon Lee, Lechao Xiao, Samuel S. Schoenholz, Yasaman Bahri, Roman Novak, Jascha Sohl-Dickstein, and Jeffrey Pennington · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Densely connected convolutional networks, 2016
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger · 2016
Cited alongside, same era.
Convolutional neural networks on graphs with fast localized spectral filtering
Michaël Defferrard, Xavier Bresson, and Pierre Vandergheynst · 2016
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling · 2016
Cited alongside, same era.
Layer normalization, 2016
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E. Hinton · 2016
Cited alongside, same era.
Deep neural networks as gaussian processes, 2017
Jaehoon Lee, Yasaman Bahri, Roman Novak, Samuel S. Schoenholz, Jeffrey Pennington, and Jascha Sohl-Dickstein · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Gaussian process behaviour in wide deep neural networks
Alexander G. de G. Matthews, Mark Rowland, Jiri Hron, Richard E. Turner, and Zoubin Ghahramani · 2018
Cited alongside, same era.
Joseph M. Antognini · 2019
Later among the works it cites.
Non-gaussian processes and neural networks at finite widths
Sho Yaida · 2019
Later among the works it cites.
A fine-grained spectral perspective on neural networks
Greg Yang and Hadi Salman · 2019
Later among the works it cites.
Tensor programs ii: Neural tangent kernel for any architecture
Greg Yang · 2020
Closest in time.
Predicting the outputs of finite networks trained with noisy gradients, 2020
Gadi Naveh, Oded Ben-David, Haim Sompolinsky, and Zohar Ringel · 2020
Closest in time.
Learning curves for deep neural networks: A gaussian field theory perspective, 2020
Omry Cohen, Or Malka, and Zohar Ringel · 2020
Closest in time.
Asymptotics of wide networks from feynman diagrams
Ethan Dyer and Guy Gur-Ari · 2020
Closest in time.
Tensor programs iii: Neural matrix laws, 2020
Greg Yang · 2020
Closest in time.
Feature learning in infinite-width neural networks, 2020
Greg Yang and Edward J. Hu · 2020
Closest in time.