A. M. Saxe, J. L. McClelland, and S. Ganguli, A mathematical theory of semantic development in deep neural networks, Proc. Natl. Acad. Sci. USA 116
2019
Later among the works it cites.
A. Garriga-Alonso, C. E. Rasmussen, and L. Aitchison, Deep convolutional networks as shallow gaussian processes, in International Conference on Learning Representations (2019)
2019
Later among the works it cites.
I. Loshchilov and F. Hutter, Decoupled weight decay regularization, in International Conference on Learning Representations (2019)
2019
Later among the works it cites.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala, Pytorch: An imperative style, high-performance deep learning library, in Adv. Neural Inf. Process. Syst. , Vol. 32, edited by H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alché-Buc, E. Fox, and R. Garnett (Curran Associates, Inc., 2019) pp. 8024–8035
2019
Later among the works it cites.
Y. Bahri, J. Kadmon, J. Pennington, S. S. Schoenholz, J. Sohl-Dickstein, and S. Ganguli, Statistical mechanics of deep learning, Annu. Rev. Condens. Matter Phys. 11
2020
Later among the works it cites.
M. Helias and D. Dahmen, Statistical Field Theory for Neural Networks (Springer International Publishing, 2020) p. 203
2020
Later among the works it cites.
E. Dyer and G. Gur-Ari, Asymptotics of wide networks from feynman diagrams, in International Conference on Learning Representations (2020)
2020
Later among the works it cites.
S. Yaida, Non-Gaussian processes and neural networks at finite widths, in Proceedings of The First Mathematical and Scientific Machine Learning Conference , Proceedings of Machine Learning Research, Vol. 107, edited by J. Lu and R. Ward (PMLR, Princeton University, Princeton, NJ, USA, 2020) pp. 165–192
2020
Later among the works it cites.
S. Goldt, M. Mézard, F. Krzakala, and L. Zdeborová, Modeling the Influence of Data Structure on Learning in Neural Networks: The Hidden Manifold Model, Phys. Rev. X 10
2020
Later among the works it cites.
M. E. A. Seddik, C. Louart, M. Tamaazousti, and R. Couillet, Random Matrix Theory Proves that Deep Learning Representations of GAN-data Behave as Gaussian Mixtures, in International Conference on Machine Learning (PMLR, 2020) pp. 8573–8582
2020
Later among the works it cites.
O. Cohen, O. Malka, and Z. Ringel, Learning curves for overparametrized deep neural networks: A field theory perspective, Phys. Rev. Res. 3
2021
Later among the works it cites.
G. Naveh, O. Ben David, H. Sompolinsky, and Z. Ringel, Predicting the outputs of finite deep neural networks trained with noisy gradients, Phys. Rev. E 104
2021
Later among the works it cites.
G. Yang and E. J. Hu, Tensor programs iv: Feature learning in infinite-width neural networks, in Proceedings of the 38th International Conference on Machine Learning , Proceedings of Machine Learning Research, Vol. 139, edited by M. Meila and T. Zhang (PMLR, 2021) pp. 11727–11737
2021
Later among the works it cites.
J. Zhou and H. Huang, Weakly correlated synapses promote dimension reduction in deep neural networks, Phys. Rev. E 103
2021
Later among the works it cites.
S. Goldt, B. Loureiro, G. Reeves, F. Krzakala, M. Mezard, and L. Zdeborova, The gaussian equivalence of generative models for learning with shallow neural networks, in Proceedings of the 2nd Mathematical and Scientific Machine Learning Conference , Proceedings of Machine Learning Research, Vol. 145, edited by J. Bruna, J. Hesthaven, and L. Zdeborova (PMLR, 2022) pp. 426–471
2022
Closest in time.
B. Loureiro, C. Gerbelot, H. Cui, S. Goldt, F. Krzakala, M. Mèzard, and L. Zdeborová, Learning curves of generic features maps for realistic datasets with a teacher-student model, J. Stat. Mech. Theory Exp. 2022
2022
Closest in time.
D. A. Roberts, S. Yaida, and B. Hanin, The Principles of Deep Learning Theory (Cambridge University Press, 2022)
2022
Closest in time.
K. Segadlo, B. Epping, A. van Meegen, D. Dahmen, M. Krämer, and M. Helias, Unified field theoretical approach to deep and recurrent neuronal networks, J. Stat. Mech. Theory Exp. 2022
2022
Closest in time.
B. Bordelon and C. Pehlevan, Self-consistent dynamical field theory of kernel evolution in wide neural networks, arXiv preprint arXiv:2205.09653 (2022)
Original
2022
Closest in time.