Fetching the paper…
Reading the bibliography…
The successes of modern deep machine learning methods are founded on their ability to transform inputs across multiple layers to build good high-level representations.
Topics in matrix analysis
Horn, R. A., Horn, R. A., and Johnson, C. R · 1994
Earlier work this paper cites.
Priors for infinite networks
Neal, R. M · 1996
Earlier work this paper cites.
Computing with infinite networks
Williams, C · 1996
Earlier work this paper cites.
Learning with kernels
Smola, A. J. and Schölkopf, B · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
Jordan, M. I., Ghahramani, Z., Jaakkola, T. S., and Saul, L. K · 1999
Earlier work this paper cites.
Probability theory: The logic of science
Jaynes, E. T · 2003
Earlier work this paper cites.
Kernel methods for pattern analysis
Shawe-Taylor, J. and Cristianini, N · 2004
Earlier work this paper cites.
Kernel methods in machine learning
Hofmann, T., Schölkopf, B., and Smola, A. J · 2008
Earlier work this paper cites.
Kernel methods for deep learning
Cho, Y. and Saul, L. K · 2009
Earlier work this paper cites.
Matrix analysis
Horn, R. A. and Johnson, C. R · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Bengio, Y., Courville, A., and Vincent, P · 2013
Earlier work this paper cites.
Deep gaussian processes
Damianou, A. and Lawrence, N. D · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
Earlier work this paper cites.
Deep learning
LeCun, Y., Bengio, Y., and Hinton, G · 2015
Earlier work this paper cites.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Earlier work this paper cites.
Variational inference: A review for statisticians
Blei, D. M., Kucukelbir, A., and McAuliffe, J. D · 2017
Cited alongside, same era.
Doubly stochastic variational inference for deep gaussian processes
Salimbeni, H. and Deisenroth, M · 2017
Cited alongside, same era.
High-dimensional gaussian copula regression: Adaptive estimation and statistical inference
Cai, T. T. and Zhang, L · 2018
Cited alongside, same era.
Deep convolutional networks as shallow gaussian processes
Garriga-Alonso, A., Rasmussen, C. E., and Aitchison, L · 2018
Cited alongside, same era.
Matrix variate distributions
Gupta, A. K. and Nagar, D. K · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A., Gabriel, F., and Hongler, C · 2018
Statistical mechanics of deep linear neural networks: The back-propagating renormalization group
Li, Q. and Sompolinsky, H · 2020
Later among the works it cites.
Predicting the outputs of finite networks trained with noisy gradients
Naveh, G., Ben-David, O., Sompolinsky, H., and Ringel, Z · 2020
Later among the works it cites.
A rigorous framework for the mean field limit of multilayer neural networks
Nguyen, P.-M. and Pham, H. T · 2020
Later among the works it cites.
Non-gaussian processes and neural networks at finite widths
Yaida, S · 2020
Later among the works it cites.
Deep kernel processes
Aitchison, L., Yang, A. X., and Ober, S. W · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep neural networks as gaussian processes
Lee, J., Bahri, Y., Novak, R., Schoenholz, S. S., Pennington, J., and Sohl-Dickstein, J · 2018
Cited alongside, same era.
Gaussian process behaviour in wide deep neural networks
Matthews, A. G. d. G., Rowland, M., Hron, J., Turner, R. E., and Ghahramani, Z · 2018
Cited alongside, same era.
A mean field view of the landscape of two-layer neural networks
Mei, S., Montanari, A., and Nguyen, P.-M · 2018
Cited alongside, same era.
Finite size corrections for neural network gaussian processes
Antognini, J. M · 2019
Cited alongside, same era.
Neural spline flows
Durkan, C., Bekasov, A., Murray, I., and Papamakarios, G · 2019
Cited alongside, same era.
Asymptotics of wide networks from feynman diagrams
Dyer, E. and Gur-Ari, G · 2019
Cited alongside, same era.
Neural networks and quantum field theory
Halverson, J., Maiti, A., and Stoner, K · 2021
Closest in time.
A self consistent theory of gaussian processes captures feature learning effects in finite cnns
Naveh, G. and Ringel, Z · 2021
Closest in time.
A variational approximate posterior for the deep wishart process
Ober, S. W. and Aitchison, L · 2021
Closest in time.
The limitations of large width in neural networks: A deep gaussian process perspective
Pleiss, G. and Cunningham, J. P · 2021
Closest in time.
The principles of deep learning theory
Roberts, D. A., Yaida, S., and Hanin, B · 2021
Closest in time.
Separation of scales and a thermodynamic description of feature learning in some cnns
Seroussi, I. and Ringel, Z · 2021
Closest in time.
Feature learning in infinite-width neural networks
Yang, G. and Hu, E. J · 2021
Closest in time.
Exact marginal prior distributions of finite bayesian neural networks
Zavatone-Veth, J. and Pehlevan, C · 2021
Closest in time.
Asymptotics of representation learning in finite bayesian neural networks
Zavatone-Veth, J. A., Canatar, A., and Pehlevan, C · 2021
Closest in time.
Deep layer-wise networks have closed-form weights
Wu, C., Masoomi, A., Gretton, A., and Dy, J · 2022
Closest in time.
Convolutional deep kernel machines
Milsom, E., Anson, B., and Aitchison, L · 2023
Closest in time.
An improved variational approximate posterior for the deep wishart process
Ober, S., Anson, B., Milsom, E., and Aitchison, L · 2023
Closest in time.