Fetching the paper…
Reading the bibliography…
Inference in deep Bayesian neural networks is only fully understood in the infinite-width limit, where the posterior flexibility afforded by increased depth washes out and the posterior predictive collapses to a shallow Gaussian process.
C. S. Herz, “Bessel functions of matrix argument,” Annals of Mathematics , pp. 474–523, 1955
1955
Earlier work this paper cites.
D. J. MacKay, “A practical Bayesian framework for backpropagation networks,” Neural Computation , vol. 4, no. 3, pp. 448–472, 1992
1992
Earlier work this paper cites.
Z. Shun and P. McCullagh, “Laplace approximation of high dimensional integrals,” Journal of the Royal Statistical Society: Series B (Methodological) , vol. 57, no. 4, pp. 749–760, 1995
1995
Earlier work this paper cites.
R. M. Neal, “Priors for infinite networks,” in Bayesian Learning for Neural Networks . Springer, 1996, pp. 29–53
1996
Earlier work this paper cites.
C. K. Williams, “Computing with infinite networks,” Advances in Neural Information Processing Systems , pp. 295–301, 1997
1997
Earlier work this paper cites.
K. Fukumizu, “Effect of batch learning in multilayer neural networks,” in Proceedings of the 5th International Conference on Neural Information Processing , 1998, pp. 67–70
1998
Earlier work this paper cites.
R. W. Butler, “Generalized inverse Gaussian distributions and their Wishart connections,” Scandinavian Journal of Statistics , vol. 25, no. 1, pp. 69–75, 1998
1998
Earlier work this paper cites.
R. W. Butler and A. T. Wood, “Laplace approximation for Bessel functions of matrix argument,” Journal of Computational and Applied Mathematics , vol. 155, no. 2, pp. 359–382, 2003
2003
Earlier work this paper cites.
C. K. Williams and C. E. Rasmussen, Gaussian processes for machine learning . MIT press Cambridge, MA, 2006, vol. 2, no. 3
2006
Earlier work this paper cites.
R. A. Horn and C. R. Johnson, Matrix Analysis . Cambridge University Press, 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei, “ImageNet Large Scale Visual Recognition Challenge,” International Journal of Computer Vision (IJCV) , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
F. Fazayeli and A. Banerjee, “The matrix generalized inverse Gaussian distribution: Properties and applications,” in Joint European Conference on Machine Learning and Knowledge Discovery in Databases . Springer, 2016, pp. 648–664
2016
Earlier work this paper cites.
J. Bun, R. Allez, J.-P. Bouchaud, and M. Potters, “Rotational invariant estimator for general noisy matrices,” IEEE Transactions on Information Theory , vol. 62, no. 12, pp. 7475–7490, 2016
2016
Cited alongside, same era.
J. Lee, J. Sohl-Dickstein, J. Pennington, R. Novak, S. Schoenholz, and Y. Bahri, “Deep neural networks as Gaussian processes,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
A. G. d. G. Matthews, J. Hron, M. Rowland, R. E. Turner, and Z. Ghahramani, “Gaussian process behaviour in wide deep neural networks,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
R. Vershynin, High-dimensional probability: An introduction with applications in data science . Cambridge University Press, 2018, vol. 47
2018
Cited alongside, same era.
2021
Closest in time.
C. Yun, S. Krishnan, and H. Mobahi, “A unifying view on implicit bias in training linear neural networks,” in International Conference on Learning Representations , 2021
2021
Closest in time.
2021
Closest in time.
Q. Li and H. Sompolinsky, “Statistical mechanics of deep linear neural networks: The backpropagating kernel renormalization,” Phys. Rev. X , vol. 11, p. 031059, 09 2021
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
J. R. Magnus and H. Neudecker, Matrix differential calculus with applications in statistics and econometrics . John Wiley & Sons, 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
J. Lee, S. Schoenholz, J. Pennington, B. Adlam, L. Xiao, R. Novak, and J. Sohl-Dickstein, “Finite versus infinite neural networks: an empirical study,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 15 156–15 172
2020
Cited alongside, same era.
L. Aitchison, “Why bigger is not always better: on finite and infinite neural networks,” in Proceedings of the 37th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, H. Daumé III and A. Singh, Eds., vol. 119. PMLR, 07 2020, pp. 156–164
2020
Cited alongside, same era.
A. G. Wilson and P. Izmailov, “Bayesian deep learning and a probabilistic perspective of generalization,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 4697–4708
2020
Cited alongside, same era.
J. A. Zavatone-Veth and C. Pehlevan, “Exact marginal prior distributions of finite Bayesian neural networks,” in Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Cited alongside, same era.
J. A. Zavatone-Veth, A. Canatar, B. S. Ruben, and C. Pehlevan, “Asymptotics of representation learning in finite Bayesian neural networks,” in Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Cited alongside, same era.
2021
Closest in time.
L. Aitchison, A. Yang, and S. W. Ober, “Deep kernel processes,” in Proceedings of the 38th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, M. Meila and T. Zhang, Eds., vol. 139. PMLR, 18–24 Jul 2021, pp. 130–140
2021
Closest in time.
“ NIST Digital Library of Mathematical Functions
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.