Fetching the paper…
Reading the bibliography…
Markov chain Monte Carlo (MCMC) methods have not been broadly adopted in Bayesian neural networks (BNNs).
[author] Metropolis, NicholasN., Rosenbluth, Arianna WA. W., Rosenbluth, Marshall NM. N., Teller, Augusta HA. H. and Teller, EdwardE. (1953). Equation of state calculations by fast computing machines. The journal of Chemical Physics 21 1087–1092
1953
Earlier work this paper cites.
[author] Rosenblatt, FrankF. (1958). The perceptron: a probabilistic model for information storage and organization in the brain. Psychological review 65 386
1958
Earlier work this paper cites.
[author] Jeffreys, HaroldH. (1962). The theory of probability, 3rd ed. OUP Oxford
1962
Earlier work this paper cites.
[author] Jaynes, E. T.E. T. (1968). Prior probabilities. IEEE Transactions on Systems Science and Cybernetics 4 227–241
1968
Earlier work this paper cites.
[author] Hastings, W. K.W. K. (1970). Monte Carlo sampling methods using Markov chains and their applications. Biometrika 57 97–109
1970
Earlier work this paper cites.
[author] Bernardo, Jose M.J. M. (1979). Reference posterior distributions for Bayesian inference. Journal of the Royal Statistical Society. Series B (Methodological) 41 113–147
1979
Earlier work this paper cites.
[author] Minsky, Marvin LM. L. and Papert, Seymour AS. A. (1988). Perceptrons: expanded edition. MIT press
1988
Earlier work this paper cites.
Smith, J. W
1988
Earlier work this paper cites.
[author] Cybenko, GeorgeG. (1989). Approximation by superpositions of a sigmoidal function. Mathematics of control, signals and systems 2 303–314
1989
Earlier work this paper cites.
[author] Hecht-Nielsen, RobertR. (1990). On the algebraic structure of feedforward network weight spaces. In Advanced Neural Computers 129–135
1990
Earlier work this paper cites.
[author] Hornik, KurtK. (1991). Approximation capabilities of multilayer feedforward networks. Neural Networks 4 251–257
1991
Earlier work this paper cites.
[author] Gelman, AndrewA. and Rubin, Donald B.D. B. (1992). Inference from iterative simulation using multiple sequences. Statistical Science 7 457–472
1992
Earlier work this paper cites.
[author] Chen, A. M.A. M., Lu, H.H. and Hecht-Nielsen, R.R. (1993). On the geometry of feedforward neural network error surfaces. Neural Computation 5 910–927
1993
Earlier work this paper cites.
[author] MacKay, David JCD. J. (1995). Developments in probabilistic modelling with neural networks—ensemble learning. In Neural Networks: Artificial Intelligence and Industrial Applications 191–198
1995
Earlier work this paper cites.
[author] Williams, Peter M.P. M. (1995). Bayesian regularization and pruning using a Laplace prior. Neural Computation 7 117–143
1995
Earlier work this paper cites.
[author] Cowles, Mary KathrynM. K. and Carlin, Bradley P.B. P. (1996). Markov chain Monte Carlo convergence diagnostics: a comparative review. Journal of the American Statistical Association 91 883–904
1996
Earlier work this paper cites.
[author] Brooks, Stephen P.S. P. and Gelman, AndrewA. (1998). General methods for monitoring convergence of iterative simulations. Journal of Computational and Graphical Statistics 7 434–455
1998
Earlier work this paper cites.
[author] Kass, Robert E.R. E., Carlin, Bradley P.B. P., Gelman, AndrewA. and Neal, Radford M.R. M. (1998). Markov chain Monte Carlo in practice: a roundtable discussion. The American Statistician 52 93–100
1998
Earlier work this paper cites.
[author] Andrieu, C.C., de Freitas, J. F. G.J. F. G. and Doucet, A.A. (1999). Sequential Bayesian estimation and model selection applied to neural networks
1999
Earlier work this paper cites.
[author] de Freitas, NandoN. (1999). Bayesian methods for neural networks, PhD thesis, University of Cambridge
1999
Earlier work this paper cites.
Andrieu, C
2000
Earlier work this paper cites.
[author] Lee, H. K. H.H. K. H. (2000). Consistency of posterior distributions for neural networks. Neural Networks 13 629–642
2000
Earlier work this paper cites.
[author] Sargent, Daniel J.D. J., Hodges, James S.J. S. and Carlin, Bradley P.B. P. (2000). Structured Markov chain Monte Carlo. Journal of Computational and Graphical Statistics 9 217–234
2000
Earlier work this paper cites.
[author] Stephens, M.M. (2000). Dealing with label switching in mixture models. Journal of the Royal Statistical Society: Series B (Statistical Methodology) 62 795–809
2000
Earlier work this paper cites.
[author] Williams, Christopher K. I.C. K. I. (2000). An MCMC approach to hierarchical mixture modelling. In Advances in Neural Information Processing Systems 12 680–686
2000
Earlier work this paper cites.
[author] de Freitas, N.N., Andrieu, C.C., Højen-Sørensen, P.P., Niranjan, M.M. and Gee, A.A. (2001). Sequential Monte Carlo methods for neural networks In Sequential Monte Carlo Methods in Practice 359–379
2001
Earlier work this paper cites.
[author] Daniels, Michael J.M. J. and Kass, Robert E.R. E. (1998). A note on first-stage approximation in two-stage hierarchical models. Sankhyā: The Indian Journal of Statistics, Series B (1960-2002) 60 19–30
2002
Earlier work this paper cites.
[author] Lee, Herbert K. H.H. K. H. (2003). A noninformative prior for neural networks. Machine Learning 50 197–212
2003
Earlier work this paper cites.
[author] Gelman, AndrewA., Carlin, John BJ. B., Stern, Hal SH. S. and Rubin, Donald BD. B. (2004). Bayesian data analysis, 2nd ed. Chapman and Hall/CRC
2004
Earlier work this paper cites.
[author] Lee, Herbert KHH. K. (2004). Priors for neural networks. In Classification, Clustering, and Data Mining Applications (DavidD. Banks, Frederick R.F. R. McMorris, PhippsP. Arabie and WolfgangW. Gaul, eds.) 141–150
2004
Earlier work this paper cites.
[author] Titterington, D. M.D. M. (2004). Bayesian methods for neural networks and related models. Statistical Science 19 128–139
2004
Earlier work this paper cites.
Lee, H. K
2005
Earlier work this paper cites.
[author] Graf, SiegfriedS. and Luschgy, HaraldH. (2007). Foundations of quantization for probability distributions. Springer
2007
Cited alongside, same era.
[author] Lee, Herbert KHH. K. (2007). Default priors for neural network classification. Journal of Classification 24 53–70
2007
Cited alongside, same era.
[author] Friel, N.N. and Pettitt, A. N.A. N. (2008). Marginal likelihood estimation via power posteriors. Journal of the Royal Statistical Society: Series B (Statistical Methodology) 70 589–607
2008
Cited alongside, same era.
Jarrett, K
2009
Cited alongside, same era.
[author] Nair, VinodV. and Hinton, Geoffrey EG. E. (2009). 3D object recognition with deep belief nets. In Advances in Neural Information Processing Systems 22 1339–1347
2009
Cited alongside, same era.
Blier, L
2018
Later among the works it cites.
De Sa, C
2018
Later among the works it cites.
Freeman, I
2018
Later among the works it cites.
[author] Nalisnick, Eric ThomasE. T. (2018). On priors for Bayesian neural networks, PhD thesis, UC Irvine
2018
Later among the works it cites.
[author] Nemeth, ChristopherC. and Sherlock, ChrisC. (2018). Merging MCMC subposteriors through Gaussian-process approximations. Bayesian Analysis 13 507–530
2018
Later among the works it cites.
[author] Nwankpa, ChigozieC., Ijomah, WinifredW., Gachagan, AnthonyA. and Marshall, StephenS. (2018). Activation functions: comparison of trends in practice and research for deep learning. arXiv
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
[author] Neal, R. M.R. M. (2011). MCMC Using Hamiltonian dynamics In Handbook of Markov Chain Monte Carlo 5. CRC Press
2011
Cited alongside, same era.
Welling, M
2011
Cited alongside, same era.
[author] Gretton, ArthurA., Borgwardt, Karsten MK. M., Rasch, Malte JM. J., Schölkopf, BernhardB. and Smola, AlexanderA. (2012). A kernel two-sample test. Journal of Machine Learning Research 13 723–773
2012
Cited alongside, same era.
[author] Krizhevsky, AlexA., Sutskever, IlyaI. and Hinton, Geoffrey EG. E. (2012). ImageNet classification with deep convolutional neural networks. In Advances in Neural Information Processing Systems 25 1097–1105
2012
Cited alongside, same era.
[author] Badrinarayanan, VijayV., Mishra, BamdevB. and Cipolla, RobertoR. (2015). Symmetry-invariant optimization in deep networks. arXiv
2015
Cited alongside, same era.
[author] Giordano, Ryan JR. J., Broderick, TamaraT. and Jordan, Michael IM. I. (2015). Linear response methods for accurate covariance estimates from mean field variational Bayes. In Advances in Neural Information Processing Systems 28 1441–1449
2015
Cited alongside, same era.
[author] Gu, Shixiang (Shane)S. S., Ghahramani, ZoubinZ. and Turner, Richard ER. E. (2015). Neural adaptive sequential Monte Carlo. In Advances in Neural Information Processing Systems 28 2629–2637
2015
Cited alongside, same era.
[author] Ong, Victor M. H.V. M. H., Nott, David J.D. J. and Smith, Michael S.M. S. (2018). Gaussian variational approximation with a factor covariance structure. Journal of Computational and Graphical Statistics 27 465–478
2018
Later among the works it cites.
[author] Robert, Christian P.C. P., Elvira, VíctorV., Tawn, NickN. and Wu, ChangyeC. (2018). Accelerating MCMC algorithms. Wiley Interdisciplinary Reviews: Computational Statistics 10 e1435
2018
Later among the works it cites.
[author] Rudolf, DanielD. and Schweizer, NikolausN. (2018). Perturbation theory for Markov chains via Wasserstein distance. Bernoulli 24 2610–2639
2018
Later among the works it cites.
Seita, D
2018
Later among the works it cites.
Truong, T.-D
2018
Later among the works it cites.
[author] Vats, DootikaD. and Flegal, James MJ. M. (2018). Lugsail lag windows and their application to MCMC. arXiv
2018
Later among the works it cites.
[author] Vats, DootikaD. and Knudson, ChristinaC. (2018). Revisiting the Gelman-Rubin diagnostic. arXiv
2018
Later among the works it cites.
[author] Brea, JohanniJ., Simsek, BerfinB., Illing, BerndB. and Gerstner, WulframW. (2019). Weight-space symmetry in deep networks gives rise to permutation saddles, connected by equal-loss valleys across the loss landscape. arXiv
2019
Closest in time.
Cannon, A
2019
Closest in time.
Chen, W. Y
2019
Closest in time.
Esmaeili, B
2019
Closest in time.
[author] Hu, Shell XuS. X., Zagoruyko, SergeyS. and Komodakis, NikosN. (2019). Exploring weight symmetry in deep neural networks. Computer Vision and Image Understanding 187 102786
2019
Closest in time.
Huang, C.-W
2019
Closest in time.
Pearce, T
2019
Closest in time.
[author] Quiroz, MatiasM., Kohn, RobertR., Villani, MattiasM. and Tran, Minh-NgocM.-N. (2019). Speeding Up MCMC by efficient data subsampling. Journal of the American Statistical Association 114 831–843
2019
Closest in time.
Titsias, M. K
2019
Closest in time.
[author] Vats, DootikaD., Flegal, James MJ. M. and Jones, Galin LG. L. (2019). Multivariate output analysis for Markov chain Monte Carlo. Biometrika 106 321–337
2019
Closest in time.
[author] Vehtari, AkiA., Gelman, AndrewA., Simpson, DanielD., Carpenter, BobB. and Burkner, Paul-ChristianP.-C. (2019). Rank-normalization, folding, and localization: an improved R for assessing convergence of MCMC. arXiv
2019
Closest in time.
Vladimirova, M
2019
Closest in time.
Ashukha, A
2020
Closest in time.
Horst, A. M
2020
Closest in time.
Izmailov, P
2020
Closest in time.
[author] Johndrow, James E.J. E., Pillai, Natesh S.N. S. and Smith, AaronA. (2020). No free lunch for approximate MCMC. arXiv
2020
Closest in time.
[author] Sen, DeborsheeD., Papamarkou, TheodoreT. and Dunson, DavidD. (2020). Bayesian neural networks and dimensionality reduction. arXiv
2020
Closest in time.
[author] Wilson, Andrew GordonA. G. and Izmailov, PavelP. (2020). Bayesian deep learning and a probabilistic perspective of generalization. arXiv
2020
Closest in time.