Fetching the paper…
Reading the bibliography…
Neural networks have achieved remarkable performance across various problem domains, but their widespread applicability is hindered by inherent limitations such as overconfidence in predictions, lack of interpretability, and vulnerability to adversarial attacks.
Yang, G. (2019a) · 1902
Earlier work this paper cites.
Deep ensembles: A loss landscape perspective
Fort, S., Hu, H., and Lakshminarayanan, B. (2019) · 1912
Earlier work this paper cites.
On the validity of Bayesian neural networks for uncertainty estimation
Mitros, J. and Mac Namee, B. (2019) · 1912
Earlier work this paper cites.
A stochastic approximation method
Robbins, H. and Monro, S. (1951) · 1951
Earlier work this paper cites.
The perceptron: a probabilistic model for information storage and organization in the brain
Rosenblatt, F. (1958) · 1958
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R. (2014) · 1958
Earlier work this paper cites.
Algebra of probable inference
Cox, R. T. (1961) · 1961
Earlier work this paper cites.
The distribution of products of beta, gamma and Gaussian random variables
Springer, M. D. and Thompson, W. E. (1970) · 1970
Earlier work this paper cites.
The foundations of statistics
Savage, L. J. (1972) · 1972
Earlier work this paper cites.
The comparison and evaluation of forecasters
DeGroot, M. H. and Fienberg, S. E. (1983) · 1983
Earlier work this paper cites.
Learning representations by back-propagating errors
Rumelhart, D. E., Hinton, G. E., and Williams, R. J. (1986) · 1986
Earlier work this paper cites.
Learnability and the vapnik-chervonenkis dimension
Blumer, A., Ehrenfeucht, A., Haussler, D., and Warmuth, M. K. (1989) · 1989
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
Cybenko, G. (1989) · 1989
Earlier work this paper cites.
On the approximate realization of continuous mappings by neural networks
Funahashi, K.-I. (1989) · 1989
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
Hornik, K., Stinchcombe, M., and White, H. (1989) · 1989
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
LeCun, Y., Boser, B., Denker, J. S., Henderson, D., Howard, R. E., Hubbard, W., and Jackel, L. D. (1989) · 1989
Earlier work this paper cites.
Inference from iterative simulation using multiple sequences
Gelman, A. and Rubin, D. B. (1992) · 1992
Earlier work this paper cites.
A practical Bayesian framework for backpropagation networks
MacKay, D. J. (1992) · 1992
Earlier work this paper cites.
Bayesian training of backpropagation networks by the hybrid Monte Carlo method
Neal, R. M. (1992) · 1992
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Hinton, G. E. and Van Camp, D. (1993) · 1993
Earlier work this paper cites.
Bayesian learning via stochastic dynamics
Neal, R. M. (1993) · 1993
Earlier work this paper cites.
Approximation and estimation bounds for artificial neural networks
Barron, A. R. (1994) · 1994
Earlier work this paper cites.
Bayes factors
Kass, R. E. and Raftery, A. E. (1995) · 1995
Earlier work this paper cites.
Bayesian learning for neural networks
Neal, R. M. (1996) · 1996
Earlier work this paper cites.
Regression shrinkage and selection via the Lasso
Tibshirani, R. (1996) · 1996
Earlier work this paper cites.
The lack of a priori distinctions between learning algorithms
Wolpert, D. H. (1996) · 1996
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Ensemble learning in Bayesian neural networks
Barber, D. and Bishop, C. M. (1998) · 1998
Earlier work this paper cites.
Ridgelets: theory and applications
Candès, E. J. (1998) · 1998
Earlier work this paper cites.
An introduction to variational methods for graphical models
Jordan, M. I., Ghahramani, Z., Jaakkola, T. S., and Saul, L. K. (1999) · 1999
Earlier work this paper cites.
Some PAC-Bayesian theorems
McAllester, D. A. (1999) · 1999
Earlier work this paper cites.
Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods
Platt, J. (1999) · 1999
Earlier work this paper cites.
An overview of statistical learning theory
Vapnik, V. N. (1999) · 1999
Earlier work this paper cites.
The case for Bayesian deep learning
Wilson, A. G. (2020) · 2001
Earlier work this paper cites.
Obtaining calibrated probability estimates from decision trees and naive Bayesian classifiers
Zadrozny, B. and Elkan, C. (2001) · 2001
Earlier work this paper cites.
Interpreting a penalty as the influence of a Bayesian prior
Wolinski, P., Charpiat, G., and Ollivier, Y. (2020) · 2002
Earlier work this paper cites.
Transforming classifier scores into accurate multiclass probability estimates
Zadrozny, B. and Elkan, C. (2002) · 2002
Earlier work this paper cites.
Information theory, inference and learning algorithms
MacKay, D. J. (2003) · 2003
Earlier work this paper cites.
Predicting the outputs of finite networks trained with noisy gradients
Naveh, G., Ben-David, O., Sompolinsky, H., and Ringel, Z. (2020) · 2004
Earlier work this paper cites.
Monte Carlo statistical methods
Robert, C. P. and Casella, G. (2004) · 2004
Earlier work this paper cites.
A comparison of tight generalization error bounds
Kääriäinen, M. and Langford, J. (2005) · 2005
Earlier work this paper cites.
Tutorial on practical prediction theory for classification
Langford, J. (2005) · 2005
Earlier work this paper cites.
Pattern recognition and machine learning
Bishop, C. M. and Nasrabadi, N. M. (2006) · 2006
Earlier work this paper cites.
Gaussian processes for machine learning
Rasmussen, C. E. and Williams, C. K. (2006) · 2006
Earlier work this paper cites.
Kernel methods for deep learning
Cho, Y. and Saul, L. K. (2009) · 2009
Earlier work this paper cites.
Aleatory or epistemic? Does it matter?
Der Kiureghian, A. and Ditlevsen, O. (2009) · 2009
Earlier work this paper cites.
Why does unsupervised pre-training help deep learning?
Erhan, D., Courville, A., Bengio, Y., and Vincent, P. (2010) · 2010
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X. and Bengio, Y. (2010) · 2010
Earlier work this paper cites.
Gelman, A., Vehtari, A., Simpson, D., Margossian, C. C., Carpenter, B., Yao, Y., Kennedy, L., Gabry, J., Bürkner, P.-C., and Modrák, M. (2020) · 2011
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A. (2011) · 2011
Earlier work this paper cites.
All you need is a good functional prior for Bayesian deep learning
Tran, B.-H., Rossi, S., Milios, D., and Filippone, M. (2020) · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient Langevin dynamics
Welling, M. and Teh, Y. W. (2011) · 2011
Earlier work this paper cites.
Bayesian posterior sampling via stochastic gradient fisher scoring
Ahn, S., Korattikara, A., and Welling, M. (2012) · 2012
Earlier work this paper cites.
The safe Bayesian
Grünwald, P. (2012) · 2012
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E. (2012) · 2012
Earlier work this paper cites.
Machine learning: a probabilistic perspective
Murphy, K. P. (2012) · 2012
Earlier work this paper cites.
Deep Gaussian processes
Damianou, A. and Lawrence, N. D. (2013) · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Graves, A., Mohamed, A., and Hinton, G. E. (2013) · 2013
Earlier work this paper cites.
Dropout training as adaptive regularization
Wager, S., Wang, S., and Liang, P. S. (2013) · 2013
Earlier work this paper cites.
Stochastic gradient hamiltonian monte carlo
Chen, T., Fox, E., and Guestrin, C. (2014) · 2014
Earlier work this paper cites.
Avoiding pathologies in very deep networks
Duvenaud, D., Rippel, O., Adams, R., and Ghahramani, Z. (2014) · 2014
Earlier work this paper cites.
Auto-encoding variational Bayes
Kingma, D. P. and Welling, M. (2014) · 2014
Earlier work this paper cites.
Automatic construction and natural-language description of nonparametric regression models
Lloyd, J., Duvenaud, D., Grosse, R., Tenenbaum, J., and Ghahramani, Z. (2014) · 2014
Earlier work this paper cites.
Asymptotically exact, embarrassingly parallel mcmc
Neiswanger, W., Wang, C., and Xing, E. P. (2014) · 2014
Earlier work this paper cites.
Weight uncertainty in neural networks
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D. (2015) · 2015
Earlier work this paper cites.
Probabilistic machine learning and artificial intelligence
Ghahramani, Z. (2015) · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on ImageNet classification
He, K., Zhang, X., Ren, S., and Sun, J. (2015) · 2015
Earlier work this paper cites.
Probabilistic backpropagation for scalable learning of Bayesian neural networks
Hernández-Lobato, J. M. and Adams, R. (2015) · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C. (2015) · 2015
Earlier work this paper cites.
Variational dropout and the local reparameterization trick
Kingma, D. P., Salimans, T., and Welling, M. (2015) · 2015
Earlier work this paper cites.
Bayesian dark knowledge
Korattikara Balan, A., Rathod, V., Murphy, K. P., and Welling, M. (2015) · 2015
Earlier work this paper cites.
A complete recipe for stochastic gradient mcmc
Ma, Y.-A., Chen, T., and Fox, E. (2015) · 2015
Earlier work this paper cites.
Optimizing neural networks with kronecker-factored approximate curvature
Martens, J. and Grosse, R. (2015) · 2015
Earlier work this paper cites.
Obtaining well calibrated probabilities using Bayesian binning
Naeini, M. P., Cooper, G., and Hauskrecht, M. (2015) · 2015
Earlier work this paper cites.
Frequentist coverage of adaptive nonparametric Bayesian credible sets
Szabó, B., van der Vaart, A. W., and van Zanten, J. (2015) · 2015
Earlier work this paper cites.
Privacy for free: Posterior sampling and stochastic gradient monte carlo
Wang, Y.-X., Fienberg, S., and Smola, A. (2015) · 2015
Earlier work this paper cites.
Robustness of classifiers: from adversarial to random noise
Fawzi, A., Moosavi-Dezfooli, S.-M., and Frossard, P. (2016) · 2016
Earlier work this paper cites.
Uncertainty in deep learning
Gal, Y. (2016) · 2016
Earlier work this paper cites.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z. (2016) · 2016
Earlier work this paper cites.
PAC-Bayesian theory meets Bayesian inference
Germain, P., Bach, F., Lacoste, A., and Lacoste-Julien, S. (2016) · 2016
Earlier work this paper cites.
Preconditioned stochastic gradient langevin dynamics for deep neural networks
Li, C., Chen, C., Carlson, D., and Carin, L. (2016) · 2016
Earlier work this paper cites.
Structured and efficient variational deep learning with matrix Gaussian posteriors
Louizos, C. and Welling, M. (2016) · 2016
Cited alongside, same era.
Deepfool: a simple and accurate method to fool deep neural networks
Moosavi-Dezfooli, S.-M., Fawzi, A., and Frossard, P. (2016) · 2016
Cited alongside, same era.
Adding gradient noise improves learning for very deep networks
Neelakantan, A., Vilnis, L., Le, Q. V., Sutskever, I., Kaiser, L., Kurach, K., and Martens, J. (2016) · 2016
Cited alongside, same era.
Exponential expressivity in deep neural networks through transient chaos
Poole, B., Lahiri, S., Raghu, M., Sohl-Dickstein, J., and Ganguli, S. (2016) · 2016
Cited alongside, same era.
Probabilistic programming in python using PyMC3
Salvatier, J., Wiecki, T. V., and Fonnesbeck, C. (2016) · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
Practical deep learning with Bayesian principles
Osawa, K., Swaroop, S., Khan, M. E. E., Jain, A., Eschenhagen, R., Turner, R. E., and Yokota, R. (2019) · 2019
Later among the works it cites.
Can you trust your model’s uncertainty? Evaluating predictive uncertainty under dataset shift
Ovadia, Y., Fertig, E., Ren, J., Nado, Z., Sculley, D., Nowozin, S., Dillon, J. V., Lakshminarayanan, B., and Snoek, J. (2019) · 2019
Later among the works it cites.
Evaluating model calibration in classification
Vaicenavicius, J., Widmann, D., Andersson, C., Lindsten, F., Roll, J., and Schön, T. (2019) · 2019
Later among the works it cites.
Understanding priors in Bayesian neural networks at the unit level
Vladimirova, M., Verbeek, J., Mesejo, P., and Arbel, J. (2019) · 2019
Later among the works it cites.
Cyclical Stochastic Gradient MCMC for Bayesian Deep Learning
Zhang, R., Li, C., Zhang, J., Chen, C., and Wilson, A. G. (2019) · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., and Lanctot, M. (2016) · 2016
Cited alongside, same era.
Robust large margin deep neural networks
Sokolić, J., Giryes, R., Sapiro, G., and Rodrigues, M. R. (2016) · 2016
Cited alongside, same era.
Benefits of depth in neural networks
Telgarsky, M. (2016) · 2016
Cited alongside, same era.
A survey of transfer learning
Weiss, K., Khoshgoftaar, T. M., and Wang, D. (2016) · 2016
Cited alongside, same era.
Deep kernel learning
Wilson, A. G., Hu, Z., Salakhutdinov, R., and Xing, E. P. (2016) · 2016
Cited alongside, same era.
Accelerating neural architecture search using performance prediction
Baker, B., Gupta, O., Raskar, R., and Naik, N. (2017) · 2017
Cited alongside, same era.
Spectrally-normalized margin bounds for neural networks
Bartlett, P. L., Foster, D. J., and Telgarsky, M. J. (2017) · 2017
Cited alongside, same era.
Aitchison, L. (2020) · 2020
Later among the works it cites.
Pitfalls of in-domain uncertainty estimation and ensembling in deep learning
Ashukha, A., Lyzhov, A., Molchanov, D., and Vetrov, D. (2020) · 2020
Later among the works it cites.
Efficient and scalable Bayesian neural nets with rank-1 factors
Dusenberry, M., Jerfel, G., Wen, Y., Ma, Y., Snoek, J., Heller, K., Lakshminarayanan, B., and Tran, D. (2020) · 2020
Later among the works it cites.
Asymptotics of wide networks from Feynman diagrams
Dyer, E. and Gur-Ari, G. (2020) · 2020
Later among the works it cites.
Stable behaviour of infinitely wide deep neural networks
Favaro, S., Fortini, S., and Stefano, P. (2020) · 2020
Later among the works it cites.
Bayesian neural networks: An introduction and survey
Goan, E. and Fookes, C. (2020) · 2020
Later among the works it cites.
Dynamics of deep neural networks and neural tangent hierarchy
Huang, J. and Yau, H.-T. (2020) · 2020
Later among the works it cites.
Being Bayesian, even just a bit, fixes overconfidence in ReLU networks
Kristiadi, A., Hein, M., and Hennig, P. (2020) · 2020
Later among the works it cites.
Finite versus infinite neural networks: an empirical study
Lee, J., Schoenholz, S., Pennington, J., Adlam, B., Xiao, L., Novak, R., and Sohl-Dickstein, J. (2020) · 2020
Later among the works it cites.
Dying ReLU and initialization: Theory and numerical examples
Lu, L., Shin, Y., Su, Y., and Karniadakis, G. E. (2020) · 2020
Later among the works it cites.
A Bayesian perspective on training speed and model selection
Lyle, C., Schut, L., Ru, R., Gal, Y., and van der Wilk, M. (2020) · 2020
Later among the works it cites.
Proving the lottery ticket hypothesis: Pruning is all you need
Malach, E., Yehudai, G., Shalev-Schwartz, S., and Shamir, O. (2020) · 2020
Later among the works it cites.
Bayesian deep convolutional networks with many channels are Gaussian processes
Novak, R., Xiao, L., Bahri, Y., Lee, J., Yang, G., Hron, J., Abolafia, D. A., Pennington, J., and Sohl-dickstein, J. (2020) · 2020
Later among the works it cites.
Uncertainty in neural networks: Approximately Bayesian ensembling
Pearce, T., Leibfried, F., and Brintrup, A. (2020) · 2020
Later among the works it cites.
Nonparametric regression using deep neural networks with ReLU activation function
Schmidt-Hieber, J. (2020) · 2020
Later among the works it cites.
Sub-Weibull distributions: Generalizing sub-Gaussian and sub-exponential properties to heavier tailed distributions
Vladimirova, M., Girard, S., Nguyen, H., and Arbel, J. (2020) · 2020
Later among the works it cites.
How good is the Bayes posterior in deep neural networks really?
Wenzel, F., Roth, K., Veeling, B., Swiatkowski, J., Tran, L., Mandt, S., Snoek, J., Salimans, T., Jenatton, R., and Nowozin, S. (2020) · 2020
Later among the works it cites.
Bayesian deep learning and a probabilistic perspective of generalization
Wilson, A. G. and Izmailov, P. (2020) · 2020
Later among the works it cites.
Non-Gaussian processes and neural networks at finite widths
Yaida, S. (2020) · 2020
Later among the works it cites.
A review of uncertainty quantification in deep learning: Techniques, applications and challenges
Abdar, M., Pourpanah, F., Hussain, S., Rezazadegan, D., Liu, L., Ghavamzadeh, M., Fieguth, P., Cao, X., Khosravi, A., and Acharya, U. R. (2021) · 2021
Later among the works it cites.
Zero-cost proxies for lightweight NAS
Abdelfattah, M. S., Mehrotra, A., Dudziak, Ł., and Lane, N. D. (2021) · 2021
Later among the works it cites.
A statistical theory of cold posteriors in deep neural networks
Aitchison, L. (2021) · 2021
Later among the works it cites.
Benchmarking bayesian deep learning on diabetic retinopathy detection tasks
Band, N., Rudner, T. G., Feng, Q., Filos, A., Nado, Z., Dusenberry, M. W., Jerfel, G., Tran, D., and Gal, Y. (2021) · 2021
Later among the works it cites.
The effect of prior Lipschitz continuity on the adversarial robustness of Bayesian neural networks
Blaas, A. and Roberts, S. J. (2021) · 2021
Later among the works it cites.
Informative Bayesian Neural Network Priors for Weak Signals
Cui, T., Havulinna, A., Marttinen, P., and Kaski, S. (2021) · 2021
Later among the works it cites.
Repulsive deep ensembles are Bayesian
D’Angelo, F. and Fortuin, V. (2021) · 2021
Later among the works it cites.
ConViT: Improving vision transformers with soft convolutional inductive biases
d’Ascoli, S., Touvron, H., Leavitt, M., Morcos, A., Biroli, G., and Sagun, L. (2021) · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., and Gelly, S. (2021) · 2021
Later among the works it cites.
On the role of data in pac-bayes bounds
Dziugaite, G. K., Hsu, K., Gharbieh, W., Arpino, G., and Roy, D. (2021) · 2021
Later among the works it cites.
Folgoc, L. L., Baltatzis, V., Desai, S., Devaraj, A., Ellis, S., Manzanera, O. E. M., Nair, A., Qiu, H., Schnabel, J., and Glocker, B. (2021) · 2021
Later among the works it cites.
Sharpness-aware minimization for efficiently improving generalization
Foret, P., Kleiner, A., Mobahi, H., and Neyshabur, B. (2021) · 2021
Later among the works it cites.
On the Choice of Priors in Bayesian Deep Learning
Fortuin, V. (2021) · 2021
Later among the works it cites.
Bayesian neural network priors revisited
Fortuin, V., Garriga-Alonso, A., Wenzel, F., Rätsch, G., Turner, R., van der Wilk, M., and Aitchison, L. (2021) · 2021
Later among the works it cites.
A survey of uncertainty in deep neural networks
Gawlikowski, J., Tassi, C. R. N., Ali, M., Lee, J., Humt, M., Feng, J., Kruspe, A., Triebel, R., Jung, P., and Roscher, R. (2021) · 2021
Later among the works it cites.
The heavy-tail phenomenon in SGD
Gurbuzbalaban, M., Şimşekli, U., and Zhu, L. (2021) · 2021
Later among the works it cites.
Can we trust Bayesian uncertainty quantification from Gaussian process priors with squared exponential covariance kernel?
Hadji, A. and Szabó, B. (2021) · 2021
Later among the works it cites.
Khan, M. E. and Rue, H. (2021) · 2021
Later among the works it cites.
Knowledge-adaptation priors
Khan, M. E. and Swaroop, S. (2021) · 2021
Later among the works it cites.
On the rate of convergence of fully connected deep neural network regression estimates
Kohler, M. and Langer, S. (2021) · 2021
Later among the works it cites.
Shape-texture debiased neural network training
Li, Y., Yu, Q., Tan, M., Mei, J., Tang, P., Shen, W., Yuille, A., and Xie, C. (2021) · 2021
Later among the works it cites.
Temporal fusion transformers for interpretable multi-horizon time series forecasting
Lim, B., Arık, S. Ö., Loeff, N., and Pfister, T. (2021) · 2021
Later among the works it cites.
The Ridgelet prior: A covariance function approach to prior specification for Bayesian neural networks
Matsubara, T., Oates, C. J., and Briol, F.-X. (2021) · 2021
Later among the works it cites.
Neural architecture search without training
Mellor, J., Turner, J., Storkey, A., and Crowley, E. J. (2021) · 2021
Later among the works it cites.
Revisiting the calibration of modern neural networks
Minderer, M., Djolonga, J., Romijnders, R., Hubis, F., Zhai, X., Houlsby, N., Tran, D., and Lucic, M. (2021) · 2021
Later among the works it cites.
PRIME: A few primitives can boost robustness to common corruptions
Modas, A., Rade, R., Ortiz-Jiménez, G., Moosavi-Dezfooli, S.-M., and Frossard, P. (2021) · 2021
Later among the works it cites.
Exploring corruption robustness: Inductive biases in vision transformers and MLP-mixers
Morrison, K., Gilby, B., Lipchak, C., Mattioli, A., and Kovashka, A. (2021) · 2021
Later among the works it cites.
Data augmentation in Bayesian neural networks and the cold posterior effect
Nabarro, S., Ganev, S., Garriga-Alonso, A., Fortuin, V., van der Wilk, M., and Aitchison, L. (2021) · 2021
Later among the works it cites.
Deep double descent: Where bigger models and more data hurt
Nakkiran, P., Kaplun, G., Bansal, Y., Yang, T., Barak, B., and Sutskever, I. (2021) · 2021
Later among the works it cites.
Predictive complexity priors
Nalisnick, E., Gordon, J., and Hernández-Lobato, J. M. (2021) · 2021
Later among the works it cites.
Precise characterization of the prior predictive distribution of deep ReLU networks
Noci, L., Bachmann, G., Roth, K., Nowozin, S., and Hofmann, T. (2021) · 2021
Later among the works it cites.
PACOH: Bayes-optimal meta-learning with PAC-guarantees
Rothfuss, J., Fortuin, V., Josifoski, M., and Krause, A. (2021) · 2021
Later among the works it cites.
Speedy performance estimation for neural architecture search
Ru, R., Lyle, C., Schut, L., Fil, M., van der Wilk, M., and Gal, Y. (2021) · 2021
Later among the works it cites.
Deep learning: a comprehensive overview on techniques, taxonomy, applications and research directions
Sarker, I. H. (2021) · 2021
Later among the works it cites.
Rank-normalization, folding, and localization: An improved R ^ \widehat{R} for assessing convergence of MCMC (with discussion)
Vehtari, A., Gelman, A., Simpson, D., Carpenter, B., and Bürkner, P.-C. (2021) · 2021
Later among the works it cites.
Asymptotics of representation learning in finite Bayesian neural networks
Zavatone-Veth, J. A., Canatar, A., and Pehlevan, C. (2021) · 2021
Later among the works it cites.
Exact priors of finite neural networks
Zavatone-Veth, J. A. and Pehlevan, C. (2021) · 2021
Later among the works it cites.
Understanding deep learning (still) requires rethinking generalization
Zhang, C., Bengio, S., Hardt, M., Recht, B., and Vinyals, O. (2021) · 2021
Later among the works it cites.
Adapting the linearised Laplace model evidence for modern deep learning
Antorán, J., Janz, D., Allingham, J. U., Daxberger, E., Barbano, R. R., Nalisnick, E., and Hernández-Lobato, J. M. (2022) · 2022
Later among the works it cites.
How Tempering Fixes Data Augmentation in Bayesian Neural Networks
Bachmann, G., Noci, L., and Hofmann, T. (2022) · 2022
Later among the works it cites.
Priors in Bayesian deep learning: A review
Fortuin, V. (2022) · 2022
Later among the works it cites.
Uncertainty quantification for nonparametric regression using empirical bayesian neural networks
Franssen, S. and Szabó, B. (2022) · 2022
Later among the works it cites.
Better uncertainty calibration via proper scores for classification and beyond
Gruber, S. and Buettner, F. (2022) · 2022
Later among the works it cites.
On the infinite-depth limit of finite-width neural networks
Hayou, S. (2022) · 2022
Later among the works it cites.
Invariance learning in deep neural networks with differentiable laplace approximations
Immer, A., van der Ouderaa, T. F., Fortuin, V., Rätsch, G., and van der Wilk, M. (2022) · 2022
Later among the works it cites.
Hands-on bayesian neural networks—a tutorial for deep learning users
Jospin, L. V., Laga, H., Boussaid, F., Buntine, W., and Bennamoun, M. (2022) · 2022
Later among the works it cites.
On Uncertainty, Tempering, and Data Augmentation in Bayesian Classification
Kapoor, S., Maddox, W. J., Izmailov, P., and Wilson, A. G. (2022) · 2022
Later among the works it cites.
Bayesian model selection, the marginal likelihood, and generalization
Lotfi, S., Izmailov, P., Benton, G., Goldblum, M., and Wilson, A. G. (2022) · 2022
Later among the works it cites.
Sam as an optimal relaxation of bayes
Möllenhoff, T. and Khan, M. E. (2022) · 2022
Later among the works it cites.
The neural testbed: Evaluating joint predictions
Osband, I., Wen, Z., Asghari, S. M., Dwaracherla, V., Lu, X., Ibrahimi, M., Lawson, D., Hao, B., O’Donoghue, B., and Van Roy, B. (2022) · 2022
Later among the works it cites.
Cold Posteriors through PAC-Bayes
Pitas, K. and Arbel, J. (2022) · 2022
Later among the works it cites.
PAC-Bayesian meta-learning: From theory to practice
Rothfuss, J., Josifoski, M., Fortuin, V., and Krause, A. (2022) · 2022
Later among the works it cites.
Efficient bayes inference in neural networks through adaptive importance sampling
Huang, Y., Chouzenoux, E., Elvira, V., and Pesquet, J.-C. (2023) · 2023
Closest in time.
Prior knowledge elicitation: The past, present, and future
Mikkola, P., Martin, O. A., Chandramouli, S., Hartmann, M., Pla, O. A., Thomas, O., Pesonen, H., Corander, J., Vehtari, A., Kaski, S., Bürkner, P.-C., and Klami, A. (2023) · 2023
Closest in time.
On the use of a local R ^ \hat{R} to improve MCMC convergence diagnostic
Moins, T., Arbel, J., Dutfoy, A., and Girard, S. (2023) · 2023
Closest in time.
Gaussian Pre-Activations in Neural Networks: Myth or Reality?
Wolinski, P. and Arbel, J. (2023) · 2023
Closest in time.