Fetching the paper…
Reading the bibliography…
Predictive distributions quantify uncertainties ignored by point estimates.
Functional variational Bayesian neural networks
Sun, S., Zhang, G., Shi, J., and Grosse, R. (2019) · 1903
Earlier work this paper cites.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
Thompson, W. R. (1933) · 1933
Earlier work this paper cites.
Inverting modified matrices
Woodbury, M. A. (1950) · 1950
Earlier work this paper cites.
Nearest neighbor pattern classification
Cover, T. and Hart, P. (1967) · 1967
Earlier work this paper cites.
Bandit processes and dynamic allocation indices
Gittins, J. C. (1979) · 1979
Earlier work this paper cites.
Bayesian deep learning and a probabilistic perspective of generalization
Wilson, A. G. and Izmailov, P. (2020) · 2002
Earlier work this paper cites.
Gaussian processes in machine learning
Rasmussen, C. E. (2003) · 2003
Earlier work this paper cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020) · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L. (2009) · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Glorot, X. and Bengio, Y. (2010) · 2010
Earlier work this paper cites.
Knows what it knows: a framework for self-aware learning
Li, L., Littman, M. L., Walsh, T. J., and Strehl, A. L. (2011) · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
Pedregosa, F., Varoquaux, G., Gramfort, A., Michel, V., Thirion, B., Grisel, O., Blondel, M., Prettenhofer, P., Weiss, R., Dubourg, V., Vanderplas, J., Passos, A., Cournapeau, D., Brucher, M., Perrot, M., and Duchesnay, E. (2011) · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient Langevin dynamics
Welling, M. and Teh, Y. W. (2011) · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E. (2012) · 2012
Earlier work this paper cites.
Machine Learning: A Probabilistic Perspective
Murphy, K. P. (2012) · 2012
Earlier work this paper cites.
The no-u-turn sampler: adaptively setting path lengths in Hamiltonian Monte Carlo
Hoffman, M. D., Gelman, A., et al. (2014) · 2014
Cited alongside, same era.
Weight uncertainty in neural network
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D. (2015) · 2015
Cited alongside, same era.
Human-level Control through Deep Reinforcement Learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al. (2015) · 2015
Cited alongside, same era.
Bootstrapped Thompson sampling and deep exploration
Osband, I. and Van Roy, B. (2015) · 2015
Cited alongside, same era.
Variational inference with normalizing flows
Rezende, D. and Mohamed, S. (2015) · 2015
Cited alongside, same era.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Riquelme, C., Tucker, G., and Snoek, J. (2018) · 2018
Later among the works it cites.
A tutorial on Thompson sampling
Russo, D. J., Van Roy, B., Kazerouni, A., Osband, I., and Wen, Z. (2018) · 2018
Later among the works it cites.
Learning with kernels: Support vector machines, regularization, optimization, and beyond
Schölkopf, B. and Smola, A. J. (2018) · 2018
Later among the works it cites.
Hypermodels for exploration
Dwaracherla, V., Lu, X., Ibrahimi, M., Osband, I., Wen, Z., and Van Roy, B. (2020) · 2020
Later among the works it cites.
Bayesian deep ensembles via the neural tangent kernel
He, B., Lakshminarayanan, B., and Teh, Y. W. (2020) · 2020
Later among the works it cites.
Deep learning: a statistical viewpoint
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gal, Y. and Ghahramani, Z. (2016) · 2016
Cited alongside, same era.
Risk versus uncertainty in deep learning: Bayes, bootstrap and the dangers of dropout
Osband, I. (2016) · 2016
Cited alongside, same era.
Controlling bias in adaptive data analysis using information theory
Russo, D. and Zou, J. (2016) · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al. (2016) · 2016
Cited alongside, same era.
The elements of statistical learning: Data mining, inference, and prediction
Friedman, J. H. (2017) · 2017
Cited alongside, same era.
Variational Gaussian dropout is not Bayesian
Hron, J., Matthews, A. G. d. G., and Ghahramani, Z. (2017) · 2017
Cited alongside, same era.
Simple and scalable predictive uncertainty estimation using deep ensembles
Lakshminarayanan, B., Pritzel, A., and Blundell, C. (2017) · 2017
Cited alongside, same era.
Bartlett, P. L., Montanari, A., and Rakhlin, A. (2021) · 2021
Closest in time.
What are Bayesian neural network posteriors really like?
Izmailov, P., Vikram, S., Hoffman, M. D., and Wilson, A. G. (2021) · 2021
Closest in time.
Reinforcement learning, bit by bit
Lu, X., Van Roy, B., Dwaracherla, V., Ibrahimi, M., Osband, I., and Wen, Z. (2021) · 2021
Closest in time.
Uncertainty Baselines: Benchmarks for uncertainty & robustness in deep learning
Nado, Z., Band, N., Collier, M., Djolonga, J., Dusenberry, M., Farquhar, S., Filos, A., Havasi, M., Jenatton, R., Jerfel, G., Liu, J., Mariet, Z., Nixon, J., Padhy, S., Ren, J., Rudner, T., Wen, Y., Wenzel, F., Murphy, K., Sculley, D., Lakshminarayanan, B., Snoek, J., Gal, Y., and Tran, D. (2021) · 2021
Closest in time.
Osband, I., Wen, Z., Asghari, M., Dwaracherla, V., Ibrahimi, M., Lu, X., and Van Roy, B. (2021) · 2021
Closest in time.
Beyond marginal uncertainty: How accurately can Bayesian regression models estimate posterior predictive correlations?
Wang, C., Sun, S., and Grosse, R. (2021) · 2021
Closest in time.
Evaluating approximate inference in Bayesian deep learning
Wilson, A. G., Izmailov, P., Hoffman, M. D., Gal, Y., Li, Y., Pradier, M. F., Vikram, S., Foong, A., Lotfi, S., and Farquhar, S. (2021) · 2021
Closest in time.
Evaluating high-order predictive distributions in deep learning
Osband, I., Wen, Z., Asghari, S. M., Dwaracherla, V., Lu, X., and Van Roy, B. (2022) · 2022
Closest in time.
From predictions to decisions: The importance of joint predictive distributions
Wen, Z., Osband, I., Qin, C., Lu, X., Ibrahimi, M., Dwaracherla, V., Asghari, M., and Van Roy, B. (2022) · 2022
Closest in time.