Fetching the paper…
Reading the bibliography…
Modern machine learning methods including deep learning have achieved great success in predictive accuracy for supervised learning tasks, but may still fall short in giving useful estimates of their predictive {\em uncertainty}.
Verification of forecasts expressed in terms of probability
Brier, G. W · 1950
Earlier work this paper cites.
The comparison and evaluation of forecasters
DeGroot, M. H. and Fienberg, S. E · 1983
Earlier work this paper cites.
Bayesian methods for adaptive models
MacKay, D. J · 1992
Earlier work this paper cites.
Novelty Detection and Neural Network Validation
Bishop, C. M · 1994
Earlier work this paper cites.
Newsweeder: Learning to filter netnews
Lang, K · 1995
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Density Networks
MacKay, D. J. and Gibbs, M. N · 1999
Earlier work this paper cites.
Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods
Platt, J. C · 1999
Earlier work this paper cites.
Evaluating predictive uncertainty challenge
Quinonero-Candela, J., Rasmussen, C. E., Sinz, F., Bousquet, O., and Schölkopf, B · 2006
Earlier work this paper cites.
Strictly proper scoring rules, prediction, and estimation
Gneiting, T. and Raftery, A. E · 2007
Earlier work this paper cites.
Reliability, sufficiency, and the decomposition of proper scores
Bröcker, J · 2009
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2009
Earlier work this paper cites.
Dataset shift in machine learning
Sugiyama, M., Lawrence, N. D., Schwaighofer, A., et al · 2009
Earlier work this paper cites.
NotMNIST dataset, 2011
Bulatov, Y · 2011
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A · 2011
Earlier work this paper cites.
Reading Digits in Natural Images with Unsupervised Feature Learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Bayesian Learning via Stochastic Gradient Langevin Dynamics
Welling, M. and Teh, Y. W · 2011
Earlier work this paper cites.
One billion word benchmark for measuring progress in statistical language modeling
Chelba, C., Mikolov, T., Schuster, M., Ge, Q., Brants, T., Koehn, P., and Robinson, T · 2013
Cited alongside, same era.
Adam: A Method for Stochastic Optimization
Kingma, D. and Ba, J · 2014
Cited alongside, same era.
Semi-supervised learning with deep generative models
Kingma, D. P., Mohamed, S., Rezende, D. J., and Welling, M · 2014
Cited alongside, same era.
Weight uncertainty in neural networks
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Cited alongside, same era.
Scalable variational gaussian process classification
Hensman, J., Matthews, A., and Ghahramani, Z · 2015
Cited alongside, same era.
Probabilistic Backpropagation for Scalable Learning of Bayesian Neural Networks
Hernández-Lobato, J. M. and Adams, R · 2015
A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks
Hendrycks, D. and Gimpel, K · 2017
Later among the works it cites.
What uncertainties do we need in Bayesian deep learning for computer vision?
Kendall, A. and Gal, Y · 2017
Later among the works it cites.
Self-normalizing neural networks
Klambauer, G., Unterthiner, T., Mayr, A., and Hochreiter, S · 2017
Later among the works it cites.
Simple and Scalable Predictive Uncertainty Estimation Using Deep Ensembles
Lakshminarayanan, B., Pritzel, A., and Blundell, C · 2017
Later among the works it cites.
Multiplicative Normalizing Flows for Variational Bayesian Neural Networks
Louizos, C. and Welling, M · 2017
Later among the works it cites.
An addendum to alchemy, 2017
Rahimi, A. and Recht, B · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Variational dropout and the local reparameterization trick
Kingma, D. P., Salimans, T., and Welling, M · 2015
Cited alongside, same era.
Obtaining Well Calibrated Probabilities Using Bayesian Binning
Naeini, M. P., Cooper, G. F., and Hauskrecht, M · 2015
Cited alongside, same era.
Training Very Deep Networks
Srivastava, R. K., Greff, K., and Schmidhuber, J · 2015
Cited alongside, same era.
Concrete problems in AI safety
Amodei, D., Olah, C., Steinhardt, J., Christiano, P., Schulman, J., and Mané, D · 2016
Cited alongside, same era.
End to end learning for self-driving cars
Bojarski, M., Testa, D. D., Dworakowski, D., Firner, B., Flepp, B., Goyal, P., Jackel, L. D., Monfort, M., Muller, U., Zhang, J., Zhang, X., Zhao, J., and Zieba, K · 2016
Cited alongside, same era.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Cited alongside, same era.
Uncertainty in the variational information bottleneck
Alemi, A. A., Fischer, I., and Dillon, J. V · 2018
Later among the works it cites.
Behrmann, J., Duvenaud, D., and Jacobsen, J.-H · 2018
Later among the works it cites.
A simple unified framework for detecting out-of-distribution samples and adversarial attacks
Lee, K., Lee, K., Lee, H., and Shin, J · 2018
Later among the works it cites.
Enhancing the Reliability of Out-of-Distribution Image Detection in Neural Networks
Liang, S., Li, Y., and Srikant, R · 2018
Later among the works it cites.
Troubling trends in machine learning scholarship
Lipton, Z. C. and Steinhardt, J · 2018
Later among the works it cites.
Deep Bayesian Bandits Showdown: An Empirical Comparison of Bayesian Deep Networks for Thompson Sampling
Riquelme, C., Tucker, G., and Snoek, J · 2018
Later among the works it cites.
Winner’s curse? On pace, progress, and empirical rigor
Sculley, D., Snoek, J., Wiltschko, A., and Rahimi, A · 2018
Later among the works it cites.
Does Your Model Know the Digit 6 Is Not a Cat? A Less Biased Evaluation of “Outlier” Detectors
Shafaei, A., Schmidt, M., and Little, J. J · 2018
Later among the works it cites.
Flipout: Efficient pseudo-independent weight perturbations on mini-batches
Wen, Y., Vicol, P., Ba, J., Tran, D., and Grosse, R · 2018
Later among the works it cites.
Benchmarking neural network robustness to common corruptions and perturbations
Hendrycks, D. and Dietterich, T · 2019
Closest in time.
Hybrid models with deep and invertible features
Nalisnick, E., Matsukawa, A., Teh, Y. W., Gorur, D., and Lakshminarayanan, B · 2019
Closest in time.
Likelihood ratios for out-of-distribution detection
Ren, J., Liu, P. J., Fertig, E., Snoek, J., Poplin, R., DePristo, M. A., Dillon, J. V., and Lakshminarayanan, B · 2019
Closest in time.
Deterministic Variational Inference for Robust Bayesian Neural Networks
Wu, A., Nowozin, S., Meeds, E., Turner, R. E., Hernandez-Lobato, J. M., and Gaunt, A. L · 2019
Closest in time.