Fetching the paper…
Reading the bibliography…
Deep neural networks (NNs) are powerful black box predictors that have recently achieved impressive performance on a wide spectrum of tasks.
Verification of forecasts expressed in terms of probability
G. W. Brier · 1950
Earlier work this paper cites.
The well-calibrated Bayesian
A. P. Dawid · 1982
Earlier work this paper cites.
The comparison and evaluation of forecasters
M. H. DeGroot and S. E. Fienberg · 1983
Earlier work this paper cites.
Adaptive mixtures of local experts
R. A. Jacobs, M. I. Jordan, S. J. Nowlan, and G. E. Hinton · 1991
Earlier work this paper cites.
Bayesian methods for adaptive models
D. J. MacKay · 1992
Earlier work this paper cites.
Stacked generalization
D. H. Wolpert · 1992
Earlier work this paper cites.
Mixture density networks
C. M. Bishop · 1994
Earlier work this paper cites.
Estimating the mean and variance of the target probability distribution
D. A. Nix and A. S. Weigend · 1994
Earlier work this paper cites.
Bagging predictors
L. Breiman · 1996
Earlier work this paper cites.
Bayesian Learning for Neural Networks
R. M. Neal · 1996
Earlier work this paper cites.
Ensemble methods in machine learning
T. G. Dietterich · 2000
Earlier work this paper cites.
Bayesian model averaging is not model combination
T. P. Minka · 2000
Earlier work this paper cites.
Random forests
L. Breiman · 2001
Earlier work this paper cites.
Comparing Bayes model averaging and stacking when model approximation error cannot be ignored
B. Clarke · 2003
Earlier work this paper cites.
Healing the relevance vector machine through augmentation
C. E. Rasmussen and J. Quinonero-Candela · 2005
Earlier work this paper cites.
Model compression
C. Bucila, R. Caruana, and A. Niculescu-Mizil · 2006
Earlier work this paper cites.
Extremely randomized trees
P. Geurts, D. Ernst, and L. Wehenkel · 2006
Earlier work this paper cites.
Evaluating predictive uncertainty challenge
J. Quinonero-Candela, C. E. Rasmussen, F. Sinz, O. Bousquet, and B. Schölkopf · 2006
Earlier work this paper cites.
Strictly proper scoring rules, prediction, and estimation
T. Gneiting and A. E. Raftery · 2007
Earlier work this paper cites.
Bayesian Theory , volume 405
J. M. Bernardo and A. F. Smith · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky · 2009
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. E. Hinton · 2010
Cited alongside, same era.
Practical variational inference for neural networks
A. Graves · 2011
Cited alongside, same era.
Bayesian learning via stochastic gradient Langevin dynamics
M. Welling and Y. W. Teh · 2011
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath, et al · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Efficient estimation of word representations in vector space
ImageNet Large Scale Visual Recognition Challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Later among the works it cites.
Predicting effects of noncoding variants with deep learning-based sequence model
J. Zhou and O. G. Troyanskaya · 2015
Later among the works it cites.
Concrete problems in AI safety
D. Amodei, C. Olah, J. Steinhardt, P. Christiano, J. Schulman, and D. Mané · 2016
Closest in time.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Y. Gal and Z. Ghahramani · 2016
Closest in time.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
A baseline for detecting misclassified and out-of-distribution examples in neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Mikolov, K. Chen, G. Corrado, and J. Dean · 2013
Cited alongside, same era.
S.-i. Maeda · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Intriguing properties of neural networks
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus · 2014
Cited alongside, same era.
Predicting the sequence specificities of DNA-and RNA-binding proteins by deep learning
B. Alipanahi, A. Delong, M. T. Weirauch, and B. J. Frey · 2015
Cited alongside, same era.
Weight uncertainty in neural networks
C. Blundell, J. Cornebise, K. Kavukcuoglu, and D. Wierstra · 2015
Cited alongside, same era.
Explaining and harnessing adversarial examples
I. J. Goodfellow, J. Shlens, and C. Szegedy · 2015
Cited alongside, same era.
D. Hendrycks and K. Gimpel · 2016
Closest in time.
Adversarial machine learning at scale
A. Kurakin, I. Goodfellow, and S. Bengio · 2016
Closest in time.
Decision trees and forests: a probabilistic perspective
B. Lakshminarayanan · 2016
Closest in time.
Stochastic multiple choice learning for training diverse deep ensembles
S. Lee, S. P. S. Prakash, M. Cogswell, V. Ranjan, D. Crandall, and D. Batra · 2016
Closest in time.
Structured and efficient variational deep learning with matrix Gaussian posteriors
C. Louizos and M. Welling · 2016
Closest in time.
Distributional smoothing by virtual adversarial examples
T. Miyato, S.-i. Maeda, M. Koyama, K. Nakae, and S. Ishii · 2016
Closest in time.
Deep exploration via bootstrapped DQN
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy · 2016
Closest in time.
Swapout: Learning an ensemble of deep architectures
S. Singh, D. Hoiem, and D. Forsyth · 2016
Closest in time.
Bayesian optimization with robust Bayesian neural networks
J. T. Springenberg, A. Klein, S. Falkner, and F. Hutter · 2016
Closest in time.
Rethinking the inception architecture for computer vision
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna · 2016
Closest in time.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra, et al · 2016
Closest in time.
Robustness to adversarial examples through an ensemble of specialists
M. Abbasi and C. Gagné · 2017
Closest in time.
On calibration of modern neural networks
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger · 2017
Closest in time.
Snapshot ensembles: Train 1, get M for free
G. Huang, Y. Li, G. Pleiss, Z. Liu, J. E. Hopcroft, and K. Q. Weinberger · 2017
Closest in time.
Ensemble adversarial training: Attacks and defenses
F. Tramèr, A. Kurakin, N. Papernot, D. Boneh, and P. McDaniel · 2017
Closest in time.