Fetching the paper…
Reading the bibliography…
Virtually any model we use in machine learning to make predictions does not perfectly represent reality.
Science and statistics
G. E. Box · 1976
Earlier work this paper cites.
Principles and techniques of applied mathematics
B. Friedman · 1990
Earlier work this paper cites.
An overview of robust Bayesian analysis
J. O. Berger, E. Moreno, L. R. Pericchi, M. J. Bayarri, J. M. Bernardo, J. A. Cano, J. De la Horra, J. Martín, D. Ríos-Insúa, B. Betrò, et al · 1994
Earlier work this paper cites.
A short introduction to boosting
Y. Freund, R. Schapire, and N. Abe · 1999
Earlier work this paper cites.
PAC-Bayesian model averaging
D. A. McAllester · 1999
Earlier work this paper cites.
Probabilistic principal component analysis
M. E. Tipping and C. M. Bishop · 1999
Earlier work this paper cites.
Ensemble methods in machine learning
T. G. Dietterich · 2000
Earlier work this paper cites.
Random forests
L. Breiman · 2001
Earlier work this paper cites.
Bounds for averaging classifiers
J. Langford and M. Seeger · 2001
Earlier work this paper cites.
Stochastic gradient boosting
J. H. Friedman · 2002
Earlier work this paper cites.
Latent Dirichlet allocation
D. M. Blei, A. Y. Ng, and M. I. Jordan · 2003
Earlier work this paper cites.
Measures of diversity in classifier ensembles and their relationship with the ensemble accuracy
L. I. Kuncheva and C. J. Whitaker · 2003
Earlier work this paper cites.
Bayesian Gaussian process models: PAC-Bayesian generalisation error bounds and sparse approximations
M. Seeger · 2003
Earlier work this paper cites.
Pattern recognition and machine learning
C. M. Bishop · 2006
Earlier work this paper cites.
Information-theoretic upper and lower bounds for statistical estimation
T. Zhang · 2006
Earlier work this paper cites.
From epsilon-entropy to kl-entropy: Analysis of minimum information complexity density estimation
T. Zhang et al · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky · 2009
Earlier work this paper cites.
The variance drain and Jensen’s inequality
R. A. Becker · 2012
Earlier work this paper cites.
The safe Bayesian: learning the learning rate via the mixability gap
P. Grünwald · 2012
Earlier work this paper cites.
Robust Bayesian Analysis
D. R. Insua and F. Ruggeri · 2012
Earlier work this paper cites.
The Bernstein-von-Mises theorem under misspecification
B. J. K. Kleijn, A. W. Van der Vaart, et al · 2012
Earlier work this paper cites.
Stochastic variational inference
M. D. Hoffman, D. M. Blei, C. Wang, and J. Paisley · 2013
Cited alongside, same era.
Bayesian inference with misspecified models
S. G. Walker · 2013
Cited alongside, same era.
Probabilistic machine learning and artificial intelligence
Z. Ghahramani · 2015
Cited alongside, same era.
Variational inference with normalizing flows
D. J. Rezende and S. Mohamed · 2015
Cited alongside, same era.
On the properties of variational approximations of Gibbs posteriors
P. Alquier, J. Ridgway, and N. Chopin · 2016
Cited alongside, same era.
A general framework for updating belief distributions
P. G. Bissiri, C. C. Holmes, and S. G. Walker · 2016
Cited alongside, same era.
Advances in variational inference
C. Zhang, J. Butepage, H. Kjellstrom, and S. Mandt · 2018
Later among the works it cites.
Mmd-Bayes: Robust Bayesian estimation via maximum mean discrepancy
B.-E. Chérief-Abdellatif and P. Alquier · 2019
Closest in time.
Deep ensembles: A loss landscape perspective
S. Fort, H. Hu, and B. Lakshminarayanan · 2019
Closest in time.
Parametric Gaussian process regressors
M. Jankowiak, G. Pleiss, and J. R. Gardner · 2019
Closest in time.
Robust deep Gaussian processes
J. Knoblauch · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
PAC-Bayesian theory meets Bayesian inference
P. Germain, F. Bach, A. Lacoste, and S. Lacoste-Julien · 2016
Cited alongside, same era.
Stein variational gradient descent: A general purpose Bayesian inference algorithm
Q. Liu and D. Wang · 2016
Cited alongside, same era.
Variational inference: A review for statisticians
D. M. Blei, A. Kucukelbir, and J. D. McAuliffe · 2017
Cited alongside, same era.
G. K. Dziugaite and D. M. Roy · 2017
Cited alongside, same era.
Inconsistency of Bayesian inference for misspecified linear models, and a proposal for repairing it
P. Grünwald and T. Van Ommen · 2017
Cited alongside, same era.
Simple and scalable predictive uncertainty estimation using deep ensembles
B. Lakshminarayanan, A. Pritzel, and C. Blundell · 2017
Cited alongside, same era.
J. Knoblauch, J. Jewson, and T. Damoulas · 2019
Closest in time.
Dichotomize and generalize: PAC-Bayesian binary activated deep neural networks
G. Letarte, P. Germain, B. Guedj, and F. Laviolette · 2019
Closest in time.
Sharpening Jensen’s inequality
J. Liao and A. Berg · 2019
Closest in time.
Accurate uncertainty estimation and decomposition in ensemble learning
J. Liu, J. Paisley, M.-A. Kioumourtzoglou, and B. Coull · 2019
Closest in time.
Probabilistic models with deep neural networks
A. R. Masegosa, R. Cabañas, H. Langseth, T. D. Nielsen, and A. Salmerón · 2019
Closest in time.
Monte Carlo gradient estimation in machine learning
S. Mohamed, M. Rosca, M. Figurnov, and A. Mnih · 2019
Closest in time.
Practical deep learning with Bayesian principles
K. Osawa, S. Swaroop, M. E. E. Khan, A. Jain, R. Eschenhagen, R. E. Turner, and R. Yokota · 2019
Closest in time.
Improved PAC-Bayesian bounds for linear regression
V. Shalaeva, A. F. Esfahani, P. Germain, and M. Petreczky · 2019
Closest in time.
Can you trust your model’s uncertainty? Evaluating predictive uncertainty under dataset shift
J. Snoek, Y. Ovadia, E. Fertig, B. Lakshminarayanan, S. Nowozin, D. Sculley, J. Dillon, J. Ren, and Z. Nado · 2019
Closest in time.
Variational Bayes under model misspecification
Y. Wang and D. Blei · 2019
Closest in time.
Second order PAC-Bayesian bounds for the weighted majority vote
A. R. Masegosa, S. S. Lorenzen, C. Igel, and Y. Seldin · 2020
Closest in time.
Pseudo-bayesian learning via direct loss minimization with applications to sparse gaussian process models
R. Sheth and R. Khardon · 2020
Closest in time.
Direct loss minimization for sparse gaussian processes
Y. Wei, R. Sheth, and R. Khardon · 2020
Closest in time.
How good is the Bayes posterior in deep neural networks really?
F. Wenzel, K. Roth, B. S. Veeling, J. Świątkowski, L. Tran, S. Mandt, J. Snoek, T. Salimans, R. Jenatton, and S. Nowozin · 2020
Closest in time.
Bayesian deep learning and a probabilistic perspective of generalization
A. G. Wilson and P. Izmailov · 2020
Closest in time.