Fetching the paper…
Reading the bibliography…
In this work we define a unified mathematical framework to deepen our understanding of the role of stochastic gradient (SG) noise on the behavior of Markov chain Monte Carlo sampling (SGMCMC) algorithms.
Stochastic Approximation and Optimization of Random Systems
L. Ljung, G. Pflug, and H. Walk · 1992
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Stochastic Approximation and Recursive Algorithms and Applications
H. Kushner and G. Yin · 2003
Earlier work this paper cites.
Handbook of stochastic methods for physics, chemistry and the natural sciences
C. W. Gardiner · 2004
Earlier work this paper cites.
Pattern recognition and machine learning
C. M. Bishop · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky, G. Hinton, et al · 2009
Earlier work this paper cites.
Mnist handwritten digit database
Y. LeCun, C. Cortes, and C. Burges · 2010
Earlier work this paper cites.
Riemann manifold langevin and hamiltonian monte carlo methods
M. Girolami and B. Calderhead · 2011
Earlier work this paper cites.
Mcmc using hamiltonian dynamics
R. M. Neal et al · 2011
Earlier work this paper cites.
Bayesian learning via stochastic gradient langevin dynamics
M. Welling and Y. W. Teh · 2011
Earlier work this paper cites.
Bayesian posterior sampling via stochastic gradient fisher scoring
S. Ahn, A. Korattikara, and M. Welling · 2012
Earlier work this paper cites.
Stochastic gradient riemannian langevin dynamics on the probability simplex
S. Patterson and Y. W. Teh · 2013
Cited alongside, same era.
Stochastic gradient hamiltonian monte carlo
T. Chen, E. Fox, and C. Guestrin · 2014
Cited alongside, same era.
A complete recipe for stochastic gradient mcmc
Y.-A. Ma, T. Chen, and E. Fox · 2015
Cited alongside, same era.
Bridging the gap between stochastic gradient mcmc and stochastic optimization
C. Chen, D. Carlson, Z. Gan, C. Li, and L. Carin · 2016
Cited alongside, same era.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning
Y. Gal and Z. Ghahramani · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Fractional langevin monte carlo: Exploring lévy driven stochastic differential equations for markov chain monte carlo
U. ŞimŠekli · 2017
Later among the works it cites.
Stochastic gradient descent performs variational inference, converges to limit cycles for deep networks
P. Chaudhari and S. Soatto · 2018
Later among the works it cites.
Essentially no barriers in neural network energy landscape
F. Draxler, K. Veschgini, M. Salmhofer, and F. Hamprecht · 2018
Later among the works it cites.
Loss surfaces, mode connectivity, and fast ensembling of dnns
T. Garipov, P. Izmailov, D. Podoprikhin, D. P. Vetrov, and A. G. Wilson · 2018
Later among the works it cites.
Explicit and combined estimators for stable distributions parameters
J. Levy Vehel, A. Philippe, and C. Robet · 2018
Later among the works it cites.
A simple baseline for bayesian uncertainty in deep learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Preconditioned stochastic gradient langevin dynamics for deep neural networks
C. Li, C. Chen, D. Carlson, and L. Carin · 2016
Cited alongside, same era.
Bayesian optimization with robust bayesian neural networks
J. T. Springenberg, A. Klein, S. Falkner, and F. Hutter · 2016
Cited alongside, same era.
Accelerating very deep convolutional networks for classification and detection
X. Zhang, J. Zou, K. He, and J. Sun · 2016
Cited alongside, same era.
UCI machine learning repository, 2017
D. Dua and C. Graff · 2017
Cited alongside, same era.
Stochastic gradient descent as approximate bayesian inference
S. Mandt, M. D. Hoffman, and D. M. Blei · 2017
Cited alongside, same era.
W. J. Maddox, P. Izmailov, T. Garipov, D. P. Vetrov, and A. G. Wilson · 2019
Later among the works it cites.
On the information bottleneck theory of deep learning
A. M. Saxe, Y. Bansal, J. Dapello, M. Advani, A. Kolchinsky, B. D. Tracey, and D. D. Cox · 2019
Later among the works it cites.
On the heavy-tailed theory of stochastic gradient descent for deep neural networks
U. Şimşekli, M. Gürbüzbalaban, T. H. Nguyen, G. Richard, and L. Sagun · 2019
Later among the works it cites.
A tail-index analysis of stochastic gradient noise in deep neural networks
U. Şimşekli, L. Sagun, and M. Gürbüzbalaban · 2019
Later among the works it cites.
Z. Zhu, J. Wu, B. Yu, L. Wu, and J. Ma · 2019
Later among the works it cites.
Cyclical stochastic gradient mcmc for bayesian deep learning
R. Zhang, C. Li, J. Zhang, C. Chen, and A. G. Wilson · 2020
Closest in time.