Fetching the paper…
Reading the bibliography…
Uncertainty computation in deep learning is essential to design robust and reliable systems.
Sampling Techniques, 3rd Edition
Cochran, W. G · 1977
Earlier work this paper cites.
A mean field theory learning algorithm for neural networks
Anderson, J. R. and Peterson, C · 1987
Earlier work this paper cites.
Transforming neural-net output levels to probability distributions
Denker, J. S. and Lecun, Y · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J · 1992
Earlier work this paper cites.
Keeping the neural networks simple by minimizing the description length of the weights
Hinton, G. E. and Van Camp, D · 1993
Earlier work this paper cites.
Bayesian learning for neural networks
Neal, R. M · 1995
Earlier work this paper cites.
Mean field theory for sigmoid belief networks
Saul, L. K., Jaakkola, T., and Jordan, M. I · 1996
Earlier work this paper cites.
Ensemble learning in Bayesian neural networks
Barber, D. and Bishop, C. M · 1998
Earlier work this paper cites.
Fast curvature matrix-vector products for second-order gradient descent
Schraudolph, N. N · 2002
Earlier work this paper cites.
Information theory, inference and learning algorithms
MacKay, D. J · 2003
Earlier work this paper cites.
Monte Carlo Statistical Methods (Springer Texts in Statistics)
Robert, C. P. and Casella, G · 2005
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Bishop, C. M · 2006
Earlier work this paper cites.
Smoothing-based optimization
Leordeanu, M. and Hebert, M · 2008
Earlier work this paper cites.
The variational Gaussian approximation revisited
Opper, M. and Archambeau, C · 2009
Earlier work this paper cites.
Exploring parameter space in reinforcement learning
Rückstieß, T., Sehnke, F., Schaul, T., Wierstra, D., Sun, Y., and Schmidhuber, J · 2010
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, J., Hazan, E., and Singer, Y · 2011
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A · 2011
Earlier work this paper cites.
Piecewise bounds for estimating Bernoulli-logistic latent Gaussian models
Marlin, B., Khan, M., and Murphy, K · 2011
Earlier work this paper cites.
Fast variational inference in the conjugate exponential family
Hensman, J., Rattray, M., and Lawrence, N. D · 2012
Earlier work this paper cites.
Machine Learning: A Probabilistic Perspective
Murphy, K. P · 2012
Earlier work this paper cites.
Lecture 6.5-RMSprop: Divide the gradient by a running average of its recent magnitude
Tieleman, T. and Hinton, G · 2012
Cited alongside, same era.
Fixed-form variational posterior approximation through stochastic linear regression
Salimans, T., Knowles, D. A., et al · 2013
Cited alongside, same era.
Optimization by variational bounding
Staines, J. and Barber, D · 2013
Cited alongside, same era.
New insights and perspectives on the natural gradient method
Martens, J · 2014
Cited alongside, same era.
Black box variational inference
Ranganath, R., Gerrish, S., and Blei, D. M · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
OpenAI Gym, 2016
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Later among the works it cites.
Entropy-sgd: Biasing gradient descent into wide valleys
Chaudhari, P., Choromanska, A., Soatto, S., LeCun, Y., Baldassi, C., Borgs, C., Chayes, J. T., Sagun, L., and Zecchina, R · 2016
Later among the works it cites.
Uncertainty in Deep Learning
Gal, Y · 2016
Later among the works it cites.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Later among the works it cites.
On graduated optimization for stochastic non-convex problems
Hazan, E., Levy, K. Y., and Shalev-Shwartz, S · 2016
Later among the works it cites.
beta-VAE: Learning basic visual concepts with a constrained variational framework
Higgins, I., Matthey, L., Pal, A., Burgess, C., Glorot, X., Botvinick, M., Mohamed, S., and Lerchner, A · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deterministic policy gradient algorithms
Silver, D., Lever, G., Heess, N., Degris, T., Wierstra, D., and Riedmiller, M. A · 2014
Cited alongside, same era.
Natural evolution strategies
Wierstra, D., Schaul, T., Glasmachers, T., Sun, Y., Peters, J., and Schmidhuber, J · 2014
Cited alongside, same era.
Gradient-based adaptive stochastic search for non-differentiable optimization
Zhou, E. and Hu, J · 2014
Cited alongside, same era.
Bayesian dark knowledge
Balan, A. K., Rathod, V., Murphy, K. P., and Welling, M · 2015
Cited alongside, same era.
Weight uncertainty in neural networks
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Cited alongside, same era.
Efficient Per-Example Gradient Computations
Goodfellow, I · 2015
Cited alongside, same era.
Later among the works it cites.
Faster stochastic variational inference using proximal-gradient methods with general divergence functions
Khan, M. E., Babanezhad, R., Lin, W., Schmidt, M., and Sugiyama, M · 2016
Later among the works it cites.
Preconditioned stochastic gradient langevin dynamics for deep neural networks
Li, C., Chen, C., Carlson, D. E., and Carin, L · 2016
Later among the works it cites.
Structured and efficient variational deep learning with matrix gaussian posteriors
Louizos, C. and Welling, M · 2016
Later among the works it cites.
Distributed Bayesian learning with stochastic natural gradient expectation propagation and the posterior server
Hasenclever, L., Webb, S., Lienart, T., Vollmer, S., Lakshminarayanan, B., Blundell, C., and Teh, Y. W · 2017
Later among the works it cites.
Conjugate-computation variational inference: converting variational inference in non-conjugate models to inferences in conjugate models
Khan, M. E. and Lin, W · 2017
Later among the works it cites.
Variational Adaptive-Newton Method for Explorative Learning
Khan, M. E., Lin, W., Tangkaratt, V., Liu, Z., and Nielsen, D · 2017
Later among the works it cites.
Stochastic gradient descent as approximate Bayesian inference
Mandt, S., Hoffman, M. D., and Blei, D. M · 2017
Later among the works it cites.
Learning structured weight uncertainty in Bayesian neural networks
Sun, S., Chen, C., and Carin, L · 2017
Later among the works it cites.
The marginal value of adaptive gradient methods in machine learning
Wilson, A. C., Roelofs, R., Stern, M., Srebro, N., and Recht, B · 2017
Later among the works it cites.
Noisy networks for exploration
Fortunato, M., Azar, M. G., Piot, B., Menick, J., Osband, I., Graves, A., Mnih, V., Munos, R., Hassabis, D., Pietquin, O., et al · 2018
Closest in time.
Parameter space noise for exploration
Plappert, M., Houthooft, R., Dhariwal, P., Sidor, S., Chen, R. Y., Chen, X., Asfour, T., Abbeel, P., and Andrychowicz, M · 2018
Closest in time.
A scalable laplace approximation for neural networks
Ritter, H., Botev, A., and Barber, D · 2018
Closest in time.
Noisy natural gradient as variational inference
Zhang, G., Sun, S., Duvenaud, D. K., and Grosse, R. B · 2018
Closest in time.