Fetching the paper…
Reading the bibliography…
Black-box variational inference (BBVI) now sees widespread use in machine learning and statistics as a fast yet flexible alternative to Markov chain Monte Carlo methods for approximate Bayesian inference.
[author] Ruppert, DD. (1988). Efficient estimations from a slowly convergent Robbins–Monro process Technical Report, Cornell University Operations Research and Industrial Engineering
1988
Earlier work this paper cites.
[author] Pflug, Georg ChG. C. (1990). Non-asymptotic confidence bounds for stochastic approximation algorithms with constant step size. Monatshefte für Mathematik 110 297–314
1990
Earlier work this paper cites.
[author] Gelman, AndrewA. and Rubin, Donald BD. B. (1992). Inference from iterative simulation using multiple sequences. Statistical Science 7 457–511
1992
Earlier work this paper cites.
[author] Geyer, C JC. J. (1992). Practical Markov Chain Monte Carlo. Statistical Science 7 473–483
1992
Earlier work this paper cites.
[author] Polyak, B TB. T. and Juditsky, A BA. B. (1992). Acceleration of stochastic approximation by averaging. SIAM Journal of Control and Optimization 30 838–855
1992
Earlier work this paper cites.
2002
Earlier work this paper cites.
[author] Bolley, FrançoisF. and Villani, CC. (2005). Weighted Csiszár-Kullback-Pinsker inequalities and applications to transportation inequalities. Annales de la faculte des sciences de Toulouse 13 331–352
2005
Earlier work this paper cites.
[author] Bishop, C. M.C. M. (2006). Pattern Recognition and Machine Learning. Springer
2006
Earlier work this paper cites.
[author] Gretton, ArthurA., Borgwardt, KarstenK., Rasch, MalteM., Schölkopf, BernhardB. and Smola, AlexA. (2006). A kernel method for the two-sample-problem. Advances in neural information processing systems 19
2006
Earlier work this paper cites.
[author] Nocedal, JorgeJ. and Wright, StephenS. (2006). Numerical optimization. Springer Science & Business Media
2006
Earlier work this paper cites.
[author] Robert, C PC. P. (2007). The Bayesian Choice, 2nd ed. Springer, New York, NY
2007
Earlier work this paper cites.
[author] Cornebise, JulienJ., Moulines, EE. and Olsson, JimmyJ. (2008). Adaptive methods for sequential importance sampling with application to state space models. Statistics and Computing 18 461–480
2008
Earlier work this paper cites.
[author] Wainwright, M. J.M. J. and Jordan, M IM. I. (2008). Graphical Models, Exponential Families, and Variational Inference. Foundations and Trends® in Machine Learning 1 1–305
2008
Earlier work this paper cites.
[author] Villani, CC. (2009). Optimal transport: old and new. Grundlehren der mathematischen Wissenschaften 338. Springer
2009
Earlier work this paper cites.
[author] Joulin, AldéricA. and Ollivier, YannY. (2010). Curvature, concentration and error estimates for Markov chain Monte Carlo. The Annals of Probability 38 2418–2442
2010
Earlier work this paper cites.
[author] Madras, NealN. and Sezer, DenizD. (2010). Quantitative bounds for Markov chain convergence: Wasserstein and total variation distances. Bernoulli 16 882–908
2010
Earlier work this paper cites.
[author] Duchi, JohnJ., Hazan, EladE. and Singer, YoramY. (2011). Adaptive Subgradient Methods for Online Learning and Stochastic Optimization. Journal of Machine Learning Research 12 2121–2159
2011
Earlier work this paper cites.
Hinton, G. E
2012
Earlier work this paper cites.
[author] Murphy, KK. (2012). Machine Learning: A Probabilistic Perspective. MIT Press, Cambridge, MA
2012
Earlier work this paper cites.
Paisley, J. W
2012
Earlier work this paper cites.
[author] Gelman, AndrewA., Carlin, JohnJ., Stern, HalH., Dunson, David BD. B., Vehtari, AkiA. and Rubin, Donald BD. B. (2013). Bayesian Data Analysis, Third ed. Chapman and Hall/CRC
2013
Earlier work this paper cites.
[author] Salimans, TimT. and Knowles, David AD. A. (2013). Fixed-form variational posterior approximation through stochastic linear regression. ArXiv
2013
Earlier work this paper cites.
Kingma, D. P
2014
Earlier work this paper cites.
Ranganath, R
2014
Earlier work this paper cites.
Rezende, D. J
2014
Earlier work this paper cites.
Titsias, M. K
2014
Cited alongside, same era.
[author] Bubeck, SébastienS. (2015). Convex Optimization: Algorithms and Complexity. Foundations and Trends® in Machine Learning 8 231 – 357. 10.1561/2200000050
2015
Cited alongside, same era.
Kingma, D. P
2015
Cited alongside, same era.
Kucukelbir, A
2015
Cited alongside, same era.
Salimans, T
2015
Cited alongside, same era.
[author] Amari, Shun-ichiS.-i. (2016). Information geometry and its applications 194. Springer
2016
Cited alongside, same era.
Burda, Y
2016
Cited alongside, same era.
[author] Domke, JustinJ. (2019). Provable gradient variance guarantees for black-box variational inference. Advances in Neural Information Processing Systems 32
2019
Later among the works it cites.
[author] Durmus, AlainA., Majewski, SzymonS. and Miasojedow, BlazejB. (2019). Analysis of Langevin Monte Carlo via convex optimization. Journal of Machine Learning Research 20 1–46
2019
Later among the works it cites.
[author] Durmus, AlainA. and Moulines, EE. (2019). High-dimensional Bayesian inference via the unadjusted Langevin algorithm. Bernoulli 25 2854–2882
2019
Later among the works it cites.
[author] Eberle, AndreasA. and Majka, Mateusz BM. B. (2019). Quantitative contraction rates for Markov chains on general state spaces. Electronic Journal of Probability 24 1–36
2019
Later among the works it cites.
Gitman, I
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hernández-Lobato, J. M
2016
Cited alongside, same era.
Johnson, M. J
2016
Cited alongside, same era.
[author] Ruiz, Francisco RF. R., AUEB, Titsias RCT. R., Blei, DavidD. et al. (2016). The generalized reparameterization gradient. Advances in neural information processing systems 29
2016
Cited alongside, same era.
[author] Vollmer, Sebastian JS. J., Zygalakis, Konstantinos CK. C. and Teh, Y WY. W. (2016). (Non-) asymptotic properties of Stochastic Gradient Langevin Dynamics. Journal of Machine Learning Research 17 1–48
2016
Cited alongside, same era.
[author] Bamler, RobertR., Zhang, ChengC., Opper, ManfredM. and Mandt, StephanS. (2017). Perturbative black box variational inference. Advances in Neural Information Processing Systems 30
2017
Cited alongside, same era.
[author] Lang, HunterH., Zhang, PengchuanP. and Xiao, LinL. (2019). Using Statistics to Automate Stochastic Optimization. Advances in neural information processing systems 32
2019
Later among the works it cites.
Maddox, W
2019
Later among the works it cites.
Yaida, S
2019
Later among the works it cites.
Agrawal, A
2020
Later among the works it cites.
Boustati, A
2020
Later among the works it cites.
Dhaka, A. K
2020
Later among the works it cites.
[author] Dieuleveut, AymericA., Durmus, AlainA. and Bach, FF. (2020). Bridging the Gap between Constant Step Size Stochastic Gradient Descent and Markov Chains. The Annals of Statistics 48 1348–1382
2020
Later among the works it cites.
[author] Geffner, TomasT. and Domke, JustinJ. (2020). Approximation based variance reduction for reparameterization gradients. Advances in Neural Information Processing Systems 33 2397–2407
2020
Later among the works it cites.
Huggins, J. H
2020
Later among the works it cites.
[author] Mohamed, ShakirS., Rosca, MihaelaM., Figurnov, MichaelM. and Mnih, AndriyA. (2020). Monte Carlo Gradient Estimation in Machine Learning. Journal of Machine Learning Research 21 1–62
2020
Later among the works it cites.
Pesme, S
2020
Later among the works it cites.
[author] Stan Development Team (2020). Stan Modeling Language Users Guide and Reference Manual
2020
Later among the works it cites.
[author] Wan, NengN., Li, DapengD. and Hovakimyan, NairaN. (2020). f-Divergence Variational Inference. Advances in neural information processing systems 33 17370–17379
2020
Later among the works it cites.
[author] Dhaka, Akash KumarA. K., Catalina, AlejandroA., Welandawe, ManushiM., Andersen, Michael RM. R., Huggins, JonathanJ. and Vehtari, AkiA. (2021). Challenges and Opportunities in High Dimensional Variational Inference. Advances in Neural Information Processing Systems 34
2021
Later among the works it cites.
[author] Liu, SifanS. and Owen, Art BA. B. (2021). Quasi-Monte Carlo Quasi-Newton in Variational Bayes. Journal of Machine Learning Research 22 1–23
2021
Later among the works it cites.
[author] Papamakarios, GeorgeG., Nalisnick, EricE., Rezende, Danilo JimenezD. J., Mohamed, ShakirS. and Lakshminarayanan, BalajiB. (2021). Normalizing Flows for Probabilistic Modeling and Inference. Journal of Machine Learning Research 22 1–64
2021
Later among the works it cites.
[author] Vats, DootikaD. and Knudson, ChristinaC. (2021). Revisiting the Gelman-Rubin Diagnostic. Statistical Science 36 518–529
2021
Later among the works it cites.
[author] Vehtari, AkiA., Gelman, AndrewA., Simpson, DanielD., Carpenter, BobB. and Bürkner, Paul-ChristianP.-C. (2021). Rank-Normalization, Folding, and Localization: An Improved R ^ \widehat{R} for Assessing Convergence of MCMC. Bayesian Analysis 16 667–718
2021
Later among the works it cites.
2022
Closest in time.