Fetching the paper…
Reading the bibliography…
Advances in variational inference enable parameterisation of probabilistic models by deep neural networks.
Latent normalizing flows for discrete sequences
Zachary M Ziegler and Alexander M Rush. 2019 · 1901
Earlier work this paper cites.
A stochastic approximation method
Herbert Robbins and Sutton Monro. 1951 · 1951
Earlier work this paper cites.
Interpolated estimation of markov source parameters from sparse data
Frederick Jelinek. 1980 · 1980
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Mitchell P Marcus, Mary Ann Marcinkiewicz, and Beatrice Santorini. 1993 · 1993
Earlier work this paper cites.
An introduction to variational methods for graphical models
MichaelI. Jordan, Zoubin Ghahramani, TommiS. Jaakkola, and LawrenceK. Saul. 1999 · 1999
Earlier work this paper cites.
A bit of progress in language modeling
Joshua T. Goodman. 2001 · 2001
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin. 2003 · 2003
Earlier work this paper cites.
Large scale online learning
Léon Bottou and Yann L. Cun. 2004 · 2004
Earlier work this paper cites.
Convex optimization
Stephen Boyd and Lieven Vandenberghe. 2004 · 2004
Earlier work this paper cites.
Gaussian Processes for Machine Learning (Adaptive Computation and Machine Learning)
Carl Edward Rasmussen and Christopher K. I. Williams. 2005 · 2005
Earlier work this paper cites.
A study of translation edit rate with targeted human annotation
Matthew Snover, Bonnie Dorr, Richard Schwartz, Linnea Micciulla, and John Makhoul. 2006 · 2006
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio. 2012 · 2012
Earlier work this paper cites.
A kernel two-sample test
Arthur Gretton, Karsten M Borgwardt, Malte J Rasch, Bernhard Schölkopf, and Alexander Smola. 2012 · 2012
Earlier work this paper cites.
Practical bayesian optimization of machine learning algorithms
Jasper Snoek, Hugo Larochelle, and Ryan P Adams. 2012 · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P. Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra. 2014 · 2014
Earlier work this paper cites.
Recurrent neural network regularization
Wojciech Zaremba, Ilya Sutskever, and Oriol Vinyals. 2014 · 2014
Earlier work this paper cites.
Importance weighted autoencoders
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov. 2015 · 2015
Earlier work this paper cites.
Dropout as a Bayesian approximation: Insights and applications
Yarin Gal and Zoubin Ghahramani. 2015 · 2015
Earlier work this paper cites.
GPyOpt: A bayesian optimization framework in python
The GPyOpt authors. 2016 · 2016
Earlier work this paper cites.
Generating sentences from a continuous space
Samuel R Bowman, Luke Vilnis, Oriol Vinyals, Andrew Dai, Rafal Jozefowicz, and Samy Bengio. 2016 · 2016
Cited alongside, same era.
Recurrent neural network grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A. Smith. 2016 · 2016
Cited alongside, same era.
Sequential neural models with stochastic layers
Marco Fraccaro, Søren Kaae Sønderby, Ulrich Paquet, and Ole Winther. 2016 · 2016
Cited alongside, same era.
Improved variational inference with inverse autoregressive flow
Diederik P Kingma, Tim Salimans, Rafal Jozefowicz, Xi Chen, Ilya Sutskever, and Max Welling. 2016 · 2016
Cited alongside, same era.
Language as a latent variable: Discrete generative models for sentence compression
Yishu Miao and Phil Blunsom. 2016 · 2016
Cited alongside, same era.
Neural variational inference for text processing
Yishu Miao, Lei Yu, and Phil Blunsom. 2016 · 2016
Automatic differentiation in machine learning: a survey
Atilim Gunes Baydin, Barak A Pearlmutter, Alexey Andreyevich Radul, and Jeffrey Mark Siskind. 2018 · 2018
Later among the works it cites.
Differentiable perturb-and-parse: Semi-supervised parsing with a structured variational autoencoder
Caio Corro and Ivan Titov. 2018 · 2018
Later among the works it cites.
Hyperspherical variational auto-encoders
Tim R Davidson, Luca Falorsi, Nicola De Cao, Thomas Kipf, and Jakub M Tomczak. 2018 · 2018
Later among the works it cites.
Semi-amortized variational autoencoders
Yoon Kim, Sam Wiseman, Andrew Miller, David Sontag, and Alexander Rush. 2018 · 2018
Later among the works it cites.
AMR parsing as graph prediction with latent alignment
Chunchuan Lyu and Ivan Titov. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Variational autoencoder for deep learning of images, labels and captions
Yunchen Pu, Zhe Gan, Ricardo Henao, Xin Yuan, Chunyuan Li, Andrew Stevens, and Lawrence Carin. 2016 · 2016
Cited alongside, same era.
Ladder variational autoencoders
Casper Kaae Sønderby, Tapani Raiko, Lars Maaløe, Søren Kaae Sønderby, and Ole Winther. 2016 · 2016
Cited alongside, same era.
Variational neural machine translation
Biao Zhang, Deyi Xiong, jinsong su, Hong Duan, and Min Zhang. 2016 · 2016
Cited alongside, same era.
Variational lossy autoencoder
Xi Chen, Diederik P Kingma, Tim Salimans, Yan Duan, Prafulla Dhariwal, John Schulman, Ilya Sutskever, and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Z-forcing: Training stochastic recurrent networks
Anirudh Goyal Alias Parth Goyal, Alessandro Sordoni, Marc-Alexandre Côté, Nan Rosemary Ke, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
beta-VAE: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, Christopher Burgess, Xavier Glorot, Matthew Botvinick, Shakir Mohamed, and Alexander Lerchner. 2017 · 2017
Cited alongside, same era.
Gábor Melis, Chris Dyer, and Phil Blunsom. 2018 · 2018
Later among the works it cites.
A hierarchical latent structure for variational conversation modeling
Yookoon Park, Jaemin Cho, and Gunhee Kim. 2018 · 2018
Later among the works it cites.
Danilo Jimenez Rezende and Fabio Viola. 2018 · 2018
Later among the works it cites.
Deep generative model for joint alignment and word representation
Miguel Rios, Wilker Aziz, and Khalil Simaan. 2018 · 2018
Later among the works it cites.
A stochastic decoder for neural machine translation
Philip Schulz, Wilker Aziz, and Trevor Cohn. 2018 · 2018
Later among the works it cites.
Wasserstein auto-encoders
Ilya Tolstikhin, Olivier Bousquet, Sylvain Gelly, and Bernhard Schoelkopf. 2018 · 2018
Later among the works it cites.
Jakub M Tomczak and Max Welling. 2018 · 2018
Later among the works it cites.
Spherical latent spaces for stable variational autoencoders
Jiacheng Xu and Greg Durrett. 2018 · 2018
Later among the works it cites.
Reweighted expectation maximization
Adji Dieng and John Paisley. 2019 · 2019
Closest in time.
Avoiding latent variable collapse with generative skip models
Adji B. Dieng, Yoon Kim, Alexander M. Rush, and David M. Blei. 2019 · 2019
Closest in time.
Auto-encoding variational neural machine translation
Bryan Eikema and Wilker Aziz. 2019 · 2019
Closest in time.
Lagging inference networks and posterior collapse in variational autoencoders
Junxian He, Daniel Spokoyny, Graham Neubig, and Taylor Berg-Kirkpatrick. 2019 · 2019
Closest in time.
A surprisingly effective fix for deep latent variable modeling of text
Bohan Li, Junxian He, Graham Neubig, Taylor Berg-Kirkpatrick, and Yiming Yang. 2019 · 2019
Closest in time.
Cyclical annealing schedule: A simple approach to mitigating KL vanishing
Xiaodong Liu, Jianfeng Gao, Asli Celikyilmaz, Lawrence Carin, et al. 2019 · 2019
Closest in time.
FlowSeq: Non-autoregressive conditional sequence generation with generative flow
Xuezhe Ma, Chunting Zhou, Xian Li, Graham Neubig, and Eduard Hovy. 2019 · 2019
Closest in time.
Preventing posterior collapse with delta-VAEs
Ali Razavi, Aäron van den Oord, Ben Poole, and Oriol Vinyals. 2019 · 2019
Closest in time.