Fetching the paper…
Reading the bibliography…
Autoregressive language models are powerful and relatively easy to train.
Hierarchical autoregressive image models with auxiliary decoders
Jeffrey De Fauw, Sander Dieleman, and Karen Simonyan. 2019 · 1903
Earlier work this paper cites.
Effective Estimation of Deep Generative Language Models
Tom Pelsmaeker and Wilker Aziz. 2019 · 1904
Earlier work this paper cites.
Unsupervised paraphrasing without translation
Aurko Roy and David Grangier. 2019 · 1905
Earlier work this paper cites.
Unsupervised controllable text generation with global variation discovery and disentanglement
Peng Xu, Yanshuai Cao, and Jackie Chi Kit Cheung. 2019 · 1905
Earlier work this paper cites.
Preventing Posterior Collapse in Sequence VAEs with Pooling
Teng Long, Yanshuai Cao, and Jackie Chi Kit Cheung. 2019 · 1911
Earlier work this paper cites.
Tensor product variable binding and the representation of symbolic structures in connectionist systems
Paul Smolensky. 1990 · 1990
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Paul J Werbos. 1990 · 1990
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Thumbs up?: sentiment classification using machine learning techniques
Bo Pang, Lillian Lee, and Shivakumar Vaithyanathan. 2002 · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin. 2003 · 2003
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton. 2010 · 2010
Earlier work this paper cites.
A first course in design and analysis of experiments
Gary W Oehlert. 2010 · 2010
Earlier work this paper cites.
Lstm neural networks for language modeling
Martin Sundermeyer, Ralf Schlüter, and Hermann Ney. 2012 · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling. 2013 · 2013
Earlier work this paper cites.
Semi-Supervised Learning with Deep Generative Models
Diederik P. Kingma, Danilo Jimenez Rezende, Shakir Mohamed, and Max Welling. 2014 · 2014
Earlier work this paper cites.
Stochastic Backpropagation and Approximate Inference in Deep Generative Models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra. 2014 · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
Importance weighted autoencoders
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov. 2015 · 2015
Cited alongside, same era.
Character-level Convolutional Networks for Text Classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Cited alongside, same era.
Generating Sentences from a Continuous Space
Samuel R. Bowman, Luke Vilnis, Oriol Vinyals, Andrew Dai, Rafal Jozefowicz, and Samy Bengio. 2016 · 2016
Cited alongside, same era.
Xi Chen, Diederik P Kingma, Tim Salimans, Yan Duan, Prafulla Dhariwal, John Schulman, Ilya Sutskever, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Semi-Amortized Variational Autoencoders
Yoon Kim, Sam Wiseman, Andrew Miller, David Sontag, and Alexander Rush. 2018 · 2018
Later among the works it cites.
Using JK fold Cross Validation to Reduce Variance When Tuning NLP Models
Henry B Moss, David S Leslie, and Paul Rayson. 2018 · 2018
Later among the works it cites.
Conditional variational autoencoder for neural machine translation
Artidoro Pagnoni, Kevin Liu, and Shangyan Li. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew E Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Unsupervised abstractive sentence summarization using length controlled variational autoencoder
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improved variational inference with inverse autoregressive flow
Durk P Kingma, Tim Salimans, Rafal Jozefowicz, Xi Chen, Ilya Sutskever, and Max Welling. 2016 · 2016
Cited alongside, same era.
Neural Variational Inference for Text Processing
Yishu Miao, Lei Yu, and Phil Blunsom. 2016 · 2016
Cited alongside, same era.
Why Neural Translations are the Right Length
Xing Shi, Kevin Knight, and Deniz Yuret. 2016 · 2016
Cited alongside, same era.
Controlling Linguistic Style Aspects in Neural Language Generation
Jessica Ficler and Yoav Goldberg. 2017 · 2017
Cited alongside, same era.
Non-autoregressive neural machine translation
Jiatao Gu, James Bradbury, Caiming Xiong, Victor OK Li, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Denoising criterion for variational auto-encoding framework
Daniel Im Jiwoong Im, Sungjin Ahn, Roland Memisevic, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Bag of Tricks for Efficient Text Classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Raphael Schumann. 2018 · 2018
Later among the works it cites.
Cyclical Annealing Schedule: A Simple Approach to Mitigating KL Vanishing
Hao Fu, Chunyuan Li, Xiaodong Liu, Jianfeng Gao, Asli Celikyilmaz, and Lawrence Carin. 2019 · 2019
Later among the works it cites.
Insertion-based decoding with automatically inferred generation order
Jiatao Gu, Qi Liu, and Kyunghyun Cho. 2019 · 2019
Later among the works it cites.
Lagging Inference Networks and Posterior Collapse in Variational Autoencoders
Junxian He, Daniel Spokoyny, Graham Neubig, and Taylor Berg-Kirkpatrick. 2019 · 2019
Later among the works it cites.
A Surprisingly Effective Fix for Deep Latent Variable Modeling of Text
Bohan Li, Junxian He, Graham Neubig, Taylor Berg-Kirkpatrick, and Yiming Yang. 2019 · 2019
Later among the works it cites.
RNNs implicitly implement tensor-product representations
R. Thomas McCoy, Tal Linzen, Ewan Dunbar, and Paul Smolensky. 2019 · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Later among the works it cites.
On the Importance of the Kullback-Leibler Divergence Term in Variational Autoencoders for Text Generation
Victor Prokhorov, Ehsan Shareghi, Yingzhen Li, Mohammad Taher Pilehvar, and Nigel Collier. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Preventing Posterior Collapse with delta-VAEs
Ali Razavi, Aaron van den Oord, Ben Poole, and Oriol Vinyals. 2019 · 2019
Later among the works it cites.
LSTM Networks Can Perform Dynamic Counting
Mirac Suzgun, Yonatan Belinkov, Stuart Shieber, and Sebastian Gehrmann. 2019 · 2019
Later among the works it cites.
Latent Normalizing Flows for Discrete Sequences
Zachary Ziegler and Alexander Rush. 2019 · 2019
Later among the works it cites.
On over-fitting in model selection and subsequent selection bias in performance evaluation
Gavin C Cawley and Nicola LC Talbot. 2010 · 2079
Closest in time.