Fetching the paper…
Reading the bibliography…
While normalizing flows have led to significant advances in modeling high-dimensional continuous distributions, their applicability to discrete distributions remains unknown.
Latent normalizing flows for discrete sequences
Ziegler, Z. M. and Rush, A. M. (2019) · 1901
Earlier work this paper cites.
Integer discrete flows and lossless compression
Hoogeboom, E., Peters, J. W., van den Berg, R., and Welling, M. (2019) · 1905
Earlier work this paper cites.
The potts model
Wu, F.-Y. (1982) · 1982
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Y., Ducharme, R., Vincent, P., and Jauvin, C. (2003) · 2003
Earlier work this paper cites.
Parallel random numbers: as easy as 1, 2, 3
Salmon, J. K., Moraes, M. A., Dror, R. O., and Shaw, D. E. (2011) · 2011
Earlier work this paper cites.
Subword language modeling with neural networks
Mikolov, T., Sutskever, I., Deoras, A., Le, H.-S., Kombrink, S., and Cernocky, J. (2012) · 2012
Earlier work this paper cites.
A fast and simple algorithm for training neural probabilistic language models
Mnih, A. and Teh, Y. W. (2012) · 2012
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Léonard, N., and Courville, A. (2013) · 2013
Earlier work this paper cites.
High-dimensional probability estimation with deep density models
Rippel, O. and Adams, R. P. (2013) · 2013
Earlier work this paper cites.
A family of nonparametric density estimation algorithms
Tabak, E. and Turner, C. V. (2013) · 2013
Earlier work this paper cites.
NICE: Non-linear independent components estimation
Dinh, L., Krueger, D., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Zaremba, W. and Sutskever, I. (2014) · 2014
Earlier work this paper cites.
Bidirectional recurrent neural networks as generative models
Berglund, M., Raiko, T., Honkala, M., Kärkkäinen, L., Vetek, A., and Karhunen, J. T. (2015) · 2015
Earlier work this paper cites.
Generating sentences from a continuous space
Bowman, S. R., Vilnis, L., Vinyals, O., Dai, A. M., Jozefowicz, R., and Bengio, S. (2015) · 2015
Earlier work this paper cites.
A fast variational approach for learning markov random field language models
Jernite, Y., Rush, A., and Sontag, D. (2015) · 2015
Cited alongside, same era.
Variational inference with normalizing flows
Rezende, D. J. and Mohamed, S. (2015) · 2015
Cited alongside, same era.
Order matters: Sequence to sequence for sets
Vinyals, O., Bengio, S., and Kudlur, M. (2015) · 2015
Cited alongside, same era.
The concrete distribution: A continuous relaxation of discrete random variables
Maddison, C. J., Mnih, A., and Teh, Y. W. (2016) · 2016
Cited alongside, same era.
Unrolled generative adversarial networks
Metz, L., Poole, B., Pfau, D., and Sohl-Dickstein, J. (2016) · 2016
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I. (2017) · 2017
Later among the works it cites.
Deliberation networks: Sequence generation beyond one-pass decoding
Xia, Y., Tian, F., Wu, L., Lin, J., Qin, T., Yu, N., and Liu, T.-Y. (2017) · 2017
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2018) · 2018
Later among the works it cites.
The importance of generation order in language modeling
Ford, N., Duckworth, D., Norouzi, M., and Dahl, G. E. (2018) · 2018
Later among the works it cites.
Non-autoregressive neural machine translation
Gu, J., Bradbury, J., Xiong, C., Li, V. O., and Socher, R. (2018) · 2018
Later among the works it cites.
Fast decoding in sequence models using discrete latent variables
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hierarchical variational models
Ranganath, R., Tran, D., and Blei, D. (2016) · 2016
Cited alongside, same era.
Architectural complexity measures of recurrent neural networks
Zhang, S., Wu, Y., Che, T., Lin, Z., Memisevic, R., Salakhutdinov, R. R., and Bengio, Y. (2016) · 2016
Cited alongside, same era.
Massive exploration of neural machine translation architectures
Britz, D., Goldie, A., Luong, M.-T., and Le, Q. (2017) · 2017
Cited alongside, same era.
Density estimation using real nvp
Dinh, L., Sohl-Dickstein, J., and Bengio, S. (2017) · 2017
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Jang, E., Gu, S., and Poole, B. (2017) · 2017
Cited alongside, same era.
Parallel wavenet: Fast high-fidelity speech synthesis
Oord, A. v. d., Li, Y., Babuschkin, I., Simonyan, K., Vinyals, O., Kavukcuoglu, K., Driessche, G. v. d., Lockhart, E., Cobo, L. C., Stimberg, F., et al. (2017) · 2017
Cited alongside, same era.
Masked autoregressive flow for density estimation
Papamakarios, G., Murray, I., and Pavlakou, T. (2017) · 2017
Cited alongside, same era.
Kaiser, Ł., Roy, A., Vaswani, A., Pamar, N., Bengio, S., Uszkoreit, J., and Shazeer, N. (2018) · 2018
Later among the works it cites.
Glow: Generative flow with invertible 1x1 convolutions
Kingma, D. P. and Dhariwal, P. (2018) · 2018
Later among the works it cites.
Deterministic non-autoregressive neural sequence modeling by iterative refinement
Lee, J., Mansimov, E., and Cho, K. (2018) · 2018
Later among the works it cites.
An analysis of neural language modeling at multiple scales
Merity, S., Keskar, N. S., and Socher, R. (2018) · 2018
Later among the works it cites.
Clarinet: Parallel wave generation in end-to-end text-to-speech
Ping, W., Peng, K., and Chen, J. (2018) · 2018
Later among the works it cites.
Waveglow: A flow-based generative network for speech synthesis
Prenger, R., Valle, R., and Catanzaro, B. (2018) · 2018
Later among the works it cites.
Blockwise parallel decoding for deep autoregressive models
Stern, M., Shazeer, N., and Uszkoreit, J. (2018) · 2018
Later among the works it cites.
Bayesian layers: A module for neural network uncertainty
Tran, D., Mike, D., van der Wilk, M., and Hafner, D. (2018) · 2018
Later among the works it cites.