Fetching the paper…
Reading the bibliography…
Recently, denoising diffusion probabilistic models and generative score matching have shown high potential in modelling complex data distributions while stochastic calculus has provided a unified point of view on these techniques allowing for flexible inference schemes.
Statistics of Random Processes , volume 5 of Stochastic Modelling and Applied Probability
Liptser, R. and Shiryaev, A · 1978
Earlier work this paper cites.
Reverse-time diffusion equation models
Anderson, B. D · 1982
Earlier work this paper cites.
A Tutorial on Hidden Markov Models and Selected Applications
Rabiner, L · 1989
Earlier work this paper cites.
Numerical Solution of Stochastic Differential Equations , volume 23 of Stochastic Modelling and Applied Probability
Kloeden, P. E. and Platen, E · 1992
Earlier work this paper cites.
Generative Adversarial Nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Earlier work this paper cites.
Variational inference with normalizing flows
Rezende, D. J. and Mohamed, S · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Deep Unsupervised Learning using Nonequilibrium Thermodynamics
Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., and Ganguli, S · 2015
Earlier work this paper cites.
WaveNet: A Generative Model for Raw Audio
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
Improved Training of Wasserstein GANs
Gulrajani, I., Ahmed, F., Arjovsky, M., Dumoulin, V., and Courville, A. C · 2017
Earlier work this paper cites.
The LJ Speech Dataset, 2017
Ito, K · 2017
Earlier work this paper cites.
Neural Ordinary Differential Equations
Chen, R. T. Q., Rubanova, Y., Bettencourt, J., and Duvenaud, D. K · 2018
Earlier work this paper cites.
Glow: Generative flow with invertible 1x1 convolutions
Kingma, D. P. and Dhariwal, P · 2018
Cited alongside, same era.
Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions
Shen, J., Pang, R., et al · 2018
Cited alongside, same era.
Parallel WaveNet: Fast High-Fidelity Speech Synthesis
van den Oord, A., Li, Y., et al · 2018
Cited alongside, same era.
MelGAN: Generative Adversarial Networks for Conditional Waveform Synthesis
Kumar, K., Kumar, R., de Boissiere, T., Gestin, L., et al · 2019
Cited alongside, same era.
Neural Speech Synthesis with Transformer Network
Li, N., Liu, S., Liu, Y., Zhao, S., and Liu, M · 2019
Cited alongside, same era.
Waveglow: A Flow-based Generative Network for Speech Synthesis
Prenger, R., Valle, R., and Catanzaro, B · 2019
Cited alongside, same era.
Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search
Kim, J., Kim, S., Kong, J., and Yoon, S · 2020
Later among the works it cites.
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
Kong, J., Kim, J., and Bae, J · 2020
Later among the works it cites.
Diffusion models for Handwriting Generation, 2020
Luhman, T. and Luhman, E · 2020
Later among the works it cites.
Permutation Invariant Graph Generation via Score-Based Generative Modeling
Niu, C., Song, Y., Song, J., Zhao, S., Grover, A., and Ermon, S · 2020
Later among the works it cites.
Shen, J., Jia, Y., Chrzanowski, M., Zhang, Y., Elias, I., Zen, H., and Wu, Y · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FastSpeech: Fast, Robust and Controllable Text to Speech
Ren, Y., Ruan, Y., Tan, X., Qin, T., et al · 2019
Cited alongside, same era.
Generative Modeling by Estimating Gradients of the Data Distribution
Song, Y. and Ermon, S · 2019
Cited alongside, same era.
High Fidelity Speech Synthesis with Adversarial Networks
Bińkowski, M., Donahue, J., Dieleman, S., Clark, A., et al · 2020
Cited alongside, same era.
Learning Gradient Fields for Shape Generation
Cai, R., Yang, G., Averbuch-Elor, H., Hao, Z., Belongie, S., Snavely, N., and Hariharan, B · 2020
Cited alongside, same era.
Parallel Tacotron: Non-Autoregressive and Controllable TTS, 2020
Elias, I., Zen, H., Shen, J., Zhang, Y., Jia, Y., Weiss, R., and Wu, Y · 2020
Cited alongside, same era.
Denoising Diffusion Probabilistic Models
Ho, J., Jain, A., and Abbeel, P · 2020
Cited alongside, same era.
Improved Techniques for Training Score-Based Generative Models
Song, Y. and Ermon, S · 2020
Later among the works it cites.
Parallel Wavegan: A Fast Waveform Generation Model Based on Generative Adversarial Networks with Multi-Resolution Spectrogram
Yamamoto, R., Song, E., and Kim, J.-M · 2020
Later among the works it cites.
WaveGrad: Estimating Gradients for Waveform Generation
Chen, N., Zhang, Y., Zen, H., Weiss, R. J., Norouzi, M., and Chan, W · 2021
Closest in time.
End-to-end Adversarial Text-to-Speech
Donahue, J., Dieleman, S., Binkowski, M., Elsen, E., and Simonyan, K · 2021
Closest in time.
DiffWave: A Versatile Diffusion Model for Audio Synthesis
Kong, Z., Ping, W., Huang, J., Zhao, K., and Catanzaro, B · 2021
Closest in time.
Score-Based Generative Modeling through Stochastic Differential Equations
Song, Y., Sohl-Dickstein, J., Kingma, D. P., Kumar, A., Ermon, S., and Poole, B · 2021
Closest in time.