Fetching the paper…
Reading the bibliography…
Denoising diffusion probabilistic models (DDPMs) have emerged as competitive generative models yet brought challenges to efficient sampling.
R. Kubichek, “Mel-cepstral distance measure for objective speech quality assessment,” in Proceedings of IEEE Pacific Rim Conference on Communications Computers and Signal Processing , vol. 1. IEEE, 1993, pp. 125–128
1993
Earlier work this paper cites.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,” in 2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No. 01CH37221) , vol. 2. IEEE, 2001, pp. 749–752
2001
Earlier work this paper cites.
G. E. Hinton, “Training products of experts by minimizing contrastive divergence. neural computation,” Neural computation , pp. 14(8):1771–1800, 2002
2002
Earlier work this paper cites.
M. A. Carreira-Perpinan and G. E. Hinton, “On contrastive divergence learning,” AISTATS , pp. 33–40, 2005
2005
Earlier work this paper cites.
A. Hyvarinen and P. Dayan, “Estimation of non-normalized statistical models by score matching,” Journal of Machine Learning Research , p. 6(4), 2005
2005
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” in Computer Science , 2009
2009
Earlier work this paper cites.
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “A short-time objective intelligibility measure for time-frequency weighted noisy speech,” in 2010 IEEE international conference on acoustics, speech and signal processing . IEEE, 2010, pp. 4214–4217
2010
Earlier work this paper cites.
P. Vincent, “A connection between score matching and denoising autoencoders,” Neural Computation , p. 23(7):1661–1674, 2011
2011
Earlier work this paper cites.
F. Protasio Ribeiro, D. Florencio, C. Zhang, and M. Seltzer, “CROWDMOS: An approach for crowdsourcing mean opinion score studies,” in ICASSP . IEEE, 2011, edition: ICASSP
2011
Earlier work this paper cites.
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
D. J. Rezende, S. Mohamed, and D. Wierstra, “Stochastic backpropagation and approximate inference in deep generative models,” ICML , pp. 1278–1286, 2014
2014
Earlier work this paper cites.
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli, “Deep unsupervised learning using nonequilibrium thermodynamics,” In International Conference on Machine Learning , pp. 2256–2265, 2015
2015
Earlier work this paper cites.
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, and K. Kavukcuoglu, “Wavenet: A generative model for raw audio,” Proc. 9th ISCA Speech Synthesis Workshop , pp. 125–125, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Ito and L. Johnson, “The lj speech dataset,” https://keithito.com/LJ-Speech-Dataset/ , 2017
2017
Earlier work this paper cites.
D. P. Kingma and P. Dhariwal, “Glow: Generative flow with invertible 1x1 convolutions,” NeurIPS , pp. 10 215–10 224, 2018
2018
Cited alongside, same era.
T. Q. Chen, Y. Rubanova, J. Bettencourt, and D. K. Duvenaud, “Neural ordinary differential equations,” NeurIPs , pp. 6571–6583, 2018
2018
Cited alongside, same era.
N. Kalchbrenner, E. Elsen, K. Simonyan, S. Noury, N. Casagrande, E. Lockhart, F. Stimberg, A. Oord, S. Dieleman, and K. Kavukcuoglu, “Efficient neural audio synthesis,” In International Conference on Machine Learning , pp. 2410–2419, 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
R. Yamamoto, E. Song, and J.-M. K. Parallel, “Wavegan: A fast waveform gen- eration model based on generative adversarial networks with multi-resolution spectrogram,” ICASSP , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
M. Bińkowski, J. Donahue, S. Dieleman, A. Clark, E. Elsen, N. Casagrande, L. C. Cobo, and K. Simonyan, “High fidelity speech synthesis with adversarial networks,” ICLR , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
W. Grathwohl, R. T. Q. Chen, J. Bettencourt, and D. Duvenaud, “Ffjord: Free-form continuous dynamics for scalable reversible generative models,” ICLR , 2019
2019
Cited alongside, same era.
J. Ho, X. Chen, A. Srinivas, Y. Duan, and P. Abbeel, “Flow++: Improving flow-based generative models with variational dequantization and architecture design,” ICML , 2019
2019
Cited alongside, same era.
L. Maaløe, M. Fraccaro, V. Liévin, and O. Winther, “Biva: A very deep hierarchy of latent variables for generative modeling,” NeurIPS , pp. 6548–6558, 2019
2019
Cited alongside, same era.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” NeurIPS , 2019
2019
Cited alongside, same era.
K. Kumar, R. Kumar, T. de Boissiere, L. Gestin, W. Z. Teoh, J. Sotelo, A. de Brebisson, Y. Bengio, and A. Courville, “Melgan: Generative adversarial networks for conditional waveform synthesis,” NeurIPS , 2019
2019
Cited alongside, same era.
J. Yamagishi, C. Veaux, and K. MacDonald, “CSTR VCTK Corpus: English multi-speaker corpus for CSTR voice cloning toolkit (version 0.92),” 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
Z. Kong, W. Ping, J. Huang, K. Zhao, and B. Catanzaro, “Diffwave: A versatile diffusion model for audio synthesis,” ICLR , 2021
2021
Closest in time.
G. Papamakarios, E. Nalisnick, D. J. Rezende, S. Mohamed, and B. Lakshminarayanan, “Normalizing flows for probabilistic modeling and inference,” JMLR , pp. 22(57):1–64, 2021
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole, “Score-based generative modeling through stochastic differential equations,” ICLR , 2021
2021
Closest in time.
J. Song, C. Meng, and S. Ermon, “Denoising diffusion implicit models,” ICLR , 2021
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.