Gansynth: Adversarial neural audio synthesis
Original
Engel, J., Agrawal, K. K., Chen, S., Gulrajani, I., Donahue, C., and Roberts, A · 1902
Earlier work this paper cites.
Gansynth: Adversarial neural audio synthesis
Original
Engel, J., Agrawal, K. K., Chen, S., Gulrajani, I., Donahue, C., and Roberts, A · 1902
Earlier work this paper cites.
Signal estimation from modified short-time fourier transform
Griffin, D. and Lim, J · 1984
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Earlier work this paper cites.
Autoencoding beyond pixels using a learned similarity metric
Original
Larsen, A. B. L., Sønderby, S. K., Larochelle, H., and Winther, O · 2015
Earlier work this paper cites.
Deep multi-scale video prediction beyond mean square error
Original
Mathieu, M., Couprie, C., and LeCun, Y · 2015
Earlier work this paper cites.
Generating images with perceptual similarity metrics based on deep networks
Dosovitskiy, A. and Brox, T · 2016
Earlier work this paper cites.
Image style transfer using convolutional neural networks
Gatys, L. A., Ecker, A. S., and Bethge, M · 2016
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution
Johnson, J., Alahi, A., and Fei-Fei, L · 2016
Earlier work this paper cites.
Samplernn: An unconditional end-to-end neural audio generation model
Original
Mehri, S., Kumar, K., Gulrajani, I., Kumar, R., Jain, S., Sotelo, J., Courville, A., and Bengio, Y · 2016
Earlier work this paper cites.
World: A vocoder-based high-quality speech synthesis system for real-time applications
MORISE, M., YOKOMORI, F., and OZAWA, K · 2016
Earlier work this paper cites.
Deconvolution and checkerboard artifacts
Odena, A., Dumoulin, V., and Olah, C · 2016
Earlier work this paper cites.
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
Salimans, T. and Kingma, D. P · 2016
Earlier work this paper cites.
Instance normalization: The missing ingredient for fast stylization
Original
Ulyanov, D., Vedaldi, A., and Lempitsky, V · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Van Den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A. W., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
Deep voice: Real-time neural text-to-speech
Arik, S. Ö., Chrzanowski, M., Coates, A., Diamos, G., Gibiansky, A., Kang, Y., Li, X., Miller, J., Ng, A., Raiman, J., et al · 2017
Earlier work this paper cites.