A hierarchical latent vector model for learning long-term structure in music
Roberts, A., Engel, J., Raffel, C., Hawthorne, C., and Eck, D · 2018
Later among the works it cites.
Natural TTS synthesis by conditioning wavenet on mel spectrogram predictions
Shen, J., Pang, R., Weiss, R. J., Schuster, M., Jaitly, N., Yang, Z., Chen, Z., Zhang, Y., Wang, Y., Skerrv-Ryan, R., Saurous, R. A., Agiomvrgiannakis, Y., and Wu, Y · 2018
Later among the works it cites.
VoiceLoop: Voice fitting and synthesis via a phonological loop
Taigman, Y., Wolf, L., Polyak, A., and Nachmani, E · 2018
Later among the works it cites.
Style Tokens: Unsupervised style modeling, control and transfer in end-to-end speech synthesis
Wang, Y., Stanton, D., Zhang, Y., Skerry-Ryan, R., Battenberg, E., Shor, J., Xiao, Y., Ren, F., Jia, Y., and Saurous, R. A · 2018
Later among the works it cites.
Dota 2 with large scale deep reinforcement learning
Original
Berner, C., Brockman, G., Chan, B., Cheung, V., Dębiak, P., Dennison, C., Farhi, D., Fischer, Q., Hashme, S., Hesse, C., et al · 2019
Later among the works it cites.
Large scale GAN training for high fidelity natural image synthesis
Brock, A., Donahue, J., and Simonyan, K · 2019
Later among the works it cites.
Generating long sequences with sparse transformers
Original
Child, R., Gray, S., Radford, A., and Sutskever, I · 2019
Later among the works it cites.
Hierarchical autoregressive image models with auxiliary decoders
Original
De Fauw, J., Dieleman, S., and Simonyan, K · 2019
Later among the works it cites.
LakhNES: Improving multi-instrumental music generation with cross-domain pre-training
Donahue, C., Mao, H. H., Li, Y. E., Cottrell, G. W., and McAuley, J. J · 2019
Later among the works it cites.
GANSynth: Adversarial neural audio synthesis
Engel, J., Agrawal, K. K., Chen, S., Gulrajani, I., Donahue, C., and Roberts, A · 2019
Later among the works it cites.
Enabling factorized piano music modeling and generation with the MAESTRO dataset
Hawthorne, C., Stasyuk, A., Roberts, A., Simon, I., Huang, C.-Z. A., Dieleman, S., Elsen, E., Engel, J., and Eck, D · 2019
Later among the works it cites.
Spleeter: A fast and state-of-the art music source separation tool with pre-trained models
Hennequin, R., Khlif, A., Voituret, F., and Moussallam, M · 2019
Later among the works it cites.
Axial attention in multidimensional transformers
Original
Ho, J., Kalchbrenner, N., Weissenborn, D., and Salimans, T · 2019
Later among the works it cites.
Hierarchical generative modeling for controllable speech synthesis
Hsu, W.-N., Zhang, Y., Weiss, R. J., Zen, H., Wu, Y., Wang, Y., Cao, Y., Jia, Y., Chen, Z., Shen, J., Nguyen, P., and Pang, R · 2019
Later among the works it cites.
Neural music synthesis for flexible timbre control
Kim, J. W., Bittner, R., Kumar, A., and Bello, J. P · 2019
Later among the works it cites.
MelGAN: Generative adversarial networks for conditional waveform synthesis
Kumar, K., Kumar, R., de Boissiere, T., Gestin, L., Teoh, W. Z., Sotelo, J., de Brébisson, A., Bengio, Y., and Courville, A. C · 2019
Later among the works it cites.
Autoencoder-based music translation
Mor, N., Wolf, L., Polyak, A., and Taigman, Y · 2019
Later among the works it cites.
Clarinet: Parallel wave generation in end-to-end text-to-speech
Ping, W., Peng, K., and Chen, J · 2019
Later among the works it cites.
WaveGlow: A flow-based generative network for speech synthesis
Prenger, R., Valle, R., and Catanzaro, B · 2019
Later among the works it cites.
Generating diverse high-fidelity images with vq-vae-2
Razavi, A., van den Oord, A., and Vinyals, O · 2019
Later among the works it cites.
MelNet: A generative model for audio in the frequency domain
Original
Vasquez, S. and Lewis, M · 2019
Later among the works it cites.
A hierarchical recurrent neural network for symbolic melody generation
Wu, J., Hu, C., Wang, Y., Hu, X., and Zhu, J · 2019
Later among the works it cites.
Self-attention generative adversarial networks
Zhang, H., Goodfellow, I., Metaxas, D., and Odena, A · 2019
Later among the works it cites.
Automatic lyrics transcription in polyphonic music: Does background music help?
Gupta, C., Yılmaz, E., and Li, H · 2020
Closest in time.
Parallel WaveGAN: A fast waveform generation model based on generative adversarial networks with multi-resolution spectrogram
Yamamoto, R., Song, E., and Kim, J.-M · 2020
Closest in time.