Fetching the paper…
Reading the bibliography…
Efficient audio synthesis is an inherently difficult machine learning task, as human perception is sensitive to both global structure and fine-scale waveform coherence.
Learning multiscale features directly fromwaveforms
Zhenyao Zhu, Jesse H Engel, and Awni Y Hannun · 1902
Earlier work this paper cites.
The phase vocoder: A tutorial
Mark Dolson · 1986
Earlier work this paper cites.
Estimating and interpreting the instantaneous frequency of a signal
Boualem Boashash · 1992
Earlier work this paper cites.
End-to-end learning for music audio
Sander Dieleman and Benjamin Schrauwen · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Earlier work this paper cites.
Fast WaveNet generation algorithm
Tom Le Paine, Pooya Khorrami, Shiyu Chang, Yang Zhang, Prajit Ramachandran, Mark A Hasegawa-Johnson, and Thomas S Huang · 2016
Earlier work this paper cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, and Soumith Chintala · 2016
Earlier work this paper cites.
Improved techniques for training gans
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Earlier work this paper cites.
A note on the evaluation of generative models
Lucas Theis, Aäron van den Oord, and Matthias Bethge · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W Senior, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Earlier work this paper cites.
BEGAN: Boundary equilibrium generative adversarial networks
David Berthelot, Thomas Schumm, and Luke Metz · 2017
Cited alongside, same era.
Neural audio synthesis of musical notes with WaveNet autoencoders
Jesse Engel, Cinjon Resnick, Adam Roberts, Sander Dieleman, Douglas Eck, Karen Simonyan, and Mohammad Norouzi · 2017
Cited alongside, same era.
Improved training of Wasserstein GANs
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron C Courville · 2017
Cited alongside, same era.
GANs trained by a two time-scale update rule converge to a local Nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Cited alongside, same era.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2017
Cited alongside, same era.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros · 2017
Later among the works it cites.
Music generation and transformation with moment matching-scattering inverse networks
Mathieu Andreux and Stephane Mallat · 2018
Later among the works it cites.
Sing: Symbol-to-instrument neural generator
Alexandre Defossez, Neil Zeghidour, Nicolas Usunier, Leon Bottou, and Francis Bach · 2018
Later among the works it cites.
Bridging audio analysis, perception and synthesis with perceptually-regularized variational timbre spaces
Philippe Esling, Axel Chemla-Romeu-Santos, and Adrien Bitton · 2018
Later among the works it cites.
A multi-discriminator cyclegan for unsupervised non-parallel speech domain adaptation
Ehsan Hosseini-Asl, Yingbo Zhou, Caiming Xiong, and Richard Socher · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yanghua Jin, Jiakai Zhang, Minjun Li, Yingtao Tian, Huachun Zhu, and Zhihao Fang · 2017
Cited alongside, same era.
On convergence and stability of gans
Naveen Kodali, Jacob Abernethy, James Hays, and Zsolt Kira · 2017
Cited alongside, same era.
SampleRNN: An unconditional end-to-end neural audio generation model
Soroush Mehri, Kundan Kumar, Ishaan Gulrajani, Rithesh Kumar, Shubham Jain, Jose Sotelo, Aaron Courville, and Yoshua Bengio · 2017
Cited alongside, same era.
Conditional image synthesis with auxiliary classifier gans
Augustus Odena, Christopher Olah, and Jonathon Shlens · 2017
Cited alongside, same era.
Tim Salimans, Andrej Karpathy, Xi Chen, and Diederik P. Kingma · 2017
Cited alongside, same era.
Char2wav: End-to-end speech synthesis
Jose Sotelo, Soroush Mehri, Kundan Kumar, Joao Felipe Santos, Kyle Kastner, Aaron Courville, and Yoshua Bengio · 2017
Cited alongside, same era.
Tacotron: Towards end-to-end speech synthesis
Yuxuan Wang, RJ Skerry-Ryan, Daisy Stanton, Yonghui Wu, Ron J Weiss, Navdeep Jaitly, Zongheng Yang, Ying Xiao, Zhifeng Chen, Samy Bengio, et al · 2017
Cited alongside, same era.
Spectral normalization for generative adversarial networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida · 2018
Later among the works it cites.
Eitan Richardson and Yair Weiss · 2018
Later among the works it cites.
Parallel WaveNet: Fast high-fidelity speech synthesis
Aaron van den Oord, Yazhe Li, Igor Babuschkin, Karen Simonyan, Oriol Vinyals, Koray Kavukcuoglu, George van den Driessche, Edward Lockhart, Luis Cobo, Florian Stimberg, Norman Casagrande, Dominik Grewe, Seb Noury, Sander Dieleman, Erich Elsen, Nal Kalchbrenner, Heiga Zen, Alex Graves, Helen King, Tom Walters, Dan Belov, and Demis Hassabis · 2018
Later among the works it cites.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2019
Closest in time.
Adversarial audio synthesis
Chris Donahue, Julian McAuley, and Miller Puckette · 2019
Closest in time.
Augmented cyclic adversarial learning for low resource domain adaptation
Ehsan Hosseini-Asl, Yingbo Zhou, Caiming Xiong, and Richard Socher · 2019
Closest in time.
Autoencoder-based music translation
Noam Mor, Lior Wolf, Adam Polyak, and Yaniv Taigman · 2019
Closest in time.