Fetching the paper…
Reading the bibliography…
Generative models in vision have seen rapid progress due to algorithmic improvements and the availability of high-quality image datasets.
The synthesis of complex audio spectra by means of frequency modulation
Chowning, John M · 1973
Earlier work this paper cites.
Signal estimation from modified short-time fourier transform
Griffin, Daniel and Lim, Jae · 1984
Earlier work this paper cites.
Calculation of a constant q spectral transform
Brown, Judith C · 1991
Earlier work this paper cites.
Estimating and interpreting the instantaneous frequency of a signal. i. fundamentals
Boashash, Boualem · 1992
Earlier work this paper cites.
The mnist database of handwritten digits, 1998
LeCun, Yann, Cortes, Corinna, and Burges, Christopher JC · 1998
Earlier work this paper cites.
Rwc music database: Music genre database and musical instrument sound database
Goto, Masataka, Hashiguchi, Hiroki, Nishimura, Takuichi, and Oka, Ryuichi · 2003
Earlier work this paper cites.
Pattern Recognition and Machine Learning (Information Science and Statistics)
Bishop, Christopher M · 2006
Earlier work this paper cites.
Xenakis, Iannis · 2006
Earlier work this paper cites.
The blizzard challenge 2008
King, Simon, Clark, Robert AJ, Mayo, Catherine, and Karaiskos, Vasilis · 2008
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Earlier work this paper cites.
Sound texture synthesis via filter statistics
McDermott, Josh H, Oxenham, Andrew J, and Simoncelli, Eero P · 2009
Cited alongside, same era.
Analog days: The invention and impact of the Moog synthesizer
Pinch, Trevor J, Trocco, Frank, and Pinch, TJ · 2009
Cited alongside, same era.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
Vincent, Pascal, Larochelle, Hugo, Lajoie, Isabelle, Bengio, Yoshua, and Manzagol, Pierre-Antoine · 2010
Cited alongside, same era.
The million song dataset
Bertin-Mahieux, Thierry, Ellis, Daniel PW, Whitman, Brian, and Lamere, Paul · 2011
Cited alongside, same era.
Reading digits in natural images with unsupervised feature learning
Netzer, Yuval, Wang, Tao, Coates, Adam, Bissacco, Alessandro, Wu, Bo, and Ng, Andrew Y · 2011
Cited alongside, same era.
A real-time system for measuring sound goodness in instrumental sounds
Romani Picas, Oriol, Parra Rodriguez, Hector, Dabiri, Dara, Tokuda, Hiroshi, Hariya, Wataru, Oishi, Koji, and Serra, Xavier · 2015
Later among the works it cites.
A note on the evaluation of generative models
Theis, Lucas, Oord, Aäron van den, and Bethge, Matthias · 2015
Later among the works it cites.
Chen, Xi, Kingma, Diederik P., Salimans, Tim, Duan, Yan, Dhariwal, Prafulla, Schulman, John, Sutskever, Ilya, and Abbeel, Pieter · 2016
Later among the works it cites.
Pixelvae: A latent variable model for natural images
Gulrajani, Ishaan, Kumar, Kundan, Ahmed, Faruk, Taiga, Adrien Ali, Visin, Francesco, Vázquez, David, and Courville, Aaron C · 2016
Later among the works it cites.
Minst, a collection of musical sound datasets, 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kingma, Diederik P and Welling, Max · 2013
Cited alongside, same era.
Generative adversarial nets
Goodfellow, Ian, Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron, and Bengio, Yoshua · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik P. and Ba, Jimmy · 2014
Cited alongside, same era.
Giorgio by morodor, 2014
Punk, Daft · 2014
Cited alongside, same era.
Musical audio synthesis using autoencoding neural nets
Sarroff, Andy M and Casey, Michael A · 2014
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Cited alongside, same era.
Wavenet: A generative model for raw audio
van den Oord, Aäron, Dieleman, Sander, Zen, Heiga, Simonyan, Karen, Vinyals, Oriol, Graves, Alex, Kalchbrenner, Nal, Senior, Andrew W., and Kavukcuoglu, Koray
Cited in the paper.
Humphrey, Eric J · 2016
Later among the works it cites.
Samplernn: An unconditional end-to-end neural audio generation model
Mehri, Soroush, Kumar, Kundan, Gulrajani, Ishaan, Kumar, Rithesh, Jain, Shubham, Sotelo, Jose, Courville, Aaron C., and Bengio, Yoshua · 2016
Later among the works it cites.
Learning-Based Methods for Comparing Sequences, with Applications to Audio-to-MIDI Alignment and Matching
Raffel, Colin · 2016
Later among the works it cites.
Improved techniques for training gans
Salimans, Tim, Goodfellow, Ian J., Zaremba, Wojciech, Cheung, Vicki, Radford, Alec, and Chen, Xi · 2016
Later among the works it cites.
Learning features of music from scratch
Thickstun, John, Harchaoui, Zaid, and Kakade, Sham · 2016
Later among the works it cites.
Sampling generative networks: Notes on a few effective techniques
White, Tom · 2016
Later among the works it cites.