Fetching the paper…
Reading the bibliography…
Synthetic creation of drum sounds (e.g., in drum machines) is commonly performed using analog or digital synthesis, allowing a musician to sculpt the desired timbre modifying various parameters.
R. Gordon, “Synthesizing drums: The snare drum,” Sound On Sound , Jan. 2002. [Online]. Available: https://www.soundonsound.com/techniques/synthesizing-drums-snare-drum
2002
Earlier work this paper cites.
——, “Synthesizing drums: The bass drum,” Sound On Sound , Jan. 2002. [Online]. Available: https://www.soundonsound.com/techniques/synthesizing-drums-bass-drum
2002
Earlier work this paper cites.
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: A Large-Scale Hierarchical Image Database,” in IEEE Conference on Computer Vision and Pattern Recognition, CVPR , June 2009
2009
Earlier work this paper cites.
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. J. Smola, “A kernel two-sample test,” Journal of Machine Learning Research , vol. 13, pp. 723–773, 2012
2012
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in Proc. of the 2nd International Conference on Learning Representations, ICLR , Banff, AB, Canada, Apr. 2014
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. C. Courville, and Y. Bengio, “Generative adversarial nets,” in Proc. of the 28th International Conference on Neural Information Processing Systems NIPS , Montreal, Quebec, Canada, Dec. 2014
2014
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in IEEE International Conference on Computer Vision, ICCV , Santiago, Chile, Dec. 2015
2015
Earlier work this paper cites.
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. W. Senior, and K. Kavukcuoglu, “Wavenet: A generative model for raw audio,” in Proc. of the 9th ISCA Speech Synthesis Workshop , Sunnyvale, CA, USA, Sept. 2016
2016
Earlier work this paper cites.
A. van den Oord, N. Kalchbrenner, and K. Kavukcuoglu, “Pixel recurrent neural networks,” in Proc. of the 33rd International Conference on Machine Learning, ICML , New York City, NY, USA, June 2016
2016
Earlier work this paper cites.
T. Salimans, I. J. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen, “Improved techniques for training gans,” in Proc. of the International Conference on Neural Information Processing Systems, NIPS , Barcelona, Spain, Dec. 2016
2016
Earlier work this paper cites.
L. Theis, A. van den Oord, and M. Bethge, “A note on the evaluation of generative models,” in Proc. of the 4th International Conference on Learning Representations, ICLR , San Juan, Puerto Rico, May 2016
2016
Earlier work this paper cites.
J. Engel, C. Resnick, A. Roberts, S. Dieleman, M. Norouzi, D. Eck, and K. Simonyan, “Neural audio synthesis of musical notes with wavenet autoencoders,” in Proc. of the 34th International Conference on Machine Learning, ICML , Sydney, NSW, Australia, Aug. 2017
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. C. Courville, “Improved training of wasserstein gans,” in Proc. of the International Conference on Neural Information Processing Systems, NIPS , Long Beach, CA, USA, Dec. 2017
2017
Cited alongside, same era.
D. Michelsanti and Z. Tan, “Conditional generative adversarial networks for speech enhancement and noise-robust speaker verification,” in Proc. of the 18th Annual Conference of the International Speech Communication Association, Interspeech , Stockholm, Sweden, Aug. 2017
2017
Cited alongside, same era.
A. Odena, C. Olah, and J. Shlens, “Conditional image synthesis with auxiliary classifier gans,” in Proc. of the 34th International Conference on Machine Learning, ICML , Sydney, NSW, Australia, Aug. 2017
2017
C. Donahue, J. McAuley, and M. Puckette, “Adversarial audio synthesis,” in Proc. of the 7th International Conference on Learning Representations, ICLR , May 2019
2019
Later among the works it cites.
C. Aouameur, P. Esling, and G. Hadjeres, “Neural drum machine: An interactive system for real-time synthesis of drum sounds,” in Proc. of the 10th International Conference on Computational Creativity, ICCC , June 2019
2019
Later among the works it cites.
J. Engel, K. K. Agrawal, S. Chen, I. Gulrajani, C. Donahue, and A. Roberts, “Gansynth: Adversarial neural audio synthesis,” in Proc. of the 7th International Conference on Learning Representations, ICLR , May 2019
2019
Later among the works it cites.
P. Esling, N. Masuda, A. Bardet, R. Despres, and A. Chemla-Romeu-Santos, “Universal audio synthesizer control with normalizing flows,” Journal of Applied Sciences , 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
T. Karras, T. Aila, S. Laine, and J. Lehtinen, “Progressive growing of gans for improved quality, stability, and variation,” in International Conference on Learning Representations, ICLR , May 2018
2018
Cited alongside, same era.
J. Shen, R. Pang, R. J. Weiss, M. Schuster, N. Jaitly, Z. Yang, Z. Chen, Y. Zhang, Y. Wang, R. Ryan, R. A. Saurous, Y. Agiomyrgiannakis, and Y. Wu, “Natural TTS synthesis by conditioning wavenet on MEL spectrogram predictions,” in IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP , Calgary, AB, Canada, Apr. 2018
2018
Cited alongside, same era.
P. Esling, A. Chemla-Romeu-Santos, and A. Bitton, “Bridging audio analysis, perception and synthesis with perceptually-regularized variational timbre spaces,” in Proc. of the 19th International Society for Music Information Retrieval Conference, ISMIR , Paris, France, Sept. 2018
2018
Cited alongside, same era.
——, “Generative timbre spaces with variational audio synthesis,” in Proc. of the 21st International Conference on Digital Audio Effects DAFx-18 , Aveiro, Portugal, Sept. 2018
2018
Cited alongside, same era.
Y. Saito, S. Takamichi, and H. Saruwatari, “Statistical parametric speech synthesis incorporating generative adversarial networks,” IEEE/ACM Trans. Audio Speech Lang. Process. , vol. 26, pp. 84–96, 2018
2018
Cited alongside, same era.
E. Hosseini-Asl, Y. Zhou, C. Xiong, and R. Socher, “A multi-discriminator cyclegan for unsupervised non-parallel speech domain adaptation,” in Proc. of the 19th Annual Conference of the International Speech Communication Association , Hyderabad, India, Sept. 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
M. Binkowski, D. J. Sutherland, M. Arbel, and A. Gretton, “Demystifying MMD gans,” in Proc. of the 6th International Conference on Learning Representations, ICLR , Vancouver, BC, Canada, Apr. 2018
2018
Cited alongside, same era.
S. Lattner and M. Grachten, “High-level control of drum track generation using learned patterns of rhythmic interaction,” in IEEE Workshop on Applications of Signal Processing to Audio and Acoustics, WASPAA , New Paltz, NY, USA, Oct. 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
S. Huang, Q. Li, C. Anil, X. Bao, S. Oore, and R. B. Grosse, “Timbretron: A wavenet(cyclegan(cqt(audio))) pipeline for musical timbre transfer,” in Proc. of the 7th International Conference on Learning Representations, ICLR , New Orleans, LA, USA, May 2019
2019
Later among the works it cites.
K. Kumar, R. Kumar, T. de Boissiere, L. Gestin, W. Z. Teoh, J. Sotelo, A. de Brébisson, Y. Bengio, and A. C. Courville, “Melgan: Generative adversarial networks for conditional waveform synthesis,” in Proc. of the Annual Conference on Neural Information Processing Systems, NIPS , Vancouver, BC, Canada, Dec. 2019
2019
Later among the works it cites.
A. Ramires, P. Chandna, X. Favory, E. Gómez, and X. Serra, “Neural percussive synthesis parameterised by high-level timbral features,” in IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP , May 2020
2020
Closest in time.
2020
Closest in time.
J. H. Engel, L. Hantrakul, C. Gu, and A. Roberts, “DDSP: differentiable digital signal processing,” in Proc. of the 8th International Conference on Learning Representations, ICLR , Addis Ababa, Ethiopia, Apr. 2020
2020
Closest in time.
J. Nistal, S. Lattner, and G. Richard, “Comparing representations for audio synthesis using generative adversarial networks,” in Proc. of the 28th European Signal Processing Conference, EUSIPCO2020 , Amsterdam, NL, Jan. 2021
2021
Closest in time.