Fetching the paper…
Reading the bibliography…
In this work, we address the problem of musical timbre transfer, where the goal is to manipulate the timbre of a sound sample from one instrument to match another instrument while preserving other musical content, such as pitch, rhythm, and loudness.
The synthesis of complex audio spectra by means of frequency modulation
J. M. Chowning · 1973
Earlier work this paper cites.
A unified approach to short-time Fourier analysis and synthesis
Jont B Allen and Lawrence R Rabiner · 1977
Earlier work this paper cites.
Hearing musical streams
Stephen McAdams and Albert Bregman · 1979
Earlier work this paper cites.
Signal estimation from modified short-time Fourier transform
Daniel Griffin and Jae Lim · 1984
Earlier work this paper cites.
Calculation of a constant Q spectral transform
Judith C Brown · 1991
Earlier work this paper cites.
Exploration of timbre by analysis and synthesis
Jean-Claude Risset and David Wessel · 1999
Earlier work this paper cites.
Separating style and content with bilinear models
Joshua B Tenenbaum and William T Freeman · 1999
Earlier work this paper cites.
Towards an inverse constant Q transform
Derry Fitzgerald, Matt Cranitch, and Marcin T Cychowski · 2006
Earlier work this paper cites.
The physics and psychophysics of music: an introduction
Juan G Roederer · 2008
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros · 2008
Earlier work this paper cites.
Physical Audio Signal Processing: for Virtual Musical Instruments and Audio Effects
Julius O. III Smith · 2010
Earlier work this paper cites.
Spectral Audio Signal Processing
Julius O. III Smith · 2011
Earlier work this paper cites.
Constructing an invertible constant-Q transform with non-stationary gabor frames
Gino Angelo Velasco, Nicki Holighaus, Monika Dörfler, and Thomas Grill · 2011
Earlier work this paper cites.
A framework for invertible, real-time constant-Q transforms
Nicki Holighaus, Monika Dörfler, Gino Angelo Velasco, and Thomas Grill · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
A neural algorithm of artistic style
Leon A. Gatys, Alexander S. Ecker, and Matthias Bethge · 2015
Cited alongside, same era.
Perceptual losses for real-time style transfer and super-resolution
Justin Johnson, Alexandre Alahi, and Li Fei-Fei · 2016
Cited alongside, same era.
Samplernn: An unconditional end-to-end neural audio generation model
Soroush Mehri, Kundan Kumar, Ishaan Gulrajani, Rithesh Kumar, Shubham Jain, Jose Sotelo, Aaron C. Courville, and Yoshua Bengio · 2016
Cited alongside, same era.
World: a vocoder-based high-quality speech synthesis system for real-time applications
Masanori Morise, Fumiya Yokomori, and Kenji Ozawa · 2016
Cited alongside, same era.
Deconvolution and checkerboard artifacts
Augustus Odena, Vincent Dumoulin, and Chris Olah · 2016
Cited alongside, same era.
Parallel-data-free voice conversion using cycle-consistent adversarial networks
Takuhiro Kaneko and Hirokazu Kameoka · 2017
Later among the works it cites.
Learning to discover cross-domain relations with generative adversarial networks
Taeksoo Kim, Moonsu Cha, Hyunsoo Kim, Jungkwon Lee, and Jiwon Kim · 2017
Later among the works it cites.
Unsupervised image-to-image translation networks
Ming-Yu Liu, Thomas Breuel, and Jan Kautz · 2017
Later among the works it cites.
Natural TTS synthesis by conditioning wavenet on mel spectrogram predictions
Jonathan Shen, Ruoming Pang, Ron J Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, RJ Skerry-Ryan, et al · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Audio texture synthesis and style transfer
Dmitry Ulyanov and Vadim Lebedev · 2016
Cited alongside, same era.
Texture networks: Feed-forward synthesis of textures and stylized images
Dmitry Ulyanov, Vadim Lebedev, Andrea Vedaldi, and Victor S. Lempitsky · 2016
Cited alongside, same era.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W. Senior, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Deep voice: Real-time neural text-to-speech
Sercan O Arik, Mike Chrzanowski, Adam Coates, Gregory Diamos, Andrew Gibiansky, Yongguo Kang, Xian Li, John Miller, Jonathan Raiman, Shubho Sengupta, et al · 2017
Cited alongside, same era.
CycleGAN, a master of steganography
Casey Chu, Andrey Zhmoginov, and Mark Sandler · 2017
Cited alongside, same era.
Neural audio synthesis of musical notes with wavenet autoencoders
Jesse Engel, Cinjon Resnick, Adam Roberts, Sander Dieleman, Douglas Eck, Karen Simonyan, and Mohammad Norouzi · 2017
Cited alongside, same era.
Eric Grinstein, Ngoc Q. K. Duong, Alexey Ozerov, and Patrick Pérez · 2017
Cited alongside, same era.
Zili Yi, Hao Zhang, Ping Tan Gong, et al · 2017
Later among the works it cites.
Modulated variational auto-encoders for many-to-many musical timbre transfer
Adrien Bitton, Philippe Esling, and Axel Chemla-Romeu-Santos · 2018
Closest in time.
Symbolic music genre transfer with CycleGAN
Gino Brunner, Yuyi Wang, Roger Wattenhofer, and Sumu Zhao · 2018
Closest in time.
Music style transfer issues: A position paper
Shuqi Dai, Zheng Zhang, and Gus Xia · 2018
Closest in time.
Synthesizing audio with generative adversarial networks
Chris Donahue, Julian McAuley, and Miller Puckette · 2018
Closest in time.
Many paths to equilibrium: GANs do not need to decrease a divergence at every step
William Fedus, Mihaela Rosca, Balaji Lakshminarayanan, Andrew M Dai, Shakir Mohamed, and Ian Goodfellow · 2018
Closest in time.
URL https://www.vsl.co.at/en/Products
Vienna Symphonic Library GmbH, 2018 · 2018
Closest in time.
Unsupervised cipher cracking using discrete GANs
Aidan N Gomez, Sicong Huang, Ivan Zhang, Bryan M Li, Muhammad Osama, and Lukasz Kaiser · 2018
Closest in time.
Chapter one: An acoustics primer, 2018
Jeffrey Hass · 2018
Closest in time.
A universal music translation network
Noam Mor, Lior Wolf, Adam Polyak, and Yaniv Taigman · 2018
Closest in time.
Neural style transfer for audio spectograms
Prateek Verma and Julius O Smith · 2018
Closest in time.