Fetching the paper…
Reading the bibliography…
Singing voice conversion is to convert a singer's voice to another one's voice without changing singing content.
“An hmm-based singing voice synthesis system,”
Keijiro Saino, Heiga Zen, Yoshihiko Nankaku, Akinobu Lee, and Keiichi Tokuda, · 2006
Earlier work this paper cites.
“Applying voice conversion to concatenative singing-voice synthesis,”
Fernando Villavicencio and Jordi Bonada, · 2010
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
“Exemplar-based voice conversion using non-negative spectrogram deconvolution,”
Zhizheng Wu, Tuomas Virtanen, Tomi Kinnunen, Eng Siong Chng, and Haizhou Li, · 2013
Earlier work this paper cites.
“Voice conversion in high-order eigen space using deep belief nets.,”
Toru Nakashika, Ryoichi Takashima, Tetsuya Takiguchi, and Yasuo Ariki, · 2013
Earlier work this paper cites.
“The nus sung and spoken lyrics corpus: A quantitative comparison of singing and speech,”
Zhiyan Duan, Haotian Fang, Bo Li, Khe Chai Sim, and Ye Wang, · 2013
Earlier work this paper cites.
“Statistical singing voice conversion with direct waveform modification based on the spectrum differential,”
Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, and Satoshi Nakamura, · 2014
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
Diederik P Kingma and Jimmy Ba, · 2014
Cited alongside, same era.
“Video-audio driven real-time facial animation,”
Yilong Liu, Feng Xu, Jinxiang Chai, Xin Tong, Lijuan Wang, and Qiang Huo, · 2015
Cited alongside, same era.
“Statistical singing voice conversion based on direct waveform modification with global variance,”
Kazuhiro Kobayashi, Tomoki Toda, Graham Neubig, Sakriani Sakti, and Satoshi Nakamura, · 2015
Cited alongside, same era.
“librosa: Audio and music signal analysis in python,”
Brian McFee, Colin Raffel, Dawen Liang, Daniel PW Ellis, Matt McVicar, Eric Battenberg, and Oriol Nieto, · 2015
Cited alongside, same era.
“Expressive singing synthesis based on unit selection for the singing synthesis challenge 2016.,”
Jordi Bonada, Martí Umbert, and Merlijn Blaauw, · 2016
Cited alongside, same era.
“Phonetic posteriorgrams for many-to-one voice conversion without parallel data training,”
Lifa Sun, Kun Li, Hao Wang, Shiyin Kang, and Helen Meng, · 2016
Later among the works it cites.
“Domain-adversarial training of neural networks,”
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky, · 2016
Later among the works it cites.
“A neural parametric singing synthesizer modeling timbre and expression from natural songs,”
Merlijn Blaauw and Jordi Bonada, · 2017
Later among the works it cites.
“Automatic differentiation in PyTorch,”
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer, · 2017
Later among the works it cites.
“A universal music translation network,”
Noam Mor, Lior Wolf, Adam Polyak, and Yaniv Taigman, · 2018
Later among the works it cites.
“Unsupervised singing voice conversion,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Wavenet: A generative model for raw audio,”
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu, · 2016
Cited alongside, same era.
Eliya Nachmani and Lior Wolf, · 2019
Closest in time.