Fetching the paper…
Reading the bibliography…
Thanks to developments in machine learning techniques, it has become possible to synthesize high-quality singing voices of a single singer.
“Singing voice synthesis based on deep neural networks,”
M. Nishimura, K. Hashimoto, K. Oura, Y. Nankaku, and K. Tokuda, · 2016
Earlier work this paper cites.
“A neural parametric singing synthesizer modeling timbre and expression from natural songs,”
M. Blaauw and J. Bonada, · 2017
Earlier work this paper cites.
“JVS corpus: free Japanese multi-speaker voice corpus,”
S. Takamichi, K. Mitsui, Y. Saito, T. Koriyama, N. Tanji, and H. Saruwatari, · 2019
Cited alongside, same era.
“HMM-based speech synthesis system (HTS),” http://hts.sp.nitech.ac.jp/
Cited in the paper.
“JSUT-song,”
Cited in the paper.
“Kiritan singing database,”
Cited in the paper.
“Celemony — what is melodyne?,” https://www.celemony.com/en/melodyne/what-is-melodyne
Cited in the paper.
“Lancers,”
Cited in the paper.
“DNN-based speaker embedding using subjective inter-speaker similarity for multi-speaker modeling in speech synthesis,”
Y. Saito, S. Takamichi, and H. Saruwatari, · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…