Fetching the paper…
Reading the bibliography…
This paper presents XiaoiceSing, a high-quality singing voice synthesis system which employs an integrated network for spectrum, F0 and duration modeling.
M. Macon, L. Jensen-Link, E. B. George, J. Oliverio, and M. Clements, “Concatenation-based midi-to-singing voice synthesis,” in
1997
Earlier work this paper cites.
K. Sjölander, “An hmm-based system for automatic segmentation and alignment of speech,” in
2003
Earlier work this paper cites.
P. Taylor, “Hidden markov models for grapheme to phoneme conversion,” in
2005
Earlier work this paper cites.
K. Saino, H. Zen, Y. Nankaku, A. Lee, and K. Tokuda, “An hmm-based singing voice synthesis system,” in
2006
Earlier work this paper cites.
H. Kenmochi and H. Ohshita, “Vocaloid-commercial singing synthesizer based on sample concatenation,” in
2007
Earlier work this paper cites.
D. Deutsch,
2013
Earlier work this paper cites.
K. Nakamura, K. Oura, Y. Nankaku, and K. Tokuda, “Hmm-based singing voice synthesis and its application to japanese and english,” in
2014
Earlier work this paper cites.
H.-Y. Gu and J.-K. He, “Singing-voice synthesis using demi-syllable unit selection,” in
2016
Cited alongside, same era.
J. Bonada, M. Umbert, and M. Blaauw, “Expressive singing synthesis based on unit selection for the singing synthesis challenge 2016.” in
2016
Cited alongside, same era.
M. Nishimura, K. Hashimoto, K. Oura, Y. Nankaku, and K. Tokuda, “Singing voice synthesis based on deep neural networks.” in
2016
Cited alongside, same era.
M. Morise, F. Yokomori, and K. Ozawa, “World: a vocoder-based high-quality speech synthesis system for real-time applications,”
2016
Cited alongside, same era.
M. Blaauw and J. Bonada, “A neural parametric singing synthesizer modeling timbre and expression from natural songs,”
2017
Cited alongside, same era.
Y. Hono, K. Hashimoto, K. Oura, Y. Nankaku, and K. Tokuda, “Singing voice synthesis based on generative adversarial networks,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
Y.-H. Yi, Y. Ai, Z.-H. Ling, and L.-R. Dai, “Singing voice synthesis using deep autoregressive neural networks for acoustic modeling,” in
2019
Later among the works it cites.
J. Lee, H.-S. Choi, C.-B. Jeon, J. Koo, and K. Lee, “Adversarially trained end-to-end korean singing voice synthesis system,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
J. Kim, H. Choi, J. Park, S. Kim, J. Kim, and M. Hahn, “Korean singing voice synthesis system based on an lstm recurrent neural network,” in
2018
Cited alongside, same era.
M. M. Association, “Midi manufacturers association,” https://www.midi.org
Cited in the paper.
2019
Later among the works it cites.
Blaauw, Merlijn and Bonada, Jordi, “Sequence-to-sequence singing synthesis using the feed-forward transformer,” in
2020
Closest in time.