Fetching the paper…
Reading the bibliography…
This paper presents a method of using autoregressive neural networks for the acoustic modeling of singing voice synthesis (SVS).
N. Ueda and R. Nakano, “Deterministic annealing em algorithm,”
1998
Earlier work this paper cites.
H. Kawahara, I. Masuda-Katsuse, and A. De Cheveigne, “Restructuring speech representations using a pitch-adaptive time–frequency smoothing and an instantaneous-frequency-based f0 extraction: Possible role of a repetitive structure in sounds,”
1999
Earlier work this paper cites.
M. Schuster, “Better generative models for sequential data problems: Bidirectional recurrent mixture density networks,” in
2000
Earlier work this paper cites.
T. Saitou, M. Unoki, and M. Akagi, “Extraction of F0 dynamic characteristics and development of F0 control model in singing voice.” Georgia Institute of Technology, 2002
2002
Earlier work this paper cites.
F. Morin and Y. Bengio, “Hierarchical probabilistic neural network language model.” in
2005
Earlier work this paper cites.
T. Saitou, M. Unoki, and M. Akagi, “Development of an F0 control model based on F0 dynamic characteristics for singing-voice synthesis,”
2005
Earlier work this paper cites.
K. Saino, H. Zen, Y. Nankaku, A. Lee, and K. Tokuda, “An HMM-based singing voice synthesis system,” in
2006
Earlier work this paper cites.
H. Zen and A. Senior, “Deep mixture density networks for acoustic modeling in statistical parametric speech synthesis,” in
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,”
2014
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
H. Zen and H. Sak, “Unidirectional long short-term memory recurrent neural network with recurrent output layer for low-latency speech synthesis,” in
2015
Cited alongside, same era.
2015
Cited alongside, same era.
H.-Y. Gu and J.-K. He, “Singing-voice synthesis using demi-syllable unit selection,” in
2016
Cited alongside, same era.
2017
Later among the works it cites.
M. Blaauw and J. Bonada, “A neural parametric singing synthesizer,”
2017
Later among the works it cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Later among the works it cites.
J. Kim, H. Choi, J. Park, S. Kim, J. Kim, and M. Hahn, “Korean singing voice synthesis system based on an LSTM recurrent neural network,” in
2018
Later among the works it cites.
X. Wang, S. Takaki, and J. Yamagishi, “Autoregressive neural F0 model for statistical parametric speech synthesis,”
2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Bonada, M. Umbert, and M. Blaauw, “Expressive singing synthesis based on unit selection for the singing synthesis challenge 2016.” in
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Nishimura, K. Hashimoto, K. Oura, Y. Nankaku, and K. Tokuda, “Singing voice synthesis based on deep neural networks.” in
2016
Cited alongside, same era.
Later among the works it cites.
2018
Later among the works it cites.
Y. Ai, J.-X. Zhang, L. Chen, and Z.-H. Ling, “DNN-based spectral enhancement for neural waveform generators with low-bit quantization,” in
2019
Closest in time.