Fetching the paper…
Reading the bibliography…
We present a voice conversion solution using recurrent sequence to sequence modeling for DNNs.
E. Nachmani and L. Wolf, “Unsupervised polyglot text to speech,”
1902
Earlier work this paper cites.
R. J. Williams and D. Zipser, “A learning algorithm for continually running fully recurrent neural networks,”
1989
Earlier work this paper cites.
A. Kain and M. Macon, “Spectral voice conversion for text-to-speech synthesis,” in
1998
Earlier work this paper cites.
J. Kominek and A. W. Black, “Cmu arctic databases for speechsynthesis,” Language Technology Institute, Carnegie Mellon University, Pittsburgh, PA, 2003. [Online]. Available: http://festvox.org/cmu arctic/index.html
2003
Earlier work this paper cites.
T. Toda, A. W. Black, and K. Tokuda, “Voice conversion based on maximum-likelihood estimation of spectral parameter trajectory,”
2007
Earlier work this paper cites.
M. Müller,
2007
Earlier work this paper cites.
S. Desai, E. V. Raghavendra, B. Yegnanarayana, A. W. Black, and K. Prahallad, “Voice conversion using artificial neural networks,” in
2009
Earlier work this paper cites.
S. Desai, A. W. Black, and B. Yegnanarayana, “Voice conversion using artificial neural networks,” in
2010
Earlier work this paper cites.
R. Takashima, T. Takiguchi, and Y. Ariki, “Exemplar-based voice conversion using sparse representation in noisy environments,”
2013
Earlier work this paper cites.
D. Kingma and M. Welling, “Autoencoding variational bayes,”
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,”
2014
Earlier work this paper cites.
L. Sun, S. Yang, K. Li, and H. Meng, “Voice conversion using deep bidirectional long short-term memory,” in
2015
Earlier work this paper cites.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell,”
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
D. J. Rezende and S. Mohamed, “Variational normalizing flows,”
2015
Cited alongside, same era.
L. Sun, K. Li, S. Kang, and H. Meng, in
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
2017
Later among the works it cites.
M. Arjrovsky, S. Chintala, and L. Bottou, “Wasserstein gan,”
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
K. Ito, “The lj speech dataset,” 2017. [Online]. Available: https://keithito.com/LJ-Speech-Dataset/
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
S. Ö. Arik, J. Chen, K. Peng, W. Ping, and Y. Zhou, “Neural voice cloning with a few samples,”
2018
Later among the works it cites.
2018
Later among the works it cites.
R. Yamamoto, “Wavenet vocoder,” 2018. [Online]. Available: https://github.com/r9y9/wavenet_vocoder
2018
Later among the works it cites.
2018
Later among the works it cites.
Y. Li, M. Sun, H. Van Hamme, X. Zhang, and J. Yang, “Robust hierarchical learning for non-negative matrix factorization with outliers,”
2019
Closest in time.