Fetching the paper…
Reading the bibliography…
Voice impersonation is not the same as voice transformation, although the latter is an essential element of it.
“Signal estimation from modified short-time fourier transform,”
Daniel Griffin and Jae Lim, · 1984
Earlier work this paper cites.
“Voice conversion through vector quantization,”
M. Abe, S. Nakamura, K. Shikano, and H. Kuwabara, · 1990
Earlier work this paper cites.
“Voice transformation using PSOLA technique,”
Hélene Valbret, Eric Moulines, and Jean-Pierre Tubach, · 1992
Earlier work this paper cites.
“TIDIGITS,” \url
Linguistic Data Consortium, · 1993
Earlier work this paper cites.
“Assessment of objective quality measures for speech intelligibility estimation,”
Wei Ming Liu, Keith A Jellyman, John SD Mason, and Nicholas WD Evans, · 2006
Earlier work this paper cites.
“Voice conversion based on maximum-likelihood estimation of spectral parameter trajectory,”
Tomoki Toda, Alan W Black, and Keiichi Tokuda, · 2007
Earlier work this paper cites.
“Robust signal-to-noise ratio estimation based on waveform amplitude distribution analysis,”
Chanwoo Kim and Richard M Stern, · 2008
Earlier work this paper cites.
“Voice conversion by mapping the speaker-specific features using pitch synchronous approach,”
K. Sreenivasa Rao, · 2010
Cited alongside, same era.
“Parametric voice conversion based on bilinear frequency warping plus amplitude scaling,”
Daniel Erro, Eva Navas, and Inma Hernaez, · 2013
Cited alongside, same era.
“Voice conversion using deep neural networks with layer-wise generative training,”
Ling-Hui Chen, Zhen-Hua Ling, Li-Juan Liu, and Li-Rong Dai, · 2014
Cited alongside, same era.
“Generative adversarial nets,”
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio, · 2014
Cited alongside, same era.
“Unsupervised representation learning with deep convolutional generative adversarial networks,”
Alec Radford, Luke Metz, and Soumith Chintala, · 2015
Cited alongside, same era.
“Empirical evaluation of rectified activations in convolutional network,”
Bing Xu, Naiyan Wang, Tianqi Chen, and Mu Li, · 2015
Later among the works it cites.
“Image style transfer using convolutional neural networks,”
Leon A Gatys, Alexander S Ecker, and Matthias Bethge, · 2016
Later among the works it cites.
“Learning to discover cross-domain relations with generative adversarial networks,”
Taeksoo Kim, Moonsu Cha, Hyunsoo Kim, Jungkwon Lee, and Jiwon Kim, · 2017
Later among the works it cites.
“Unpaired image-to-image translation using cycle-consistent adversarial networks,”
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros, · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Batch normalization: Accelerating deep network training by reducing internal covariate shift,”
Sergey Ioffe and Christian Szegedy, · 2015
Cited alongside, same era.
“Audio examples,” \url
Yolanda Gao,
Cited in the paper.
Ahmed Elgammal, Bingchen Liu, Mohamed Elhoseiny, and Marian Mazzone, · 2017
Later among the works it cites.
“Unsupervised image-to-image translation networks,”
Ming-Yu Liu, Thomas Breuel, and Jan Kautz, · 2017
Later among the works it cites.