Fetching the paper…
Reading the bibliography…
Traditional voice conversion methods rely on parallel recordings of multiple speakers pronouncing the same sentences.
Signal estimation from modified short-time fourier transform
Daniel W. Griffin and Jae S. Lim · 1983
Earlier work this paper cites.
Spectral voice conversion for text-to-speech synthesis
A. Kain and M. W. Macon · 1998
Earlier work this paper cites.
Continuous probabilistic transform for voice conversion
Yannis Stylianou, Olivier Cappé, and Eric Moulines · 1998
Earlier work this paper cites.
Improving the intelligibility of dysarthric speech
Alexander B Kain, John-Paul Hosom, Xiaochuan Niu, Jan PH Van Santen, Melanie Fried-Oken, and Janice Staehely · 2007
Earlier work this paper cites.
Voice conversion based on maximum-likelihood estimation of spectral parameter trajectory
Tomoki Toda, Alan W Black, and Keiichi Tokuda · 2007
Earlier work this paper cites.
Text-independent voice conversion based on state mapped codebook
and Meng Zhang, Jianhua Tao, Jilei Tian, and Xia Wang · 2008
Earlier work this paper cites.
Data-driven emotion conversion in spoken english
Zeynep Inanoglu and Steve Young · 2009
Earlier work this paper cites.
Evaluation of expressive speech synthesis with voice conversion and copy resynthesis techniques
Oytun Turk and Marc Schroder · 2010
Earlier work this paper cites.
Voice conversion using partial least squares regression
Elina Helander, Tuomas Virtanen, Jani Nurminen, and Moncef Gabbouj · 2010
Earlier work this paper cites.
Spectral mapping using artificial neural networks for voice conversion
S. Desai, A. W. Black, B. Yegnanarayana, and K. Prahallad · 2010
Earlier work this paper cites.
Speaking-aid systems using gmm-based voice conversion for electrolaryngeal speech
Keigo Nakamura, Tomoki Toda, Hiroshi Saruwatari, and Kiyohiro Shikano · 2012
Earlier work this paper cites.
Statistical voice conversion techniques for body-conducted unvoiced speech enhancement
Tomoki Toda, Mikihiro Nakagiri, and Kiyohiro Shikano · 2012
Earlier work this paper cites.
Voice conversion using deep neural networks with layer-wise generative training
Ling-Hui Chen, Zhen-Hua Ling, Li-Juan Liu, and Li-Rong Dai · 2014
Earlier work this paper cites.
Voice conversion based on speaker-dependent restricted boltzmann machines
Toru Nakashika, Tetsuya Takiguchi, and Yasuo Ariki · 2014
Cited alongside, same era.
Voice conversion using deep neural networks with speaker-independent pre-training
S. H. Mohammadi and A. Kain · 2014
Cited alongside, same era.
Generative adversarial networks
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Voice conversion using deep bidirectional long short-term memory based recurrent neural networks
L. Sun, S. Kang, K. Li, and H. Meng · 2015
Cited alongside, same era.
Mapping frames with dnn-HMM recognizer for non-parallel voice conversion
M. Dong, C. Yang, Y. Lu, J. W. Ehnes, D. Huang, H. Ming, R. Tong, S. W. Lee, and H. Li · 2015
Cited alongside, same era.
Siamese network of deep fisher-vector descriptors for image retrieval
Eng-Jon Ong, Sameed Husain, and Miroslaw Bober · 2017
Later among the works it cites.
Unsupervised cross-domain image generation
Yaniv Taigman, Adam Polyak, and Lior Wolf · 2017
Later among the works it cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Later among the works it cites.
Cyclegan-vc: Non-parallel voice conversion using cycle-consistent adversarial networks
T. Kaneko and H. Kameoka · 2018
Later among the works it cites.
Stargan-vc: non-parallel many-to-many voice conversion using star generative adversarial networks
H. Kameoka, T. Kaneko, K. Tanaka, and N. Hojo · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, and Soumith Chintala · 2016
Cited alongside, same era.
Siamese network features for image matching
I. Melekhov, J. Kannala, and E. Rahtu · 2016
Cited alongside, same era.
Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network
Wenzhe Shi, Jose Caballero, Ferenc Huszar, Johannes Totz, Andrew P. Aitken, Rob Bishop, Daniel Rueckert, and Zehan Wang · 2016
Cited alongside, same era.
Sequence-to-sequence voice conversion with similarity metric learned using generative adversarial networks
Takuhiro Kaneko, Hirokazu Kameoka, Kaoru Hiramatsu, and Kunio Kashino · 2017
Cited alongside, same era.
Image-to-image translation with conditional adversarial networks
P. Isola, J. Zhu, T. Zhou, and A. A. Efros · 2017
Cited alongside, same era.
Unpaired image-to-image translation using cycle-consistent adversarial networks
J. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Cited alongside, same era.
Tero Karras, Samuli Laine, and Timo Aila · 2018
Later among the works it cites.
Chris Donahue, Julian McAuley, and Miller Puckette · 2018
Later among the works it cites.
Self-attention generative adversarial networks
Han Zhang, Ian Goodfellow, Dimitris Metaxas, and Augustus Odena · 2018
Later among the works it cites.
Spectral normalization for generative adversarial networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida · 2018
Later among the works it cites.
Cyclegan-vc2: Improved cyclegan-based non-parallel voice conversion
T. Kaneko, H. Kameoka, K. Tanaka, and N. Hojo · 2019
Closest in time.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2019
Closest in time.
Few-shot unsupervised image-to-image translation
Ming-Yu Liu, Xun Huang, Arun Mallya, Tero Karras, Timo Aila, Jaakko Lehtinen, and Jan Kautz · 2019
Closest in time.
Travelgan: Image-to-image translation by transformation vector learning
Matthew Amodio and Smita Krishnaswamy · 2019
Closest in time.