Fetching the paper…
Reading the bibliography…
Voice conversion (VC) aims at conversion of speaker characteristic without altering content.
“Calculation of a constant q spectral transform,”
Judith Brown, · 1991
Earlier work this paper cites.
“An experimental study of speaker verification sensitivity to computer voice-altered imposters,”
Bryan L. Pellom and John H. L. Hansen, · 1999
Earlier work this paper cites.
“Voice conversion based on maximum-likelihood estimation of spectral parameter trajectory,”
Tomoki Toda, Alan W. Black, and Keiichi Tokuda, · 2007
Earlier work this paper cites.
“Voice transformation: A survey,”
Yannis Stylianou, · 2009
Earlier work this paper cites.
“Vulnerability of speaker verification systems against voice conversion spoofing attacks: The case of telephone speech,”
Tomi Kinnunen, Zhizheng Wu, Kong-Aik Lee, Filip Sedlak, Engsiong Chng, and Haizhou Li, · 2012
Earlier work this paper cites.
“The distribution of calibrated likelihood-ratios in speaker recognition,”
David A. van Leeuwen and Niko Brümmer, · 2013
Earlier work this paper cites.
“MSR identity toolbox: A MATLAB toolbox for speaker recognition research (v 1.0),”
Seyed Omid Sadjadi, Malcolm Slaney, and Larry Heck, · 2013
Earlier work this paper cites.
“A one-class classification approach to generalised speaker verification spoofing countermeasures using local binary patterns,”
Federico Alegre, Asmaa Amehraye, and Nicholas W. D. Evans, · 2013
Earlier work this paper cites.
“Spoofing and countermeasures for speaker verification: A survey,”
Zhizheng Wu, Nicholas W. D. Evans, Tomi Kinnunen, Junichi Yamagishi, Federico Alegre, and Haizhou Li, · 2015
Earlier work this paper cites.
“Can we automatically transform speech recorded on common consumer devices in real-world environments into professional production quality speech? - A dataset, insights, and challenges,”
Gautham J. Mysore, · 2015
Cited alongside, same era.
“The voice conversion challenge 2016,”
Tomoki Toda, Ling-Hui Chen, Daisuke Saito, Fernando Villavicencio, Mirjam Wester, Zhizheng Wu, and Junichi Yamagishi, · 2016
Cited alongside, same era.
“Overview of BTAS 2016 speaker anti-spoofing competition,”
Pavel Korshunov, Sébastien Marcel, Hannah Muckenhirn, Andre R. Goncalves, A. G. Souza Mello, Ricardo P. Velloso Violato, Flávio O. Simões, M. U. Neto, Marcus de Assis Angeloni, José Augusto Stuchi, Heinrich Dinkel, Nanxin Chen, Yanmin Qian, Dipjyoti Paul, Goutam Saha, and Md. Sahidullah, · 2016
Cited alongside, same era.
“A new feature for automatic speaker verification anti-spoofing: Constant q cepstral coefficients,”
Massimiliano Todisco, Héctor Delgado, and Nicholas Evans, · 2016
Cited alongside, same era.
“Spoofing detection goes noisy: An analysis of synthetic speech detection in the presence of additive noise,”
“Asvspoof: The automatic speaker verification spoofing and countermeasures challenge,”
Zhizheng Wu, Junichi Yamagishi, Tomi Kinnunen, Cemal Hanilçi, Md. Sahidullah, Aleksandr Sizov, Nicholas W. D. Evans, and Massimiliano Todisco, · 2017
Later among the works it cites.
“Front-end for antispoofing countermeasures in speaker verification: Scattering spectral decomposition,”
Kaavya Sriskandaraja, Vidhyasaharan Sethu, Eliathamby Ambikairajah, and Haizhou Li, · 2017
Later among the works it cites.
“Audio replay attack detection with deep learning frameworks,”
Galina Lavrentyeva, Sergey Novoselov, Egor Malykh, Alexander Kozlov, Oleg Kudashev, and Vadim Shchemelinin, · 2017
Later among the works it cites.
“Impact of score fusion on voice biometrics and presentation attack detection in cross-database evaluations,”
Pavel Korshunov and Sébastien Marcel, · 2017
Later among the works it cites.
“Impact of bandwidth and channel variation on presentation attack detection for speaker verification,”
Héctor Delgado, Massimiliano Todisco, Nicholas W. D. Evans, Md. Sahidullah, Wei Ming Liu, Federico Alegre, Tomi Kinnunen, and Benoit G. B. Fauve, · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cemal Hanilçi, Tomi Kinnunen, Md. Sahidullah, and Aleksandr Sizov, · 2016
Cited alongside, same era.
“WaveNet: A generative model for raw audio,”
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu, · 2016
Cited alongside, same era.
“An overview of voice conversion systems,”
Seyed Hamidreza Mohammadi and Alexander Kain, · 2017
Cited alongside, same era.
“Voice conversion using input-to-output highway networks,”
Yuki Saito, Shinnosuke Takamichi, and Hiroshi Saruwatari, · 2017
Cited alongside, same era.
“Statistical voice conversion with wavenet-based waveform generation,”
Kazuhiro Kobayashi, Tomoki Hayashi, Akira Tamamori, and Tomoki Toda, · 2017
Cited alongside, same era.
Later among the works it cites.
“The voice conversion challenge 2018: Promoting development of parallel and nonparallel methods,”
Jaime Lorenzo-Trueba, Junichi Yamagishi, Tomoki Toda, Daisuke Saito, Fernando Villavicencio, Tomi Kinnunen, and Zhenhua Ling., · 2018
Closest in time.
“Combining evidences from mel cepstral, cochlear filter cepstral and instantaneous frequency features for detection of natural vs. spoofed speech,”
Tanvina B. Patel and Hemant A. Patil, · 2066
Closest in time.
“A comparison of features for synthetic speech detection,”
Md. Sahidullah, Tomi Kinnunen, and Cemal Hanilçi, · 2091
Closest in time.