Fetching the paper…
Reading the bibliography…
The social media revolution has produced a plethora of web services to which users can easily upload and share multimedia documents.
J. S. Garofolo, “Timit acoustic phonetic continuous speech corpus,”
1993
Earlier work this paper cites.
C. Gussenhoven,
2004
Earlier work this paper cites.
I. Rec, “P. 563: Single-ended method for objective speech quality assessment in narrow-band telephony applications,”
2004
Earlier work this paper cites.
J. Chen, D. T. Huy, K. Phua, J. Biswas, and M. Jayachandran, “Using keyword spotting and replacement for speech anonymization,” in
2007
Earlier work this paper cites.
S. J. Prince and J. H. Elder, “Probabilistic linear discriminant analysis for inferences about identity,” in
2007
Earlier work this paper cites.
Q. Jin, A. R. Toth, T. Schultz, and A. W. Black, “Speaker de-identification via voice transformation,” in
2009
Earlier work this paper cites.
M. Akagi and Y. Irie, “Privacy protection for speech based on concepts of auditory scene analysis,”
2012
Earlier work this paper cites.
M. Pobar and I. Ipšić, “Online speaker de-identification using voice transformation,” in
2014
Earlier work this paper cites.
F. Alegre, G. Soldi, and N. Evans, “Evasion and obfuscation in automatic speaker verification,” in
2014
Cited alongside, same era.
2014
Cited alongside, same era.
Z. Wu, T. Kinnunen, N. Evans, J. Yamagishi, C. Hanilçi, M. Sahidullah, and A. Sizov, “ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge,” in
2015
Cited alongside, same era.
T. Justin, V. Štruc, S. Dobrišek, B. Vesnicer, I. Ipšić, and F. Mihelič, “Speaker de-identification using diphone recognition and speech synthesis,” in
2015
Cited alongside, same era.
L. Sun, K. Li, H. Wang, S. Kang, and H. Meng, “Phonetic posteriorgrams for many-to-one voice conversion without parallel data training,” in
C. Magariños, P. Lopez-Otero, L. Docio-Fernandez, E. Rodriguez-Banga, D. Erro, and C. Garcia-Mateo, “Reversible speaker de-identification using pre-trained transformation functions,”
2017
Later among the works it cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in
2017
Later among the works it cites.
J. Lorenzo-Trueba, F. Fang, X. Wang, I. Echizen, J. Yamagishi, and T. Kinnunen, “Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama’s voice using GAN, WaveNet and low-quality found data,” in
2018
Later among the works it cites.
F. Fang, J. Yamagishi, I. Echizen, M. Sahidullah, and T. Kinnunen, “Transforming acoustic characteristics to deceive playback spoofing countermeasures of speaker verification systems,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
K. Hashimoto, J. Yamagishi, and I. Echizen, “Privacy-preserving sound to degrade automatic speaker verification performance,” in
2016
Cited alongside, same era.
L. Juvela, X. Wang, S. Takaki, S. Kim, M. Airaksinen, and J. Yamagishi, “The NII speech synthesis entry for Blizzard Challenge 2016,” in
2016
Cited alongside, same era.
T. Kinnunen, M. Sahidullah, H. Delgado, M. Todisco, N. Evans, J. Yamagishi, and K. A. Lee, “The ASVspoof 2017 challenge: Assessing the limits of replay spoofing attack detection,”
2017
Cited alongside, same era.
Bagher BabaAli and Karel Vesely, “Kaldi TIMIT recipe,”
Cited in the paper.
2018
Later among the works it cites.
F. Bahmaninezhad, C. Zhang, and J. Hansen, “Convolutional neural network based speaker de-identification,” in
2018
Later among the works it cites.
V. Vestman, B. Soomro, A. Kanervisto, V. Hautamäki, and T. Kinnunen, “Who do i sound like? showcasing speaker recognition technology by youtube voice search,” in
2019
Closest in time.
X. Wang, S. Takaki, and J. Yamagishi, “Neural source-filter-based waveform model for statistical parametric speech synthesis,” in
2019
Closest in time.