Fetching the paper…
Reading the bibliography…
Diverse promising datasets have been designed to hold back the development of fake audio detection, such as ASVspoof databases.
Y. W. Lau, M. Wagner, and D. Tran, “Vulnerability of speaker verification to voice mimicking,” in Proc. of Int. Symposium on Intelligent Multimedia,Video and Speech Processing , 2004
2004
Earlier work this paper cites.
J.-F. Bonastre, D. Matrouf, and C. Fredouille, “Artificial impostor voice transformation effects on false acceptance rates,” in Proc. of INTERSPEECH , 2007
2007
Earlier work this paper cites.
P. L. D. Leon, M. Pucher, and J. Yamagishi, “Evaluation of the vulnerability of speaker verification to synthetic speech,” in Proc. of Odyssey: The Speaker and Language Recognition Workshop , 2010
2010
Earlier work this paper cites.
T. Kinnune, Z.-Z. Wu, K. A. Lee, and et al., “Vulnerability of speaker verification systems against voice conversion spoofing attacks: The case of telephone speech,” in Proc. of ICASSP , 2012
2012
Earlier work this paper cites.
R. G. Hautamaki, T. Kinnunen, V. Hautamaki, T. Leino, and A. M. Laukkanen, “I-vectors meet imitators: on vulnerability of speaker verification systems against voice mimicry,” in Proc. of INTERSPEECH , 2013
2013
Earlier work this paper cites.
F. Alegre, A. Amehraye, and N. Evans, “A one-class classification approach to generalised speaker verification spoofing countermeasures using local binary patterns,” in Proc. of Int. Conf. on Biometrics: Theory, Applications and Systems (BTAS) , 2013
2013
Earlier work this paper cites.
Z. Kons and H. Aronowitz, “Voice transformation-based spoofing of text dependent speaker verification systems,” in Proc. of INTERSPEECH , 2013
2013
Earlier work this paper cites.
Z. Wu, A. Larcher, K. A. Lee, and et al., “Vulnerability evaluation of speaker verification under voice conversion spoofing: the effect of text constraints,” in Proc. of INTERSPEECH , 2013
2013
Earlier work this paper cites.
T. Gulzar, A. Singh, and S. Sharma, “Comparative analysis of lpcc, mfcc and bfcc for the recognition of hindi words using artificial neural networks,” International Journal of Computer Applications , vol. 101, no. 12, pp. 22–27, 2014
2014
Earlier work this paper cites.
Z. Wu, T. Kinnunen, N. Evans, J. Yamagishi, C. Hanilc¸i, and et al., “Asvspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge,” in Proc. of INTERSPEECH , 2015
2015
Earlier work this paper cites.
Z. Wu, A. Khodabakhsh, C. Demiroglu, and et al., “Sas : A speaker verification spoofing database containing diverse attacks,” in Proc. of ICASSP , 2015
2015
Cited alongside, same era.
Y. Wang, R. J. Skerry-Ryan, D. Stanton, Y. Wu, and R. A. Saurous, “Tacotron: Towards end-to-end speech synthesis,” in Proc. of INTERSPEECH , 2017
2017
Cited alongside, same era.
T. Kinnunen, M. Sahidullah, H. Delgado, N. E. M. Todisco, and et al., “The asvspoof 2017 challenge: Assessing the limits of replay spoofing attack detection,” in Proc. of INTERSPEECH , 2017
2017
Cited alongside, same era.
M. Todisco, H. Delgado, and N. Evans, “Constant q cepstral coefficients: A spoofing countermeasure for automatic speaker verification,” Computer Speech and Language , vol. 45, pp. 516–535, 2017
2017
Cited alongside, same era.
G. Lavrentyeva, S. Novoselov, E. Malykh, A. Kozlov, and V. Shchemelinin, “Audio replay attack detection with deep learning frameworks,” in Interspeech , 2017
M. Todisco, X. Wang, V. Vestman, M. Sahidullah, and K. A. Lee, “Asvspoof 2019: Future horizons in spoofed and fake audio detection,” in Proc. of INTERSPEECH , 2019
2019
Later among the works it cites.
G. Lavrentyeva1, S. Novoselov1, M. Volkova, and et al., “Phonespoof: A new dataset for spoofing attack detection in telephone channel,” in Proc. of ICASSP , 2019
2019
Later among the works it cites.
J. Valin and J. Skoglund, “LPCNET: improving neural speech synthesis through linear prediction,” in Proc. of ICASSP , 2019
2019
Later among the works it cites.
T. Chen, A. Kumar, P. Nagarsheth, G. Sivaraman, and E. Khoury, “Generalization of audio deepfake detection,” in Proc. of Odyssey: The Speaker and Language Recognition Workshop , 2020
2020
Later among the works it cites.
R. Wang, F. Juefei-Xu, Y. Huang, Q. Guo, and et al., “Deepsonar: Towards effective and robust detection of ai-synthesized fake voices,” in Proc. of ACM MM , 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
J. Shen, R. Pang, R. J. Weiss, and et al., “Natural tts synthesis by conditioning wavenet on mel spectrogram predictions,” in Proc. of ICASSP , 2018
2018
Cited alongside, same era.
Y. Wang, D. Stanton, Y. Zhang, and et al., “Style tokens: Unsupervised style modeling, control and transfer in end-to-end speech synthesis,” in Proc. of ICML , 2018
2018
Cited alongside, same era.
R. J. Skerry-Ryan, E. Battenberg, X. Ying, Y. Wang, and R. A. Saurous, “Towards end-to-end prosody transfer for expressive speech synthesis with tacotron,” in Proc. of ICML , 2018
2018
Cited alongside, same era.
X. Wu, R. He, Z. Sun, and T. Tan, “A light cnn for deep face representation with noisy labels,” IEEE Transactions on Information Forensics and Security , vol. 13, no. 11, pp. 2884–2896, 2018
2018
Cited alongside, same era.
2020
Later among the works it cites.
Z. Wu, R. K. Das1, J. Yang, and H. Li, “Light convolutional neural network with feature genuinization for detection of synthetic speech attacks,” in Proc. of INTERSPEECH , 2020
2020
Later among the works it cites.
X. Wang, J. Yamagishi1, M. Todisco1c, and et al., “Asvspoof 2019: A large-scale public database of synthesized, converted and replayed speech,” Journal of Computer Speech and Language , vol. 64, 2020
2020
Later among the works it cites.
R. Reimao and V. Tzerpos, “For: A dataset for synthetic speech detection,” 2020
2020
Later among the works it cites.