Fetching the paper…
Reading the bibliography…
Audio DeepFakes (DF) are artificially generated utterances created using deep learning, with the primary aim of fooling the listeners in a highly convincing manner.
R. Kubichek, “Mel-cepstral distance measure for objective speech quality assessment,” in Proceedings of IEEE Pacific Rim Conference on Communications Computers and Signal Processing , vol. 1, 1993, pp. 125–128 vol.1
1993
Earlier work this paper cites.
X. Zhou, D. Garcia-Romero, R. Duraiswami, C. Espy-Wilson, and S. Shamma, “Linear versus mel frequency cepstral coefficients for speaker recognition,” in 2011 IEEE workshop on automatic speech recognition & understanding . IEEE, 2011, pp. 559–564
2011
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in 2nd International Conference on Learning Representations, ICLR , 2014
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
I. Loshchilov and F. Hutter, “SGDR: stochastic gradient descent with warm restarts,” in 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings . OpenReview.net, 2017
2017
Earlier work this paper cites.
X. Wu, R. He, Z. Sun, and T. Tan, “A light cnn for deep face representation with noisy labels,” IEEE Transactions on Information Forensics and Security , vol. 13, no. 11, pp. 2884–2896, 2018
2018
Earlier work this paper cites.
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=rJzIBfZAb
2018
Earlier work this paper cites.
F. Tramèr, A. Kurakin, N. Papernot, I. Goodfellow, D. Boneh, and P. McDaniel, “Ensemble adversarial training: Attacks and defenses,” in International Conference on Learning Representations , 2018
2018
Earlier work this paper cites.
S. Liu, H. Wu, H.-Y. Lee, and H. Meng, “Adversarial attacks on spoofing countermeasures of automatic speaker verification,” in 2019 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU) , 2019, pp. 312–319
2019
Earlier work this paper cites.
L. Schönherr, K. Kohls, S. Zeiler, T. Holz, and D. Kolossa, “Adversarial attacks against automatic speech recognition systems via psychoacoustic hiding,” in 26th Annual Network and Distributed System Security Symposium, NDSS 2019, San Diego, California, USA, February 24-27, 2019 . The Internet Society, 2019
2019
Earlier work this paper cites.
T. Kinnunen, H. Delgado, N. Evans, K. A. Lee, V. Vestman, A. Nautsch, M. Todisco, X. Wang, M. Sahidullah, J. Yamagishi, and D. A. Reynolds, “Tandem assessment of spoofing countermeasures and automatic speaker verification: Fundamentals,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 28, pp. 2195–2210, 2020
2020
Earlier work this paper cites.
H. Wu, S. Liu, H. Meng, and H.-y. Lee, “Defense against adversarial attacks on spoofing countermeasures of asv,” in 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2020, pp. 6564–6568
2020
Cited alongside, same era.
F. Croce and M. Hein, “Minimally distorted adversarial examples with a fast adaptive boundary attack,” in International Conference on Machine Learning . PMLR, 2020, pp. 2196–2205
2020
Cited alongside, same era.
2021
Cited alongside, same era.
H. Khalid, S. Tariq, M. Kim, and S. S. Woo, “FakeAVCeleb: A novel audio-video multimodal deepfake dataset,” in Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 2) , 2021
2021
Cited alongside, same era.
X. Wang and J. Yamagishi, “Investigating Self-Supervised Front Ends for Speech Spoofing Countermeasures,” in Proc. The Speaker and Language Recognition Workshop (Odyssey 2022) , 2022, pp. 100–106
2022
Closest in time.
J. weon Jung, Y. Kim, H.-S. Heo, B.-J. Lee, Y. Kwon, and J. S. Chung, “Pushing the limits of raw waveform speaker recognition,” in Proc. Interspeech 2022 , 2022, pp. 2228–2232
2022
Closest in time.
P. Kawa, M. Plata, and P. Syga, “Specrnet: Towards faster and more accessible audio deepfake detection,” in 2022 IEEE 21st International Conference on Trust, Security and Privacy in Computing and Communications (TrustCom) . IEEE, 2022
2022
Closest in time.
——, “Attack Agnostic Dataset: Towards Generalization and Stabilization of Audio DeepFake Detection,” in Proc. Interspeech 2022 , 2022, pp. 4023–4027
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Tak, J. Patino, M. Todisco, A. Nautsch, N. Evans, and A. Larcher, “End-to-end anti-spoofing with rawnet2,” 2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Jun 2021. [Online]. Available: http://dx.doi.org/10.1109/ICASSP39728.2021.9414234
2021
Cited alongside, same era.
2021
Cited alongside, same era.
J. Frank and L. Schönherr, “WaveFake: A Data Set to Facilitate Audio Deepfake Detection,” in Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track , 2021
2021
Cited alongside, same era.
D. Hendrycks, S. Basart, N. Mu, S. Kadavath, F. Wang, E. Dorundo, R. Desai, T. Zhu, S. Parajuli, M. Guo, D. Song, J. Steinhardt, and J. Gilmer, “The many faces of robustness: A critical analysis of out-of-distribution generalization,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2021, pp. 8340–8349
2021
Cited alongside, same era.
Y. Bai, J. Mei, A. L. Yuille, and C. Xie, “Are transformers more robust than cnns?” in Advances in Neural Information Processing Systems , M. Ranzato, A. Beygelzimer, Y. Dauphin, P. Liang, and J. W. Vaughan, Eds., vol. 34. Curran Associates, Inc., 2021, pp. 26 831–26 843
2021
Cited alongside, same era.
T. Bai, J. Luo, J. Zhao, B. Wen, and Q. Wang, “Recent advances in adversarial training for adversarial robustness,” in Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21 , 8 2021, pp. 4312–4321. [Online]. Available: https://doi.org/10.24963/ijcai.2021/591
2021
Cited alongside, same era.
M. Masood, M. Nawaz, K. M. Malik, A. Javed, A. Irtaza, and H. Malik, “Deepfakes generation and detection: State-of-the-art, open challenges, countermeasures, and way forward,” Applied Intelligence , pp. 1–53, 2022
2022
Cited alongside, same era.
N. Müller, P. Czempin, F. Diekmann, A. Froghyar, and K. Böttinger, “Does Audio Deepfake Detection Generalize?” in Proc. Interspeech 2022 , 2022, pp. 2783–2787
2022
Closest in time.
European Commission, “2022 Strengthened Code of Practice on Disinformation,” 2022. [Online]. Available: https://digital-strategy.ec.europa.eu/en/library/2022-strengthened-code-practice-disinformation
2022
Closest in time.
S. Joshi, S. Kataria, Y. Shao, P. Żelasko, J. Villalba, S. Khudanpur, and N. Dehak, “Defense against Adversarial Attacks on Hybrid Speech Recognition System using Adversarial Fine-tuning with Denoiser,” in Proc. Interspeech 2022 , 2022, pp. 5035–5039
2022
Closest in time.
Y. Shao, J. Villalba, S. Joshi, S. Kataria, S. Khudanpur, and N. Dehak, “Chunking Defense for Adversarial Attacks on ASR,” in Proc. Interspeech 2022 , 2022, pp. 5045–5049
2022
Closest in time.
H. Tan, J. Zhang, H. Zhang, L. Wang, Y. Qian, and Z. Gu, “NRI-FGSM: An Efficient Transferable Adversarial Attack for Speaker Recognition Systems,” in Proc. Interspeech 2022 , 2022, pp. 4386–4390
2022
Closest in time.
Y. Liu, Y. Cheng, L. Gao, X. Liu, Q. Zhang, and J. Song, “Practical evaluation of adversarial robustness via adaptive auto attack,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2022, pp. 15 105–15 114
2022
Closest in time.