Fetching the paper…
Reading the bibliography…
Adversarial examples to speaker recognition (SR) systems are generated by adding a carefully crafted noise to the speech signal to make the system fail while being imperceptible to humans.
F. Kreuk, Y. Adi, M. Cisse, and J. Keshet, “Fooling End-To-End Speaker Verification With Adversarial Examples,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2018, pp. 1962–1966
1966
Earlier work this paper cites.
G. R. Doddington, M. A. Przybocki, A. F. Martin, and D. A. Reynolds, “The NIST speaker recognition evaluation - Overview, methodology, systems, results, perspective,” Speech Communication , 2000
2000
Earlier work this paper cites.
J. Kominek and A. W. Black, “The CMU Arctic speech databases,” in ISCA workshop on speech synthesis , 2004
2004
Earlier work this paper cites.
N. Dehak, P. J. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 19, no. 4, pp. 788–798, 2010
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz et al. , “The kaldi speech recognition toolkit,” in IEEE workshop on automatic speech recognition and understanding , 2011
2011
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in International Conference on Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” Advances in Neural Information Processing Systems , vol. 27, pp. 2672–2680, 2014
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in International Conference of Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
Z. Wu, N. Evans, T. Kinnunen, J. Yamagishi, F. Alegre, and H. Li, “Spoofing and countermeasures for speaker verification: A survey,” Speech Communication , vol. 66, pp. 130–153, 2015
2015
Earlier work this paper cites.
I. J. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and Harnessing Adversarial Examples,” in International Conference on Learning Representations (ICLR) , 2015
2015
Earlier work this paper cites.
T. Vaidya, Y. Zhang, M. Sherr, and C. Shields, “Cocaine noodles: exploiting the gap between human and machine speech recognition,” in USENIX Workshop on Offensive Technologies (WOOT) , 2015
2015
Earlier work this paper cites.
I. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in International Conference on Learning Representations , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An asr corpus based on public domain audio books,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2015
2015
Earlier work this paper cites.
N. Carlini, P. Mishra, T. Vaidya, Y. Zhang, M. Sherr, C. Shields, D. Wagner, and W. Zhou, “Hidden voice commands,” in USENIX Security Symposium , 2016, pp. 513–530
2016
Earlier work this paper cites.
D. Amodei, S. Ananthanarayanan, R. Anubhai, J. Bai, E. Battenberg, C. Case, J. Casper, B. Catanzaro, Q. Cheng, G. Chen et al. , “Deep speech 2: End-to-end speech recognition in english and mandarin,” in International conference on machine learning , 2016, pp. 173–182
2016
Earlier work this paper cites.
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, and K. Kavukcuoglu, “Wavenet: A generative model for raw audio,” in ISCA Speech Synthesis Workshop , 2016, pp. 125–125
2016
Earlier work this paper cites.
W. Shi, J. Caballero, F. Huszár, J. Totz, A. P. Aitken, R. Bishop, D. Rueckert, and Z. Wang, “Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network,” in IEEE conference on computer vision and pattern recognition , 2016
2016
Earlier work this paper cites.
N. Carlini and D. Wagner, “Towards Evaluating the Robustness of Neural Networks,” in IEEE Symposium on Security and Privacy , 2017
2017
Earlier work this paper cites.
A. Kurakin, I. J. Goodfellow, and S. Bengio, “Adversarial examples in the physical world,” in Workshop Track of the International Conference on Learning Representations (ICLR) , 2017
2017
Earlier work this paper cites.
D. Iter, J. Huang, and M. Jermann, “Generating adversarial examples for speech recognition,” Stanford Technical Report , 2017
2017
Earlier work this paper cites.
M. Cisse, Y. Adi, N. Neverova, and J. Keshet, “Houdini: Fooling Deep Structured Prediction Models,” in NIPS , 2017, pp. 6977—-6987
2017
Earlier work this paper cites.
A. Kurakin, I. Goodfellow, and S. Bengio, “Adversarial machine learning at scale,” in International Conference on Learning Representations, ICLR , 2017
2017
Earlier work this paper cites.
D. Snyder, D. Garcia-Romero, D. Povey, and S. Khudanpur, “Deep Neural Network Embeddings for Text-Independent Speaker Verification,” in Interspeech , 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in neural information processing systems , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
S.-M. Moosavi-Dezfooli, A. Fawzi, O. Fawzi, and P. Frossard, “Universal adversarial perturbations,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 1765–1773
2017
Earlier work this paper cites.
T. Ko, V. Peddinti, D. Povey, M. L. Seltzer, and S. Khudanpur, “A study on data augmentation of reverberant speech for robust speech recognition,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2017, pp. 5220–5224
2017
Earlier work this paper cites.
M. Arjovsky, S. Chintala, and L. Bottou, “Wasserstein generative adversarial networks,” in Proceedings of the 34th International Conference on Machine Learning-Volume 70 , 2017, pp. 214–223
2017
Cited alongside, same era.
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. Courville, “Improved Training of Wasserstein GANs,” in International Conference on Neural Information Processing Systems (NIPS) , 2017
2017
Cited alongside, same era.
A. A. Alemi, I. Fischer, J. V. Dillon, and K. Murphy, “Deep variational information bottleneck,” International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
N. Carlini and D. Wagner, “Audio adversarial examples: Targeted attacks on speech-to-text,” in IEEE Security and Privacy Workshops (SPW) , 2018
2018
Cited alongside, same era.
J. Cohen, E. Rosenfeld, and Z. Kolter, “Certified adversarial robustness via randomized smoothing,” in International Conference on Machine Learning , 2019, pp. 1310–1320
2019
Later among the works it cites.
C. Donahue, J. McAuley, and M. Puckette, “Adversarial Audio Synthesis,” International Conference on Learning Representations , 2019
2019
Later among the works it cites.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala, “PyTorch: An Imperative Style, High-Performance Deep Learning Library,” in NeurIPS , 2019
2019
Later among the works it cites.
R. K. Das, X. Tian, T. Kinnunen, and H. Li, “The Attacker’s Perspective on Automatic Speaker Verification: An Overview,” Interspeech , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Gong and C. Poellabauer, “Crafting Adversarial Examples For Speech Paralinguistics Applications,” in DYnamic and Novel Advances in Machine Learning and Intelligent Cyber Security (DYNAMICS) Workshop , 2018
2018
Cited alongside, same era.
X. Yuan, Y. Chen, Y. Zhao, Y. Long, X. Liu, K. Chen, S. Zhang, H. Huang, X. Wang, and C. A. Gunter, “Commandersong: A systematic approach for practical adversarial voice recognition,” in USENIX Security Symposium , 2018, pp. 49–64
2018
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-Vectors : Robust DNN Embeddings for Speaker Recognition,” in ICASSP , 2018, pp. 5329–5333
2018
Cited alongside, same era.
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
Y. Song, T. Kim, S. Nowozin, S. Ermon, and N. Kushman, “Pixeldefend: Leveraging generative models to understand and defend against adversarial examples,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
P. Samangouei, M. Kabkab, and R. Chellappa, “Defense-GAN: Protecting Classifiers Against Adversarial Attacks Using Generative Models,” International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
K. Rajaratnam and J. Kalita, “Noise flooding for detecting audio adversarial examples against automatic speech recognition,” in IEEE International Symposium on Signal Processing and Information Technology (ISSPIT) , 2018, pp. 197–201
2018
Cited alongside, same era.
A. Athalye, N. Carlini, and D. Wagner, “Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples,” in International Conference on Machine Learning , 2018, pp. 274–283
2018
Cited alongside, same era.
M. R. Kamble, H. B. Sailor, H. A. Patil, and H. Li, “Advances in anti-spoofing: from the perspective of asvspoof challenges,” APSIPA Transactions on Signal and Information Processing , vol. 9, 2020
2020
Later among the works it cites.
J. Villalba, Y. Zhang, and N. Dehak, “x-vectors meet adversarial attacks: Benchmarking adversarial robustness in speaker verification,” Interspeech , 2020
2020
Later among the works it cites.
Y. Dong, Q.-A. Fu, X. Yang, T. Pang, H. Su, Z. Xiao, and J. Zhu, “Benchmarking adversarial robustness on image classification,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 321–331
2020
Later among the works it cites.
H. X. Y. M. Hao-Chen, L. D. Deb, H. L. J.-L. T. Anil, and K. Jain, “Adversarial attacks and defenses in images, graphs and text: A review,” International Journal of Automation and Computing , vol. 17, no. 2, pp. 151–178, 2020
2020
Later among the works it cites.
D. Wang, R. Wang, L. Dong, D. Yan, X. Zhang, and Y. Gong, “Adversarial examples attack and countermeasure for speech recognition system: A survey,” in International Conference on Security and Privacy in Digital Economy , 2020, pp. 443–468
2020
Later among the works it cites.
X. Li, J. Zhong, X. Wu, J. Yu, X. Liu, and H. Meng, “Adversarial Attacks on GMM I-Vector Based Speaker Verification Systems,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2020
2020
Later among the works it cites.
Y. Xie, C. Shi, Z. Li, J. Liu, Y. Chen, and B. Yuan, “Real-Time, Universal, and Robust Adversarial Attacks Against Speaker Recognition Systems,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2020, pp. 1738–1742
2020
Later among the works it cites.
J. Li, X. Zhang, C. Jia, J. Xu, L. Zhang, Y. Wang, S. Ma, and W. Gao, “Universal adversarial perturbations generative network for speaker recognition,” in IEEE International Conference on Multimedia and Expo (ICME) , 2020, pp. 1–6
2020
Later among the works it cites.
Q. Wang, P. Guo, and L. Xie, “Inaudible adversarial perturbations for targeted attack in speaker recognition,” Interspeech , 2020
2020
Later among the works it cites.
Y. Zhang, Z. Jiang, J. Villalba, and N. Dehak, “Black-box attacks on spoofing countermeasures using transferability of adversarial examples,” Interspeech , 2020
2020
Later among the works it cites.
X. Li, N. Li, J. Zhong, X. Wu, X. Liu, D. Su, D. Yu, and H. Meng, “Investigating robustness of adversarial samples detection for automatic speaker verification,” Interspeech , 2020
2020
Later among the works it cites.
I. Andronic, L. Kürzinger, E. R. C. Rosas, G. Rigoll, and B. U. Seeber, “Mp3 compression to diminish adversarial noise in end-to-end speech recognition,” in International Conference on Speech and Computer , 2020, pp. 22–34
2020
Later among the works it cites.
J. Villalba, N. Chen, D. Snyder, D. Garcia-Romero, A. McCree, G. Sell, J. Borgstrom, L. P. García-Perera, F. Richardson, R. Dehak, P. A. Torres-Carrasquillo, and N. Dehak, “State-of-the-art speaker recognition with neural network embeddings in NIST SRE18 and Speakers in the Wild evaluations,” Computer Speech & Language , vol. 60, p. 101026, 2020
2020
Later among the works it cites.
F. Tramer, N. Carlini, W. Brendel, and A. Madry, “On adaptive attacks to adversarial example defenses,” Advances in Neural Information Processing Systems , vol. 33, 2020
2020
Later among the works it cites.
R. Yamamoto, E. Song, and J.-M. Kim, “Parallel wavegan: A fast waveform generation model based on generative adversarial networks with multi-resolution spectrogram,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020
2020
Later among the works it cites.
A. Nagrani, J. S. Chung, W. Xie, and A. Zisserman, “VoxCeleb: Large-scale speaker verification in the wild,” Computer Speech and Language , 2020
2020
Later among the works it cites.
H. Abdullah, K. Warren, V. Bindschaedler, N. Papernot, and P. Traynor, “SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems,” IEEE Symposium on Security and Privacy , 2021
2021
Closest in time.
A. Jati, C.-C. Hsu, M. Pal, R. Peri, W. AbdAlmageed, and S. Narayanan, “Adversarial attack and defense strategies for deep speaker recognition systems,” Computer Speech & Language , vol. 68, p. 101199, 2021
2021
Closest in time.
M. Pal, A. Jati, R. Peri, C.-C. Hsu, W. AbdAlmageed, and S. Narayanan, “Adversarial defense for deep speaker recognition using hybrid adversarial training,” IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , pp. 6164–6168, 2021
2021
Closest in time.
C. Laidlaw, S. Singla, and S. Feizi, “Perceptual adversarial robustness: Defense against unseen threat models,” in International Conference on Learning Representations , 2021
2021
Closest in time.
M. Esmaeilpour, P. Cardinal, and A. L. Koerich, “Class-conditional defense gan against end-to-end speech attacks,” in IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2021
2021
Closest in time.
D. J. Im, S. Ahn, R. Memisevic, and Y. Bengio, “Denoising criterion for variational auto-encoding framework,” in Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence , 2017, pp. 2059–2065
2065
Closest in time.