Fetching the paper…
Reading the bibliography…
Speaker recognition (SR) is widely used in our daily life as a biometric authentication or identification mechanism.
H. Hermansky, “Perceptual linear predictive (PLP) analysis of speech,” The Journal of the Acoustical Society of America , vol. 87, no. 4, 1990
1990
Earlier work this paper cites.
D. A. Reynolds and R. C. Rose, “Robust text-independent speaker identification using gaussian mixture speaker models,” IEEE Trans. Speech and Audio Processing , vol. 3, no. 1, pp. 72–83, 1995
1995
Earlier work this paper cites.
R. Eberhart and J. Kennedy, “A new optimizer using particle swarm theory,” in MHS , 1995
1995
Earlier work this paper cites.
J. Sohn, N. Kim, and W. Sung, “A statistical model-based voice activity detection,” IEEE Signal Processing Letters , 1999
1999
Earlier work this paper cites.
D. A. Reynolds, T. F. Quatieri, and R. B. Dunn, “Speaker verification using adapted gaussian mixture models,” Digit. Signal Process. , 2000
2000
Earlier work this paper cites.
N. P. H. Thian, C. Sanderson, and S. Bengio, “Spectral subband centroids as complementary features for speaker authentication,” in ICB , 2004
2004
Earlier work this paper cites.
J. Fortuna, P. Sivakumaran, A. Ariyaeeinia, and A. Malegaonkar, “Open-set speaker identification using adapted gaussian mixture models,” in INTERSPEECH , 2005
2005
Earlier work this paper cites.
T. Kinnunen and H. Li, “An overview of text-independent speaker recognition: From features to supervectors,” Speech Commun. , 2010
2010
Earlier work this paper cites.
N. Dehak, P. J. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,” IEEE Trans. on Audio, Speech, and Language Processing , 2010
2010
Earlier work this paper cites.
L. Muda, M. Begam, and I. Elamvazuthi, “Voice recognition algorithms using mel frequency cepstral coefficient (MFCC) and dynamic time warping (dtw) techniques,” Journal of Computing , 2010
2010
Earlier work this paper cites.
H. Beigi, Fundamentals of Speaker Recognition . Springer, 12 2011
2011
Earlier work this paper cites.
P. L. De Leon, M. Pucher, J. Yamagishi, I. Hernaez, and I. Saratxaga, “Evaluation of speaker verification security and detection of hmm-based synthetic speech,” IEEE/ACM Trans. Audio, Speech & Language Processing , 2012
2012
Earlier work this paper cites.
B. Biggio, I. Corona, D. Maiorca, B. Nelson, N. Srndic, P. Laskov, G. Giacinto, and F. Roli, “Evasion attacks against machine learning at test time,” in ECML/PKDD , 2013
2013
Earlier work this paper cites.
L. M. Rios and N. V. Sahinidis, “Derivative-free optimization: a review of algorithms and comparison of software implementations,” Journal of Global Optimization , vol. 56, no. 3, pp. 1247–1293, 2013
2013
Earlier work this paper cites.
R. G. Hautamäki, T. Kinnunen, V. Hautamäki, T. Leino, and A.-M. Laukkanen, “I-vectors meet imitators: on vulnerability of speaker verification systems against voice mimicry.” in INTERSPEECH , 2013
2013
Earlier work this paper cites.
Z. Wu and H. Li, “Voice conversion and spoofing attack on speaker verification systems,” in APSIPA , 2013
2013
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in ICLR , 2014
2014
Earlier work this paper cites.
T. Liu and S. Guan, “Factor analysis method for text-independent speaker identification,” JSW , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Z. Wu, S. Gao, E. S. Cling, and H. Li, “A study on replay attack and anti-spoofing for text-dependent speaker verification,” in APSIPA , 2014
2014
Earlier work this paper cites.
M. Shirvanian and N. Saxena, “Wiretapping via mimicry: Short voice imitation mitm attacks on crypto phones,” in ACM CCS , 2014
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in ICLR , 2015
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An asr corpus based on public domain audio books,” in ICASSP , 2015
2015
Earlier work this paper cites.
S. Shiota, F. Villavicencio, J. Yamagishi, N. Ono, I. Echizen, and T. Matsui, “Voice liveness detection algorithms based on pop noise caused by human breath for automatic speaker verification,” in INTERSPEECH , 2015
2015
Earlier work this paper cites.
Z. Wu, N. Evans, T. Kinnunen, J. Yamagishi, F. Alegre, and H. Li, “Spoofing and countermeasures for speaker verification: A survey,” Speech Commun. , 2015
2015
Earlier work this paper cites.
D. Mukhopadhyay, M. Shirvanian, and N. Saxena, “All your voices are belong to us: Stealing voices to fool humans and machines,” in ESORICS , 2015
2015
Earlier work this paper cites.
H. Ren, Y. Song, S. Yang, and F. Situ, “Secure smart home: A voiceprint and internet based authentication system for remote accessing,” in ICCSE , 2016
2016
Earlier work this paper cites.
Citi uses voice prints to authenticate customers quickly and effortlessly. https://www.forbes.com/sites/tomgroenfeldt/2016/06/27/citi-uses-voice-prints-to-authenticate-customers-quickly-and-effortlessly/#7b01dea1109c
2016
Earlier work this paper cites.
N. Carlini, P. Mishra, T. Vaidya, Y. Zhang, M. Sherr, C. Shields, D. Wagner, and W. Zhou, “Hidden voice commands,” in USENIX Security , 2016
2016
Cited alongside, same era.
G. Heigold, I. Moreno, S. Bengio, and N. Shazeer, “End-to-end text-dependent speaker verification,” in ICASSP , 2016, pp. 5115–5119
2016
Cited alongside, same era.
S. Sremath Tirumala and S. R. Shahamiri, “A review on deep learning approaches in speaker identification,” in ICSPS , 2016
2016
Cited alongside, same era.
D. Amodei, S. Ananthanarayanan, R. Anubhai, J. Bai, E. Battenberg, C. Case, J. Casper, B. Catanzaro, Q. Cheng, G. Chen et al. , “Deep speech 2: End-to-end speech recognition in english and mandarin,” in International conference on machine learning , 2016, pp. 173–182
2016
Cited alongside, same era.
M. Sharif, S. Bhagavatula, L. Bauer, and M. K. Reiter, “Accessorize to a crime: Real and stealthy attacks on state-of-the-art face recognition,” in ACM CCS , 2016, pp. 1528–1540
Fosafer VPR. http://caijing.chinadaily.com.cn/chanye/2018-06/06/content_36337667.htm
2018
Later among the works it cites.
V. Vestman, D. Gowda, M. Sahidullah, P. Alku, and T. Kinnunen, “Speaker recognition from whispered speech: A tutorial survey and an application of time-varying linear prediction,” Speech Commun. , 2018
2018
Later among the works it cites.
Y. Dong, F. Liao, T. Pang, H. Su, J. Zhu, X. Hu, and J. Li, “Boosting adversarial attacks with momentum,” in CVPR , 2018, pp. 9185–9193
2018
Later among the works it cites.
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in ICLR , 2018
2018
Later among the works it cites.
J. S. Chung, A. Nagrani, and A. Zisserman, “Voxceleb2: Deep speaker recognition,” in INTERSPEECH , 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
L. Zhang, S. Tan, J. Yang, and Y. Chen, “VoiceLive: A phoneme localization based liveness detection for voice authentication on smartphones,” in ACM CCS , 2016
2016
Cited alongside, same era.
A. Kurakin, I. J. Goodfellow, and S. Bengio, “Adversarial examples in the physical world,” in ICLR , 2017
2017
Cited alongside, same era.
N. Carlini and D. Wagner, “Towards evaluating the robustness of neural networks,” in IEEE S&P , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. Papernot, P. McDaniel, I. Goodfellow, S. Jha, Z. B. Celik, and A. Swami, “Practical black-box attacks against machine learning,” in AsiaCCS , 2017, pp. 506–519
2017
Cited alongside, same era.
P.-Y. Chen, H. Zhang, Y. Sharma, J. Yi, and C.-J. Hsieh, “Zoo: Zeroth order optimization based black-box attacks to deep neural networks without training substitute models,” in AISec , 2017, pp. 15–26
2017
Cited alongside, same era.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in INTERSPEECH , 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
M. Shirvanian, N. Saxena, and D. Mukhopadhyay, “Short voice imitation man-in-the-middle attacks on crypto phones: Defeating humans and machines,” Journal of Computer Security , 2018
2018
Later among the works it cites.
S. Chen, M. Xue, L. Fan, S. Hao, L. Xu, H. Zhu, and B. Li, “Automated poisoning attacks and defenses in malware detection systems: An adversarial machine learning approach,” Computers & Security , 2018
2018
Later among the works it cites.
D. Ribas and E. Vincent, “An improved uncertainty propagation method for robust i-vector based speaker recognition,” in ICASSP , 2019
2019
Closest in time.
R. Taori, A. Kamsetty, B. Chu, and N. Vemuri, “Targeted adversarial examples for black box audio systems,” in IEEE S&P Workshops , 2019
2019
Closest in time.
H. Yakura and J. Sakuma, “Robust audio adversarial example for a physical attack,” in IJCAI , 2019
2019
Closest in time.
Y. Qin, N. Carlini, G. W. Cottrell, I. J. Goodfellow, and C. Raffel, “Imperceptible, robust, and targeted adversarial examples for automatic speech recognition,” in ICML , 2019
2019
Closest in time.
D. Snyder, D. Garcia-Romero, G. Sell, A. McCree, D. Povey, and S. Khudanpur, “Speaker recognition for multi-speaker conversations using x-vectors,” in IEEE ICASSP , 2019
2019
Closest in time.
Z. Yang, B. Li, P. Chen, and D. Song, “Characterizing audio adversarial examples using temporal dependency,” in ICLR , 2019
2019
Closest in time.
M. K. Nandwana, L. Ferrer, M. McLaren, D. Castan, and A. Lawson, “Analysis of critical metadata factors for the calibration of speaker recognition systems,” in INTERSPEECH , 2019
2019
Closest in time.
P. S. Nidadavolu, V. Iglesias, J. Villalba, and N. Dehak, “Investigation on neural bandwidth extension of telephone speech for improved speaker recognition,” in ICASSP , 2019
2019
Closest in time.
K. A. Lee, Q. Wang, and T. Koshinaka, “The CORAL+ algorithm for unsupervised domain adaptation of PLDA,” in ICASSP , 2019
2019
Closest in time.
2019
Closest in time.
M. Alzantot, Y. Sharma, S. Chakraborty, H. Zhang, C. Hsieh, and M. B. Srivastava, “Genattack: practical black-box attacks with gradient-free optimization,” in GECCO , 2019, pp. 1111–1119
2019
Closest in time.
2019
Closest in time.
X. Du, X. Xie, Y. Li, L. Ma, Y. Liu, and J. Zhao, “Deepstellar: Model-based quantitative analysis of stateful deep learning systems,” in ESEC/FSE , 2019
2019
Closest in time.
L. Schönherr, K. Kohls, S. Zeiler, T. Holz, and D. Kolossa, “Adversarial attacks against automatic speech recognition systems via psychoacoustic hiding,” in NDSS , 2019
2019
Closest in time.
H. Abdullah, W. Garcia, C. Peeters, P. Traynor, K. R. B. Butler, and J. Wilson, “Practical hidden voice attacks against speech and speaker recognition systems,” in NDSS , 2019
2019
Closest in time.
M. Shirvanian, S. Vo, and N. Saxena, “Quantifying the breakability of voice assistants,” in PerCom , 2019
2019
Closest in time.
2020
Closest in time.
X. Zhang, X. Xie, L. Ma, X. Du, Q. Hu, Y. Liu, J. Zhao, and M. Sun, “Towards characterizing adversarial defects of deep learning software from the lens of uncertainty,” in ICSE , 2020
2020
Closest in time.