Fetching the paper…
Reading the bibliography…
Voice interfaces are becoming accepted widely as input methods for a diverse set of devices.
ISO, “Information Technology – Coding of moving pictures and associated audio for digital storage media at up to 1.5 Mbits/s – Part3: Audio,” International Organization for Standardization, ISO 11172-3, 1993
1993
Earlier work this paper cites.
L. Rabiner and B.-H. Juang, Fundamentals of Speech Recognition . Prentice-Hall, Inc., 1993
1993
Earlier work this paper cites.
R. B. Wolfgang, C. I. Podilchuk, and E. J. Delp, “Perceptual watermarks for digital images and video,” Proceedings of the IEEE , vol. 87, no. 7, pp. 1108–1126, 1999
1999
Earlier work this paper cites.
M. Barni, F. Bartolini, and A. Piva, “Improved wavelet-based watermarking through pixel-wise masking,” IEEE transactions on image processing , vol. 10, no. 5, pp. 783–791, 2001
2001
Earlier work this paper cites.
G. Navarro, “A guided tour to approximate string matching,” ACM Computing Surveys , vol. 33, no. 1, pp. 31–88, Mar. 2001
2001
Earlier work this paper cites.
J. W. Seok and J. W. Hong, “Audio watermarking for copyright protection of digital audio data,” Electronics Letters , vol. 37, no. 1, pp. 60–61, 2001
2001
Earlier work this paper cites.
S. Wu, J. Huang, D. Huang, and Y. Q. Shi, “Self-synchronized audio watermark in dwt domain,” in 2004 IEEE International Symposium on Circuits and Systems (IEEE Cat. No.04CH37512) , vol. 5, May 2004, pp. V–V
2004
Earlier work this paper cites.
D. Lowd and C. Meek, “Adversarial learning,” in Conference on Knowledge Discovery in Data Mining . ACM, Aug. 2005, pp. 641–647
2005
Earlier work this paper cites.
M. Barreno, B. Nelson, R. Sears, A. D. Joseph, and J. D. Tygar, “Can machine learning be secure?” in Symposium on Information, Computer and Communications Security . ACM, Mar. 2006, pp. 16–25
2006
Earlier work this paper cites.
I. Cox, M. Miller, J. Bloom, J. Fridrich, and T. Kalker, Digital watermarking and steganography . Morgan kaufmann, 2007
2007
Earlier work this paper cites.
E. Zwicker and H. Fastl, Psychoacoustics: Facts and Models , 3rd ed. Springer, 2007
2007
Earlier work this paper cites.
M. Barreno, B. Nelson, A. D. Joseph, and J. D. Tygar, “The security of machine learning,” Machine Learning , vol. 81, no. 2, pp. 121–148, Nov. 2010
2010
Earlier work this paper cites.
M. Nutzinger, C. Fabian, and M. Marschalek, “Secure hybrid spread spectrum system for steganography in auditive media,” in Conference on Intelligent Information Hiding and Multimedia Signal Processing . IEEE, Oct. 2010, pp. 78–81
2010
Earlier work this paper cites.
M. Asad, J. Gilani, and A. Khalid, “An enhanced Least Significant Bit modification technique for audio steganography,” in Conference on Computer Networks and Information Technology . IEEE, Jul. 2011, pp. 143–147
2011
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The Kaldi speech recognition toolkit,” in Workshop on Automatic Speech Recognition and Understanding . IEEE, Dec. 2011
2011
Earlier work this paper cites.
J. Liu, K. Zhou, and H. Tian, “Least-Significant-Digit steganography in low bitrate speech,” in International Conference on Communications . IEEE, Jun. 2012, pp. 1133–1137
2012
Earlier work this paper cites.
N. Schinkel-Bielefeld, N. Lotze, and F. Nagel, “Audio quality evaluation by experienced and inexperienced listeners,” in International Congress on Acoustics . ASA, Jun. 2013, pp. 6–16
2013
Earlier work this paper cites.
W. Diao, X. Liu, Z. Zhou, and K. Zhang, “Your voice assistant is mine: How to abuse speakers to steal information and control your phone,” in Workshop on Security and Privacy in Smartphones & Mobile Devices . ACM, Nov. 2014, pp. 63–74
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Nguyen, J. Yosinski, and J. Clune, “Deep neural networks are easily fooled: High confidence predictions for unrecognizable images,” in Conference on Computer Vision and Pattern Recognition . IEEE, Jun. 2015, pp. 427–436
2015
Earlier work this paper cites.
T. Vaidya, Y. Zhang, M. Sherr, and C. Shields, “Cocaine Noodles: Exploiting the gap between human and machine speech recognition,” in Workshop on Offensive Technologies . USENIX, Aug. 2015
2015
Earlier work this paper cites.
N. Carlini, P. Mishra, T. Vaidya, Y. Zhang, M. Sherr, C. Shields, D. A. Wagner, and W. Zhou, “Hidden voice commands,” in USENIX Security Symposium . USENIX, Aug. 2016, pp. 513–530
2016
Cited alongside, same era.
A. Fawzi, S.-M. Moosavi-Dezfooli, and P. Frossard, “Robustness of classifiers: From adversarial to random noise,” in Conference on Neural Information Processing Systems . Curran Associates, Inc., Dec. 2016, pp. 1632–1640
2016
Cited alongside, same era.
K. Kohls, T. Holz, D. Kolossa, and C. Pöpper, “SkypeLine: Robust hidden data transmission for VoIP,” in Asia Conference on Computer and Communications Security . ACM, May 2016, pp. 877–888
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Ravanelli, P. Brakel, M. Omologo, and Y. Bengio, “A network of deep neural networks for distant speech recognition,” in 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , March 2017, pp. 4880–4884
2017
Later among the works it cites.
N. Roy, H. Hassanieh, and R. Roy Choudhury, “BackDoor: Making microphones hear inaudible sounds,” in Conference on Mobile Systems, Applications, and Services . ACM, Jun. 2017, pp. 2–14
2017
Later among the works it cites.
2017
Later among the works it cites.
L. Song and P. Mittal, “Inaudible voice commands,” CoRR , vol. abs/1708.07238, pp. 1–3, Aug. 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
N. Papernot, P. McDaniel, S. Jha, M. Fredrikson, Z. B. Celik, and A. Swami, “The limitations of deep learning in adversarial settings,” in European Symposium on Security and Privacy . IEEE, Mar. 2016, pp. 372–387
2016
Cited alongside, same era.
N. Papernot, P. McDaniel, X. Wu, S. Jha, and A. Swami, “Distillation as a defense to adversarial perturbations against deep neural networks,” in Symposium on Security and Privacy . IEEE, May 2016, pp. 582–597
2016
Cited alongside, same era.
2016
Cited alongside, same era.
F. Tramèr, F. Zhang, A. Juels, M. K. Reiter, and T. Ristenpart, “Stealing machine learning models via prediction APIs,” in USENIX Security Symposium . USENIX, Aug. 2016, pp. 601–618
2016
Cited alongside, same era.
A. Ali, S. Vogel, and S. Renals, “Speech recognition challenge in the wild: Arabic mgb-3,” in 2017 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU) , Dec 2017, pp. 316–322
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. Carlini and D. Wagner, “Towards evaluating the robustness of neural networks,” in Symposium on Security and Privacy . IEEE, May 2017, pp. 39–57
2017
Cited alongside, same era.
J. Trmal, M. Wiesner, V. Peddinti, X. Zhang, P. Ghahremani, Y. Wang, V. Manohar, H. Xu, D. Povey, and S. Khudanpur, “The kaldi openkws system: Improving low resource keyword search,” Proc. Interspeech , pp. 3597–3601, Aug. 2017
2017
Later among the works it cites.
P. Upadhyaya, O. Farooq, M. R. Abidi, and Y. V. Varshney, “Continuous hindi speech recognition model based on kaldi asr toolkit,” in 2017 International Conference on Wireless Communications, Signal Processing and Networking (WiSPNET) , March 2017, pp. 786–789
2017
Later among the works it cites.
J. Villalba, N. Brümmer, and N. Dehak, “Tied variational autoencoder backends for i-vector speaker recognition,” Interspeech, Stockholm , pp. 1005–1008, Aug. 2017
2017
Later among the works it cites.
W. Xiong, J. Droppo, X. Huang, F. Seide, M. Seltzer, A. Stolcke, D. Yu, and G. Zweig, “Toward human parity in conversational speech recognition,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 25, no. 12, pp. 2410–2423, Dec. 2017
2017
Later among the works it cites.
——, “The microsoft 2016 conversational speech recognition system,” in Acoustics, Speech and Signal Processing (ICASSP), 2017 IEEE International Conference on . IEEE, 2017, pp. 5255–5259
2017
Later among the works it cites.
V. Zantedeschi, M.-I. Nicolae, and A. Rawat, “Efficient defenses against adversarial attacks,” in Workshop on Artificial Intelligence and Security . ACM, Nov. 2017, pp. 39–49
2017
Later among the works it cites.
G. Zhang, C. Yan, X. Ji, T. Zhang, T. Zhang, and W. Xu, “DolphinAttack: Inaudible voice commands,” in Conference on Computer and Communications Security . ACM, Oct. 2017, pp. 103–117
2017
Later among the works it cites.
K. Audhkhasi, B. Kingsbury, B. Ramabhadran, G. Saon, and M. Picheny, “Building competitive direct acoustics-to-word models for english conversational speech recognition,” 2018
2018
Closest in time.
2018
Closest in time.
A. Fawzi, O. Fawzi, and P. Frossard, “Analysis of classifiers’ robustness to adversarial perturbations,” Machine Learning , vol. 107, no. 3, pp. 481–508, Mar. 2018
2018
Closest in time.
2018
Closest in time.
2018
Closest in time.
M. Schoeffler, S. Bartoschek, F.-R. Stöter, M. Roess, S. Westphal, B. Edler, and J. Herre, “webMUSHRA – A comprehensive framework for web-based listening tests,” Journal of Open Research Software , vol. 6, no. 1, Feb. 2018
2018
Closest in time.
U. Shaham, Y. Yamada, and S. Negahban, “Understanding adversarial training: Increasing local stability of supervised models through robust optimization,” Neurocomputing , 2018
2018
Closest in time.
B. Wang and N. Z. Gong, “Stealing hyperparameters in machine learning,” in Symposium on Security and Privacy . IEEE, May 2018
2018
Closest in time.
2018
Closest in time.
T. Moynihan, “How to keep Amazon Echo and Google Home from responding to your TV,” Feb. 2017, https://www.wired.com/2017/02/keep-amazon-echo-google-home-responding-tv/
2026
Closest in time.