Fetching the paper…
Reading the bibliography…
Voice Processing Systems (VPSes), now widely deployed, have been made significantly more accurate through the application of recent advances in machine learning.
H. Fletcher and W. A. Munson, “Loudness, its definition, measurement and calculation,” Bell System Technical Journal , vol. 12, no. 4, pp. 377–430, 1933
1933
Earlier work this paper cites.
T. D. Hanley and G. Draegert, “Effect of level of distracting noise upon speaking rate, duration, and intensity.” PURDUE RESEARCH FOUNDATION LAFAYETTE IND, Tech. Rep., 1949
1949
Earlier work this paper cites.
E. C. Cherry, “Some experiments on the recognition of speech, with one and with two ears,” in The Journal of the Acoustical Society of America, 25th . http://www.ee.columbia.edu/~dpwe/papers/Cherry53-cpe.pdf
1953
Earlier work this paper cites.
L. R. Rabiner and R. W. Schafer, Digital processing of speech signals . Prentice Hall, 1978
1978
Earlier work this paper cites.
J. S. Garofolo et al. , “Getting started with the darpa timit cd-rom: An acoustic phonetic continuous speech database,” National Institute of Standards and Technology (NIST), Gaithersburgh, MD , vol. 107, p. 16, 1988
1988
Earlier work this paper cites.
W. G. on Speech Understanding and Aging, “Speech understanding and aging,” The Journal of the Acoustical Society of America , vol. 83, no. 3, pp. 859–895, 1988
1988
Earlier work this paper cites.
L. R. Rabiner, “A tutorial on hidden markov models and selected applications in speech recognition,” in PROCEEDINGS OF THE IEEE , 1989, pp. 257–286
1989
Earlier work this paper cites.
D. Angluin, “Computational learning theory: Survey and selected bibliography,” in Proceedings of the Twenty-fourth Annual ACM Symposium on Theory of Computing , ser. STOC ’92. New York, NY, USA: ACM, 1992, pp. 351–369. [Online]. Available: http://doi.acm.org/10.1145/129712.129746
1992
Earlier work this paper cites.
R. Mannell, “The perceptual and auditory implications of parametric scaling in synthetic speech,” Macquarie University , p. (Chapter 2), 1994. [Online]. Available: \url{http://clas.mq.edu.au/speech/acoustics/auditory_representations/pitchdiscrim.html}
1994
Earlier work this paper cites.
D. J. Plude, J. T. Enns, and D. Brodeur, “The development of selective attention: A life-span overview,” Acta Psychologica , vol. 86, no. 2, pp. 227 – 272, 1994. [Online]. Available: http://www.sciencedirect.com/science/article/pii/0001691894900043
1994
Earlier work this paper cites.
“ISO 226:2003,” https://www.iso.org/standard/34222.html
2003
Earlier work this paper cites.
P. Lamere, P. Kwok, W. Walker, E. Gouvea, R. Singh, B. Raj, and P. Wolf, “Design of the cmu sphinx-4 decoder,” in Eighth European Conference on Speech Communication and Technology , 2003
2003
Earlier work this paper cites.
C. Cieri, D. Miller, and K. Walker, “The fisher corpus: a resource for the next generations of speech-to-text.” in LREC , vol. 4, 2004, pp. 69–71
2004
Earlier work this paper cites.
N. Dalvi, P. Domingos, Mausam, S. Sanghai, and D. Verma, “Adversarial classification,” Proceedings of the 2004 ACM SIGKDD international conference on Knowledge discovery and data mining - KDD ’04 , p. 99, 2004. [Online]. Available: http://portal.acm.org/citation.cfm?doid=1014052.1014066
2004
Earlier work this paper cites.
M. Barreno, B. Nelson, R. Sears, A. D. Joseph, and J. D. Tygar, “Can machine learning be secure?” in Proceedings of the 2006 ACM Symposium on Information, computer and communications security - ASIACCS ’06 , 2006, p. 16. [Online]. Available: https://www.cs.drexel.edu/{~}greenie/cs680/asiaccs06.pdfhttp://portal.acm.org/citation.cfm?doid=1128817.1128824
2006
Earlier work this paper cites.
J. Newsome, B. Karp, and D. Song, “Paragraph: Thwarting signature learning by training maliciously,” in Recent Advances in Intrusion Detection , D. Zamboni and C. Kruegel, Eds. Berlin, Heidelberg: Springer Berlin Heidelberg, 2006, pp. 81–105
2006
Earlier work this paper cites.
F. Pulvermuller and Y. Shtyrov, “Language outside the focus of attention: The mismatch negativity as a tool for studying higher cognitive processes,” Progress in Neurobiology , vol. 79, no. 1, pp. 49 – 71, 2006. [Online]. Available: http://www.sciencedirect.com/science/article/pii/S0301008206000323
2006
Earlier work this paper cites.
J. Ramírez, J. M. Górriz, and J. C. Segura, “Voice activity detection. fundamentals and speech recognition system robustness,” in Robust speech recognition and understanding , 2007
2007
Earlier work this paper cites.
R. L. Diehl, “Acoustic and auditory phonetics: the adaptive design of speech sound systems,” Philosophical Transactions of the Royal Society B: Biological Sciences , vol. 363, p. 965–978, 2008. [Online]. Available: /url{https://www.ncbi.nlm.nih.gov/pmc/articles/PMC2606790/pdf/rstb20072153.pdf}
2008
Earlier work this paper cites.
B. G. Shinn-Cunningham, “Object-based auditory and visual attention,” in Trends in Cognitive Science , 2008, pp. 182–186. [Online]. Available: \url{http://www.cns.bu.edu/~shinn/resources/pdfs/2008/2008TICS_Shinn.pdf}
2008
Earlier work this paper cites.
S. A. Gelfand, Hearing: An Introduction to Psychological and Physiological Acoustics , 5th ed. Informa Healthcare, 2009
2009
Earlier work this paper cites.
B. P. Lathi and Z. Ding, Modern Digital and Analog Communication Systems , 4th ed. Oxford University Press, 2009
2009
Cited alongside, same era.
M. Barreno, B. Nelson, A. D. Joseph, and J. D. Tygar, “The security of machine learning,” Machine Learning , vol. 81, no. 2, pp. 121–148, 2010
2010
Cited alongside, same era.
L. Torrey and J. Shavlik, “Transfer learning,” in Handbook of Research on Machine Learning Applications and Trends: Algorithms, Methods, and Techniques . IGI Global, 2010, pp. 242–264
2010
Cited alongside, same era.
L. Huang, A. D. Joseph, B. Nelson, B. I. Rubinstein, and J. D. Tygar, “Adversarial machine learning,” in Proceedings of the 4th ACM Workshop on Security and Artificial Intelligence , ser. AISec ’11. New York, NY, USA: ACM, 2011, pp. 43–58. [Online]. Available: http://doi.acm.org/10.1145/2046684.2046692
2011
Cited alongside, same era.
S. M. Moosavi Dezfooli, A. Fawzi, and P. Frossard, “Deepfool: a simple and accurate method to fool deep neural networks,” in Proceedings of 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , no. EPFL-CONF-218057, 2016
2016
Later among the works it cites.
H. Newton and S. Schoen, Newton’s Telecom Dictionary , 30th ed. Harry Newton, 2016
2016
Later among the works it cites.
N. Papernot, P. McDaniel, S. Jha, M. Fredrikson, Z. B. Celik, and A. Swami, “The limitations of deep learning in adversarial settings,” in Security and Privacy (EuroS&P), 2016 IEEE European Symposium on . IEEE, 2016, pp. 372–387
2016
Later among the works it cites.
M. Sharif, S. Bhagavatula, L. Bauer, and M. K. Reiter, “Accessorize to a crime: Real and stealthy attacks on state-of-the-art face recognition,” in Proceedings of the 2016 ACM SIGSAC Conference on Computer and Communications Security . ACM, 2016, pp. 1528–1540
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Limb, “Building the musical muscle,” TEDMED , 2011. [Online]. Available: \url={https://www.ted.com/talks/charles_limb_building_the_musical_muscle#t-367224}
2011
Cited alongside, same era.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The kaldi speech recognition toolkit,” in IEEE 2011 Workshop on Automatic Speech Recognition and Understanding . IEEE Signal Processing Society, 2011, iEEE Catalog No.: CFP11SRW-USB
2011
Cited alongside, same era.
B. Hartpence, Packet Guide to Voice over IP , 1st ed. O‘Reilly Media, Inc, 2013
2013
Cited alongside, same era.
R. Pascanu, T. Mikolov, and Y. Bengio, “On the difficulty of training recurrent neural networks,” in Proceedings of the 30th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, S. Dasgupta and D. McAllester, Eds., vol. 28, no. 3. Atlanta, Georgia, USA: PMLR, 17–19 Jun 2013, pp. 1310–1318. [Online]. Available: http://proceedings.mlr.press/v28/pascanu13.html
2013
Cited alongside, same era.
2013
Cited alongside, same era.
2014
Cited alongside, same era.
A. Graves and N. Jaitly, “Towards end-to-end speech recognition with recurrent neural networks,” in International Conference on Machine Learning , 2014, pp. 1764–1772
2014
Cited alongside, same era.
2014
Cited alongside, same era.
A. Stojanow and J. Liebetrau, “A review on conventional psychoacoustic evaluation tools, methods and algorithms,” in 2016 Eighth International Conference on Quality of Multimedia Experience (QoMEX) , June 2016, pp. 1–6. [Online]. Available: https://ieeexplore.ieee.org/document/7498923/
2016
Later among the works it cites.
L. Zhang, S. Tan, J. Yang, and Y. Chen, “Voicelive: A phoneme localization based liveness detection for voice authentication on smartphones,” in Proceedings of the 2016 ACM SIGSAC Conference on Computer and Communications Security . ACM, 2016, pp. 1080–1091
2016
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
N. Carlini and D. Wagner, “Towards evaluating the robustness of neural networks,” in Security and Privacy (SP), 2017 IEEE Symposium on . IEEE, 2017, pp. 39–57
2017
Later among the works it cites.
S. Getzmann, J. Jasny, and M. Falkenstein, “Switching of auditory attention in “cocktail-party” listening: Erp evidence of cueing effects in younger and older adults,” Brain and Cognition , vol. 111, pp. 1 – 12, 2017. [Online]. Available: http://www.sciencedirect.com/science/article/pii/S0278262616302408
2017
Later among the works it cites.
Y. Liu, S. Ma, Y. Aafer, W.-C. Lee, J. Zhai, W. Wang, and X. Zhang, “Trojaning attack on neural networks,” Proceedings of the 2017 Network and Distributed System Security Symposium (NDSS) , 2017
2017
Later among the works it cites.
S. Maheshwari, “Burger King ‘O.K. Google’ Ad Doesn’t Seem O.K. With Google,” https://www.nytimes.com/2017/04/12/business/burger-king-tv-ad-google-home.html
2017
Later among the works it cites.
C. Martin, “72% Want Voice Control In Smart-Home Products,” Media Post – https://www.mediapost.com/publications/article/292253/72-want-voice-control-in-smart-home-products.html?edition=99353
2017
Later among the works it cites.
S. Nichols, “TV anchor says live on-air ‘Alexa, order me a dollhouse’- Guess what happens next,” https://www.theregister.co.uk/2017/01/07/tv-anchor-says-alexa-buy-me-a-dollhouse-and-she-does/
2017
Later among the works it cites.
2017
Later among the works it cites.
G. Zhang, C. Yan, X. Ji, T. Zhang, T. Zhang, and W. Xu, “Dolphinattack: Inaudible voice commands,” in Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security . ACM, 2017, pp. 103–117
2017
Later among the works it cites.
L. Zhang, S. Tan, and J. Yang, “Hearing your voice is not enough: An articulatory gesture based liveness detection for voice authentication,” in Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security . ACM, 2017, pp. 57–71
2017
Later among the works it cites.
L. Blue, L. Vargas, and P. Traynor, “Hello, is it me you‘re looking for? differentiating between human and electronic speakers for voice interface security,” in 11th ACM Conference on Security and Privacy in Wireless and Mobile Networks , 2018
2018
Later among the works it cites.
H. Stephenson, “UX design trends 2018: from voice interfaces to a need to not trick people,” Digital Arts - https://www.digitalartsonline.co.uk/features/interactive-design/ux-design-trends-2018-from-voice-interfaces-need-not-trick-people/
2018
Later among the works it cites.
X. Yuan, Y. Chen, Y. Zhao, Y. Long, X. Liu, K. Chen, S. Zhang, H. Huang, X. Wang, and C. A. Gunter, “Commandersong: A systematic approach for practical adversarial voice recognition,” in Proceedings of the USENIX Security Symposium , 2018
2018
Later among the works it cites.