Fetching the paper…
Reading the bibliography…
We explore the application of end-to-end stateless temporal modeling to small-footprint keyword spotting as opposed to recurrent networks that model long-term temporal dependencies using internal states.
“A hidden markov model based keyword recognition system,”
Richard C Rose and Douglas B Paul, · 1990
Earlier work this paper cites.
“Automatic recognition of keywords in unconstrained speech using hidden markov models,”
Jay G Wilpon, Lawrence R Rabiner, C-H Lee, and ER Goldman, · 1990
Earlier work this paper cites.
“Improvements and applications for key word recognition using hidden markov modeling techniques,”
JG Wilpon, LG Miller, and P Modi, · 1991
Earlier work this paper cites.
“An application of recurrent neural networks to discriminative keyword spotting,”
Santiago Fernández, Alex Graves, and Jürgen Schmidhuber, · 2007
Earlier work this paper cites.
“Understanding the difficulty of training deep feedforward neural networks,”
Xavier Glorot and Yoshua Bengio, · 2010
Earlier work this paper cites.
“Small-footprint keyword spotting using deep neural networks,”
Guoguo Chen, Carolina Parada, and Georg Heigold, · 2014
Cited alongside, same era.
“Online word-spotting in continuous speech with recurrent neural networks,”
Pallavi Baljekar, Jill Fain Lehman, and Rita Singh, · 2014
Cited alongside, same era.
“Convolutional neural networks for small-footprint keyword spotting,”
Tara N Sainath and Carolina Parada, · 2015
Cited alongside, same era.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Cited alongside, same era.
“Max-pooling loss training of long short-term memory networks for small-footprint keyword spotting,”
Ming Sun, Anirudh Raju, George Tucker, Sankaran Panchapagesan, Gengshen Fu, Arindam Mandal, Spyros Matsoukas, Nikko Strom, and Shiv Vitaladevuni, · 2016
Cited alongside, same era.
“Wavenet: A generative model for raw audio.,”
Aäron Van Den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W Senior, and Koray Kavukcuoglu, · 2016
Later among the works it cites.
“Hello edge: Keyword spotting on microcontrollers,”
Yundong Zhang, Naveen Suda, Liangzhen Lai, and Vikas Chandra, · 2017
Later among the works it cites.
“Deep residual learning for small-footprint keyword spotting,”
Raphael Tang and Jimmy Lin, · 2017
Later among the works it cites.
“Temporal modeling using dilated convolution and gating for voice-activity-detection,”
Shuo-Yiin Chang, Bo Li, Gabor Simko, Tara N Sainath, Anshuman Tripathi, Aäron van den Oord, and Oriol Vinyals, · 2018
Closest in time.
“Speech commands: A dataset for limited-vocabulary speech recognition,”
Pete Warden, · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Closest in time.