Fetching the paper…
Reading the bibliography…
We present an MatchboxNet - an end-to-end neural network for speech command recognition.
A. Waibel, T. Hanazawa, G. Hinton, K. Shirano, and K. Lang, “A time-delay neural network architecture for isolated word recognition,”
1989
Earlier work this paper cites.
K. Lang, A. Waibel, and G. Hinton, “A time-delay neural network architecture for isolated word recognition,”
1990
Earlier work this paper cites.
Y. Bengio, R. De Mori, G. Flammia, and R. Kompe, “Global optimization of a neural network-hidden Markov model hybrid,”
1992
Earlier work this paper cites.
T. Robinson, M. Hochberg, and S. Renals, “IPA: improved phone modelling with recurrent neural networks,”
1994
Earlier work this paper cites.
H. Hermansky, D. Ellis, and S. Sharma, “Tandem connectionist feature extraction for conventional hmm systems,”
2000
Earlier work this paper cites.
A. Graves, D. Eck, N. Beringer, and J. Schmidhuber, “Biologically plausible speech recognition with LSTM neural nets,” in
2004
Earlier work this paper cites.
A. Graves and J. Schmidhuber, “Framewise phoneme classification with bidirectional LSTM and other neural network architectures,”
2005
Earlier work this paper cites.
G. Hinton
2012
Earlier work this paper cites.
F. Font, G. Roma, and X. Serra, “Freesound technical demo,” in
2013
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,”
2015
Earlier work this paper cites.
T. Sainath and C. Parada, “Convolutional neural networks for small-footprint keyword spotting,” in
2015
Earlier work this paper cites.
Y. Qian and P. C. Woodland, “Very deep convolutional neural networks for robust speech recognition,” in
2016
Cited alongside, same era.
2016
Cited alongside, same era.
F. Chollet, “Xception: Deep learning with depthwise separable convolutions,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
2019
Later among the works it cites.
A. Coucke
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. DeVries and G. W. Taylor, “Improved regularization of convolutional neural networks with cutout,”
2017
Cited alongside, same era.
2017
Cited alongside, same era.
P. Warden, “Speech commands: A dataset for limited-vocabulary speech recognition,”
2018
Cited alongside, same era.
J. Tang, Y. Song, L. Dai, and I. McLoughlin, “Acoustic modeling with densely connected residual network for multichannel speech recognition,” in
2018
Cited alongside, same era.
A. Kusupati
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Won, S. Chun, O. Nieto, and X. Serra, “Data-driven harmonic filters for audio representation learning,” in
2020
Closest in time.
2020
Closest in time.