Fetching the paper…
Reading the bibliography…
This paper proposes a Sub-band Convolutional Neural Network for spoken term classification.
O. Abdel-Hamid, A. Mohamed, H. Jiang, and G. Penn, “Applying convolutional neural networks concepts to hybrid nn-hmm model for speech recognition,” in
2012
Earlier work this paper cites.
O. Abdel-Hamid, L. Deng, and D. Yu, “Exploring convolutional neural network structures and optimization techniques for speech recognition,” in
2013
Earlier work this paper cites.
G. Chen, C. Parada, and G. Heigold, “Small-footprint keyword spotting using deep neural networks,” in
2014
Earlier work this paper cites.
T. N. Sainath and C. Parada, “Convolutional neural networks for small-footprint keyword spotting,” in
2015
Earlier work this paper cites.
M. Abadi, A. Agarwal et al., “TensorFlow: Large-scale machine learning on heterogeneous systems,” 2015, software available from tensorflow.org. [Online]. Available: https://www.tensorflow.org/
2015
Earlier work this paper cites.
Y. Qian, M. Bi, T. Tan, and K. Yu, “Very deep convolutional neural networks for noise robust speech recognition,”
2016
Earlier work this paper cites.
N. Takahashi, M. Gygli, B. Pfister, and L. V. Gool, “Deep convolutional neural networks and data augmentation for acoustic event recognition,” in
2016
Earlier work this paper cites.
G. Tucker, M. Wu, M. Sun, S. Panchapagesan, G. Fu, and S. Vitaladevuni, “Model compression applied to small-footprint keyword spotting,” in
2016
Earlier work this paper cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in
2017
Earlier work this paper cites.
J. F. Gemmeke, D. P. W. Ellis, D. Freedman, A. Jansen, W. Lawrence, R. C. Moore, M. Plakal, and M. Ritter, “Audio set: An ontology and human-labeled dataset for audio events,” in
2017
Cited alongside, same era.
A. Mesaros, T. Heittola, A. Diment, B. Elizalde, A. Shah, E. Vincent, B. Raj, and T. Virtanen, “DCASE 2017 challenge setup: Tasks, datasets and baseline system,” in
2017
Cited alongside, same era.
S. Hershey, S. Chaudhuri, D. P. W. Ellis, J. F. Gemmeke, A. Jansen, R. C. Moore, M. Plakal, D. Platt, R. A. Saurous, B. Seybold, M. Slaney, R. J. Weiss, and K. Wilson, “Cnn architectures for large-scale audio classification,” in
2017
Cited alongside, same era.
H. Lim, J. Park, and Y. Han, “Rare sound event detection using 1D convolutional recurrent neural networks,” DCASE2017 Challenge, Tech. Rep., September 2017
2017
Cited alongside, same era.
P. Warden, “Speech commands: A dataset for limited-vocabulary speech recognition,”
2018
Later among the works it cites.
J. S. Chung, A. Nagrani, and A. Zisserman, “Voxceleb2: Deep speaker recognition,” in
2018
Later among the works it cites.
C. Kao, W. Wang, M. Sun, and C. Wang, “R-CRNN: region-based convolutional recurrent neural network for audio event detection,” in
2018
Later among the works it cites.
R. Tang and J. Lin, “Deep residual learning for small-footprint keyword spotting,” in
2018
Later among the works it cites.
R. Pang, T. Sainath, R. Prabhavalkar, S. Gupta, Y. Wu, S. Zhang, and C.-C. Chiu, “Compression of end-to-end models,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
Y. He, R. Prabhavalkar, K. Rao, W. Li, A. Bakhtin, and I. McGraw, “Streaming small-footprint keyword spotting using sequence-to-sequence models,” in
2017
Cited alongside, same era.
S. O. Arik, M. Kliegl, R. Child, J. Hestness, A. Gibiansky, C. Fougner, R. Prenger, and A. Coates, “Convolutional recurrent neural networks for small-footprint keyword spotting,” in
2017
Cited alongside, same era.
M. Sun, D. Snyder, Y. Gao, V. Nagaraja, M. Rodehorst, S. Panchapagesan, N. Strom, S. Matsoukas, and S. Vitaladevuni, “Compressed time delay neural network for small-footprint keyword spotting,” in
2017
Cited alongside, same era.
L. Lu, M. Guo, and S. Renals, “Knowledge distillation for small-footprint highway networks,” in
2017
Cited alongside, same era.
“TensorFlow: Simple audio recognition.” [Online]. Available: https://www.tensorflow.org/tutorials/sequences/audio_recognition/
Cited in the paper.
B. Shi, M. Sun, C. Kao, V. Rozgic, S. Matsoukas, and C. Wang, “Semi-supervised acoustic event detection based on tri-training,” in
2019
Closest in time.
Q. Tang, M. Sun, C. Kao, V. Rozgic, and C. Wang, “Hierarchical residual-pyramidal model for large context based media presence detection,” in
2019
Closest in time.
B. Shi, M. Sun, C. Kao, V. Rozgic, S. Matsoukas, and C. Wang, “Compression of acoustic event detection models with quantized distillation,” to appear in INTERSPEECH, 2019
2019
Closest in time.
S. S. R. Phaye, E. Benetos, and Y. Wang, “Subspectralnet - using sub-spectrogram based convolutional neural networks for acoustic scene classification,” in
2019
Closest in time.