Fetching the paper…
Reading the bibliography…
The intensive computation of Automatic Speech Recognition (ASR) models obstructs them from being deployed on mobile devices.
A. Graves, S. Fernández, F. J. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
J. Cohen, “Embedded speech recognition applications in mobile phones: Status, trends, and challenges,” in
2008
Earlier work this paper cites.
A. Mohamed, G. E. Dahl, and G. E. Hinton, “Acoustic modeling using deep belief networks,”
2012
Earlier work this paper cites.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in
2015
Earlier work this paper cites.
V. Peddinti, D. Povey, and S. Khudanpur, “A time delay neural network architecture for efficient modeling of long temporal contexts,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
R. C., C. P., and G. S., “Wav2letter: an end-to-end convnet-based speech recognition system,”
2016
Earlier work this paper cites.
L. A. and G. S., “Fast algorithms for convolutional neural networks,” in
2016
Earlier work this paper cites.
G. inc., “Tensorflow lite: an open source deep learning framework for on-device inference,”
2017
Earlier work this paper cites.
T. inc., “Ncnn: high-performance neural network inference computing framework optimized for mobile platforms,”
2017
Cited alongside, same era.
H. Bu, J. Du, X. Na, B. Wu, and H. Zheng, “Aishell-1: An open-source mandarin speech corpus and a speech recognition baseline,” in
2017
Cited alongside, same era.
J. Wang, B. Cao, P. Yu, L. Sun, W. Bao, and X. Zhu, “Deep learning towards mobile applications,” in
2018
Cited alongside, same era.
S. Zhang, M. Lei, Z. Yan, and L. Dai, “Deep-fsmn for large vocabulary continuous speech recognition,” in
2018
Cited alongside, same era.
L. Dong, S. Xu, and B. Xu, “Speech-transformer: a no-recurrence sequence-to-sequence model for speech recognition,” in
2018
Cited alongside, same era.
Z. Gao, S. Zhang, M. Lei1, and I. McLoughlin2, “San-m: Memory equipped self-attention for end-to-end speech recognition,”
2020
Closest in time.
S. K. Esser, J. L. McKinstry, D. Bablani, R. Appuswamy, and D. S. Modha, “Learned step size quantization,” in
2020
Closest in time.
X. Jiang, H. Wang, Y. Chen, Z. Wu, L. Wang, B. Zou, Y. Yang, Z. Cui, Y. Cai, T. Yu, C. Lv, and Z. Wu, “Mnn: A universal and efficient inference engine,” in
2020
Closest in time.
J. Fernandez-Marques, P. N. Whatmough, A. Mundy, and M. Mattina, “Searching for winograd-aware quantized networks,” in
2020
Closest in time.
G. Li, L. Liu, X. Wang, X. Ma, and X. Feng, “Lance: efficient low-precision quantized winograd convolution for neural networks based on graphics processing units,” in
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Krishnamoorthi, “Quantizing deep convolutional networks for efficient inference: A whitepaper,”
2018
Cited alongside, same era.
J. C. et al., “Pact: Parameterized clipping activation for quantized neural networks,” in
2018
Cited alongside, same era.
M. Nagel, M. van Baalen, T. Blankevoort, and M. Welling, “Data-free quantization through weight equalization and bias correction,” in
2019
Cited alongside, same era.
M. Cheng, C. Wang, X. Hu, J. Huang, and X. Wang, “Weakly supervised construction of ASR systems with massive video data,”
2020
Closest in time.
C. Wang, M. Cheng, X. Hu, and J. Huang, “Easyasr: A distributed machine learning platform for end-to-end automatic speech recognition,”
2020
Closest in time.
N. inc., “Tensorrt: programmable inference accelerator,”
2020
Closest in time.