Fetching the paper…
Reading the bibliography…
The number of parameters in deep neural networks (DNNs) is rapidly increasing to support complicated tasks and to improve model accuracy.
Pearson Education India, 1974
A. V. Aho and J. E. Hopcroft, The design and analysis of computer algorithms · 1974
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” nature
1986
Earlier work this paper cites.
J. Von Neumann, “First draft of a report on the EDVAC,” IEEE Annals of the History of Computing
1993
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation
1997
Earlier work this paper cites.
C. Nvidia, “cuBLAS library,” NVIDIA Corporation, Santa Clara, California
2008
Earlier work this paper cites.
V. Volkov and J. W. Demmel, “Benchmarking GPUs to tune dense linear algebra,” in SC’08: Proceedings of the 2008 ACM/IEEE conference on Supercomputing
2008
Earlier work this paper cites.
G. Guennebaud, B. Jacob, et al
2010
Earlier work this paper cites.
Elsevier, 2011
J. L. Hennessy and D. A. Patterson, Computer architecture: a quantitative approach · 2011
Earlier work this paper cites.
T. N. Sainath, B. Kingsbury, V. Sindhwani, E. Arisoy, and B. Ramabhadran, “Low-rank matrix factorization for deep neural network training with high-dimensional output targets,” in 2013 IEEE international conference on acoustics, speech and signal processing
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
E. Wang, Q. Zhang, B. Shen, G. Zhang, X. Lu, Q. Wu, and Y. Wang, “Intel math kernel library,” in High-Performance Computing on the Intel® Xeon Phi™
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Y. Li, Y. Liang, and A. Risteski, “Recovery guarantee of weighted low-rank approximation via alternating minimization,” in International Conference on Machine Learning
2016
Earlier work this paper cites.
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi, “XNOR-Net: Imagenet classification using binary convolutional neural networks,” in European conference on computer vision
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
Y. Guo, A. Yao, H. Zhao, and Y. Chen, “Network sketching: Exploiting binary structure in deep CNNs,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in neural information processing systems
2017
Cited alongside, same era.
2017
Cited alongside, same era.
D. He, Z. Wang, and J. Liu, “A survey to predict the trend of AI-able server evolution in the cloud,” IEEE Access
2018
Cited alongside, same era.
E. Eleftheriou, ““in-memory computing”: Accelerating ai applications,” in 2018 48th European Solid-State Device Research Conference (ESSDERC)
2018
Cited alongside, same era.
2019
Later among the works it cites.
Z. Zhou, X. Chen, E. Li, L. Zeng, K. Luo, and J. Zhang, “Edge intelligence: Paving the last mile of artificial intelligence with edge computing,” Proceedings of the IEEE
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
D. Zhang, J. Yang, D. Ye, and G. Hua, “LQ-Nets: Learned quantization for highly accurate and compact deep neural networks,” in Proceedings of the European conference on computer vision (ECCV)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. R. Salakhutdinov, and Q. V. Le, “XLNet: Generalized autoregressive pretraining for language understanding,” in Advances in neural information processing systems
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
S. Karita, N. E. Y. Soplin, S. Watanabe, M. Delcroix, A. Ogawa, and T. Nakatani, “Improving Transformer-based end-to-end speech recognition with connectionist temporal classification and language model integration,” Proc. Interspeech 2019
2019
Later among the works it cites.
K. Irie, R. Prabhavalkar, A. Kannan, A. Bruguier, D. Rybach, and P. Nguyen, “On the choice of modeling unit for sequence-to-sequence speech recognition,” Proc. Interspeech 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
T. Choudhary, V. Mishra, A. Goswami, and J. Sarangapani, “A comprehensive survey on model compression and acceleration,” Artificial Intelligence Review
2020
Closest in time.