Fetching the paper…
Reading the bibliography…
The enormous size of modern deep neural networks makes it challenging to deploy those models in memory and communication limited scenarios.
Coding theorems for a discrete source with a fidelity criterion
Shannon, C. E. (1959) · 1959
Earlier work this paper cites.
Information rates of gaussian signals under criteria constraining the error spectrum
McDonald, R. and Schultheiss, P. (1964) · 1964
Earlier work this paper cites.
Rate distortion theory: A mathematical basis for data compression
Berger, T. (1971) · 1971
Earlier work this paper cites.
Comparing biases for minimal network construction with back-propagation
Hanson, S. J. and Pratt, L. Y. (1989) · 1989
Earlier work this paper cites.
Optimal brain damage
LeCun, Y., Denker, J. S., and Solla, S. A. (1990) · 1990
Earlier work this paper cites.
Second order derivatives for network pruning: Optimal brain surgeon
Hassibi, B. and Stork, D. G. (1993) · 1993
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P. (1998) · 1998
Earlier work this paper cites.
Improving the speed of neural networks on cpus
Vanhoucke, V., Senior, A., and Mao, M. Z. (2011) · 2011
Earlier work this paper cites.
Elements of information theory
Cover, T. M. and Thomas, J. A. (2012) · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E. (2012) · 2012
Earlier work this paper cites.
Estimating the hessian by back-propagating curvature
Martens, J., Sutskever, I., and Swersky, K. (2012) · 2012
Earlier work this paper cites.
Deep learning with cots hpc systems
Coates, A., Huval, B., Wang, T., Wu, D., Catanzaro, B., and Andrew, N. (2013) · 2013
Cited alongside, same era.
Exploiting linear structure within convolutional networks for efficient evaluation
Denton, E. L., Zaremba, W., Bruna, J., LeCun, Y., and Fergus, R. (2014) · 2014
Cited alongside, same era.
Compressing deep convolutional networks using vector quantization
Gong, Y., Liu, L., Yang, M., and Bourdev, L. (2014) · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A. (2014) · 2014
Cited alongside, same era.
Compressing neural networks with the hashing trick
Chen, W., Wilson, J., Tyree, S., Weinberger, K., and Chen, Y. (2015) · 2015
Cited alongside, same era.
Learning one-hidden-layer neural networks with landscape design
Ge, R., Lee, J. D., and Ma, T. (2017) · 2017
Later among the works it cites.
On calibration of modern neural networks
Guo, C., Pleiss, G., Sun, Y., and Weinberger, K. Q. (2017) · 2017
Later among the works it cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Howard, A. G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., and Adam, H. (2017) · 2017
Later among the works it cites.
The nearest neighbor information estimator is adaptively near minimax rate-optimal
Jiao, J., Gao, W., and Han, Y. (2017) · 2017
Later among the works it cites.
Bayesian compression for deep learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
An exploration of parameter redundancy in deep networks with circulant projections
Cheng, Y., Yu, F. X., Feris, R. S., Kumar, S., Choudhary, A., and Chang, S.-F. (2015) · 2015
Cited alongside, same era.
Towards the limit of network quantization
Choi, Y., El-Khamy, M., and Lee, J. (2016) · 2016
Cited alongside, same era.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and¡ 0.5 mb model size
Iandola, F. N., Han, S., Moskewicz, M. W., Ashraf, K., Dally, W. J., and Keutzer, K. (2016) · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., et al. (2016) · 2016
Cited alongside, same era.
Federici, M., Ullrich, K., and Welling, M. (2017) · 2017
Cited alongside, same era.
Han, S., Mao, H., and Dally, W. J. (2015a)
Cited in the paper.
Learning both weights and connections for efficient neural network
Han, S., Pool, J., Tran, J., and Dally, W. (2015b)
Cited in the paper.
Louizos, C., Ullrich, K., and Welling, M. (2017) · 2017
Later among the works it cites.
Stochastic gradient descent as approximate bayesian inference
Mandt, S., Hoffman, M. D., and Blei, D. M. (2017) · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al. (2017) · 2017
Later among the works it cites.
Soft weight-sharing for neural network compression
Ullrich, K., Meeds, E., and Welling, M. (2017) · 2017
Later among the works it cites.
A review on neural networks with random weights
Cao, W., Wang, X., Ming, Z., and Gao, J. (2018) · 2018
Closest in time.
Amc: Automl for model compression and acceleration on mobile devices
He, Y., Lin, J., Liu, Z., Wang, H., Li, L.-J., and Han, S. (2018) · 2018
Closest in time.