Fetching the paper…
Reading the bibliography…
Deep neural network pruning and quantization techniques have demonstrated it is possible to achieve high levels of compression with surprisingly little degradation to test set accuracy.
The state of sparsity in deep neural networks
Gale, T., Elsen, E., and Hooker, S · 1902
Earlier work this paper cites.
Towards compact and robust deep neural networks
Sehwag, V., Wang, S., Mittal, P., and Jana, S · 1906
Earlier work this paper cites.
The generalization of ‘Student’s’ problem when several different population variances are involved
Welch, B. L · 1947
Earlier work this paper cites.
A test of goodness of fit
Anderson, T. W. and Darling, D. A · 1954
Earlier work this paper cites.
Goodness-of-fit Techniques
D’Agostino, R. B. and Stephens, M. A. (eds.) · 1986
Earlier work this paper cites.
Optimal brain damage
Cun, Y. L., Denker, J. S., and Solla, S. A · 1990
Earlier work this paper cites.
Generalization by weight-elimination with application to forecasting
Weigend, A. S., Rumelhart, D. E., and Huberman, B. A · 1991
Earlier work this paper cites.
Simplifying neural networks by soft weight-sharing
Nowlan, S. J. and Hinton, G. E · 1992
Earlier work this paper cites.
Selecting typical instances in instance-based learning
Zhang, J · 1992
Earlier work this paper cites.
Optimal brain surgeon and general network pruning
Hassibi, B., Stork, D. G., and Wolff, G. J · 1993
Earlier work this paper cites.
Synaptic development of the cerebral cortex: implications for learning, memory, and mental illness
Rakic, P., Bourgeois, J.-P., and Goldman-Rakic, P. S · 1994
Earlier work this paper cites.
Sparse connection and pruning in large dynamic artificial neural networks, 1997
Ström, N · 1997
Earlier work this paper cites.
Case-based explanation for artificial neural nets
Caruana, R · 2000
Earlier work this paper cites.
Structural and functional brain development and its relation to cognitive development
Casey, B., Giedd, J. N., and Thomas, K. M · 2000
Earlier work this paper cites.
Goodness-of-Fit Tests and Model Validity
Huber-Carol, C., Balakrishnan, N., Nikulin, M., and Mesbah, M · 2002
Earlier work this paper cites.
Longitudinal mapping of cortical thickness and brain growth in normal children
Sowell, E. R., Thompson, P. M., Leonard, C. M., Welcome, S. E., Kan, E., and Toga, A. W · 2004
Earlier work this paper cites.
Classification with a reject option using a hinge loss
Bartlett, P. L. and Wegkamp, M. H · 2008
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Fundamentals of Human Neuropsychology
Kolb, B. and Whishaw, I · 2009
Earlier work this paper cites.
Improving the speed of neural networks on cpus
Vanhoucke, V., Senior, A., and Mao, M. Z · 2011
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2012
Earlier work this paper cites.
Memory Bounded Deep Convolutional Networks
Collins, M. D. and Kohli, P · 2014
Earlier work this paper cites.
Memory bounded deep convolutional networks
Collins, M. D. and Kohli, P · 2014
Earlier work this paper cites.
Training deep neural networks with low precision multiplications
Courbariaux, M., Bengio, Y., and David, J.-P · 2014
Earlier work this paper cites.
Deep learning with limited numerical precision
Gupta, S., Agrawal, A., Gopalakrishnan, K., and Narayanan, P · 2015
Cited alongside, same era.
Learning both Weights and Connections for Efficient Neural Network
Han, S., Pool, J., Tran, J., and Dally, W. J · 2015
Cited alongside, same era.
Deep Residual Learning for Image Recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Cited alongside, same era.
Distilling the Knowledge in a Neural Network
Hinton, G., Vinyals, O., and Dean, J · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C · 2015
Cited alongside, same era.
Learning Sparse Neural Networks through L _ 0 L\_0 Regularization
Louizos, C., Welling, M., and Kingma, D. P · 2017
Later among the works it cites.
Micikevicius, P., Narang, S., Alben, J., Diamos, G., Elsen, E., Garcia, D., Ginsburg, B., Houston, M., Kuchaiev, O., Venkatesh, G., and Wu, H · 2017
Later among the works it cites.
Exploring Sparsity in Recurrent Neural Networks
Narang, S., Elsen, E., Diamos, G., and Sengupta, S · 2017
Later among the works it cites.
Soft Weight-Sharing for Neural Network Compression
Ullrich, K., Meeds, E., and Welling, M · 2017
Later among the works it cites.
To prune, or not to prune: exploring the efficacy of pruning for model compression
Zhu, M. and Gupta, S · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep learning face attributes in the wild
Liu, Z., Luo, P., Wang, X., and Tang, X · 2015
Cited alongside, same era.
On the efficient representation and execution of deep acoustic models
Alvarez, R., Prabhavalkar, R., and Bakhtin, A · 2016
Cited alongside, same era.
An Analysis of Deep Neural Network Models for Practical Applications
Canziani, A., Paszke, A., and Culurciello, E · 2016
Cited alongside, same era.
Eyeriss: A spatial architecture for energy-efficient dataflow for convolutional neural networks
Chen, Y., Emer, J., and Sze, V · 2016
Cited alongside, same era.
Boosting with abstention
Cortes, C., DeSalvo, G., and Mohri, M · 2016
Cited alongside, same era.
Dynamic network surgery for efficient dnns
Guo, Y., Yao, A., and Chen, Y · 2016
Cited alongside, same era.
Quantized neural networks: Training neural networks with low precision weights and activations
Hubara, I., Courbariaux, M., Soudry, D., El-Yaniv, R., and Bengio, Y · 2016
Cited alongside, same era.
Morphnet: Fast & simple resource-constrained structure learning of deep networks
Gordon, A., Eban, E., Nachum, O., Chen, B., Wu, H., Yang, T.-J., and Choi, E · 2018
Later among the works it cites.
Sparse dnns with improved adversarial robustness
Guo, Y., Zhang, C., Zhang, C., and Chen, Y · 2018
Later among the works it cites.
Quantization and training of neural networks for efficient integer-arithmetic-only inference
Jacob, B., Kligys, S., Chen, B., Zhu, M., Tang, M., Howard, A., Adam, H., and Kalenichenko, D · 2018
Later among the works it cites.
Efficient Neural Audio Synthesis
Kalchbrenner, N., Elsen, E., Simonyan, K., Noury, S., Casagrande, N., Lockhart, E., Stimberg, F., van den Oord, A., Dieleman, S., and Kavukcuoglu, K · 2018
Later among the works it cites.
SNIP: single-shot network pruning based on connection sensitivity
Lee, N., Ajanthan, T., and Torr, P. H. S · 2018
Later among the works it cites.
Evolutionary pruning of transfer learned deep convolutional neural network for breast cancer diagnosis in digital breast tomosynthesis
Samala, R. K., Chan, H.-P., Hadjiiski, L. M., Helvie, M. A., Richter, C., and Cha, K · 2018
Later among the works it cites.
Faster gaze prediction with dense networks and Fisher pruning
Theis, L., Korshunova, I., Tejani, A., and Huszár, F · 2018
Later among the works it cites.
Lpcnet: Improving Neural Speech Synthesis Through Linear Prediction
Valin, J. and Skoglund, J · 2018
Later among the works it cites.
Variable generalization performance of a deep learning model to detect pneumonia in chest radiographs: A cross-sectional study
Zech, J. R., Badgeley, M. A., Liu, M., Costa, A. B., Titano, J. J., and Oermann, E. K · 2018
Later among the works it cites.
Rigging the lottery: Making all tickets winners, 2019
Evci, U., Gale, T., Menick, J., Castro, P. S., and Elsen, E · 2019
Closest in time.
Benchmarking neural network robustness to common corruptions and perturbations
Hendrycks, D. and Dietterich, T · 2019
Closest in time.
Hendrycks, D., Zhao, K., Basart, S., Steinhardt, J., and Song, D · 2019
Closest in time.
A benchmark for interpretability methods in deep neural networks
Hooker, S., Erhan, D., Kindermans, P.-J., and Kim, B · 2019
Closest in time.
Energy and Policy Considerations for Deep Learning in NLP
Strubell, E., Ganesh, A., and McCallum, A · 2019
Closest in time.
Non-vacuous generalization bounds at the imagenet scale: a pac-bayesian compression approach
Zhou, W., Veitch, V., Austern, M., Adams, R. P., and Orbanz, P · 2019
Closest in time.
What Neural Networks Memorize and Why: Discovering the Long Tail via Influence Estimation
Feldman, V. and Zhang, C · 2020
Closest in time.
Mlir: A compiler infrastructure for the end of moore’s law, 2020
Lattner, C., Amini, M., Bondhugula, U., Cohen, A., Davis, A., Pienaar, J., Riddle, R., Shpeisman, T., Vasilache, N., and Zinenko, O · 2020
Closest in time.
Train large, then compress: Rethinking model size for efficient training and inference of transformers, 2020
Li, Z., Wallace, E., Shen, S., Lin, K., Keutzer, K., Klein, D., and Gonzalez, J. E · 2020
Closest in time.
Keep the gradients flowing: Using gradient flow to study sparse network optimization
Tessera, K., Hooker, S., and Rosman, B · 2021
Closest in time.