The state of sparsity in deep neural networks
Original
Gale, T., Elsen, E., and Hooker, S · 1902
Earlier work this paper cites.
The lottery ticket hypothesis at scale
Original
Frankle, J., Dziugaite, G. K., Roy, D. M., and Carbin, M · 1903
Earlier work this paper cites.
Deconstructing lottery tickets: Zeros, signs, and the supermask
Original
Zhou, H., Lan, J., Liu, R., and Yosinski, J · 1905
Earlier work this paper cites.
The difficulty of training sparse neural networks
Original
Evci, U., Pedregosa, F., Gomez, A. N., and Elsen, E · 1906
Earlier work this paper cites.
Discovering neural wirings
Original
Wortsman, M., Farhadi, A., and Rastegari, M · 1906
Earlier work this paper cites.
Sparse networks from scratch: Faster training without losing performance
Original
Dettmers, T. and Zettlemoyer, L · 1907
Earlier work this paper cites.
Fast sparse convnets
Original
Elsen, E., Dukhan, M., Gale, T., and Simonyan, K · 1911
Earlier work this paper cites.
Skeletonization: A technique for trimming the fat from a network via relevance assessment
Mozer, M. C. and Smolensky, P · 1989
Earlier work this paper cites.
Optimal Brain Damage
LeCun, Y., Denker, J. S., and Solla, S. A · 1990
Earlier work this paper cites.
Second order derivatives for network pruning: Optimal Brain Surgeon
Hassibi, B. and Stork, D · 1993
Earlier work this paper cites.
Evaluating pruning methods
Thimm, G. and Fiesler, E · 1995
Earlier work this paper cites.
Sparse Connection and Pruning in Large Dynamic Artificial Neural Networks
Ström, N · 1997
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2009
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., van Merrienboer, B., Gulcehre, C., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M., Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., and Zheng, X · 2015
Earlier work this paper cites.
Learning both weights and connections for efficient neural network
Han, S., Pool, J., Tran, J., and Dally, W · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., Berg, A. C., and Fei-Fei, L · 2015
Earlier work this paper cites.