Fetching the paper…
Reading the bibliography…
Deep neural networks are increasingly used on mobile devices, where computational resources are limited.
Optimal brain damage
Y. LeCun, J. S. Denker, S. A. Solla, R. E. Howard, and L. D. Jackel · 1989
Earlier work this paper cites.
Optimal brain surgeon and general network pruning
B. Hassibi, D. G. Stork, and G. J. Wolff · 1993
Earlier work this paper cites.
Model compression
C. Bucilua, R. Caruana, and A. Niculescu-Mizil · 2006
Earlier work this paper cites.
Model selection and estimation in regression with grouped variables
M. Yuan and Y. Lin · 2006
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
V. Nair and G. E. Hinton · 2010
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Distilling the knowledge in a neural network
G. Hinton, O. Vinyals, and J. Dean · 2014
Earlier work this paper cites.
Network in network
M. Lin, Q. Chen, and S. Yan · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
J. T. Springenberg, A. Dosovitskiy, T. Brox, and M. Riedmiller · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
N. Srivastava, G. E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
Compressing neural networks with the hashing trick
W. Chen, J. Wilson, S. Tyree, K. Weinberger, and Y. Chen · 2015
Earlier work this paper cites.
S. Han, H. Mao, and W. J. Dally · 2015
Earlier work this paper cites.
Learning both weights and connections for efficient neural network
S. Han, J. Pool, J. Tran, and W. Dally · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Deeply-supervised nets
C.-Y. Lee, S. Xie, P. Gallagher, Z. Zhang, and Z. Tu · 2015
Cited alongside, same era.
Fitnets: Hints for thin deep nets
A. Romero, N. Ballas, S. E. Kahou, A. Chassang, C. Gatta, and Y. Bengio · 2015
Cited alongside, same era.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Cited alongside, same era.
Training very deep networks
R. K. Srivastava, K. Greff, and J. Schmidhuber · 2015
Cited alongside, same era.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Cited alongside, same era.
Pruning filters for efficient convnets
H. Li, A. Kadav, I. Durdanovic, H. Samet, and H. P. Graf · 2016
Later among the works it cites.
Xnor-net: Imagenet classification using binary convolutional neural networks
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi · 2016
Later among the works it cites.
Aggregated residual transformations for deep neural networks
S. Xie, R. Girshick, P. Dollár, Z. Tu, and K. He · 2016
Later among the works it cites.
S. Zagoruyko and N. Komodakis · 2016
Later among the works it cites.
Deep convolutional neural networks with merge-and-run mappings
L. Zhao, J. Wang, X. Li, Z. Tu, and W. Zeng · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning the number of neurons in deep networks
J. M. Alvarez and M. Salzmann · 2016
Cited alongside, same era.
Xception: Deep learning with depthwise separable convolutions
F. Chollet · 2016
Cited alongside, same era.
Spatially adaptive computation time for residual networks
M. Figurnov, M. D. Collins, Y. Zhu, L. Zhang, J. Huang, D. Vetrov, and R. Salakhutdinov · 2016
Cited alongside, same era.
Adaptive computation time for recurrent neural networks
A. Graves · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Identity mappings in deep residual networks
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Adaptive neural networks for fast test-time prediction
T. Bolukbasi, J. Wang, O. Dekel, and V. Saligrama · 2017
Closest in time.
Channel pruning for accelerating very deep neural networks
Y. He, X. Zhang, and J. Sun · 2017
Closest in time.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko, W. Wang, T. Weyand, M. Andreetto, and H. Adam · 2017
Closest in time.
Snapshot ensembles: Train 1, get m for free
G. Huang, Y. Li, G. Pleiss, Z. Liu, J. E. Hopcroft, and K. Q. Weinberger · 2017
Closest in time.
Densely connected convolutional networks
G. Huang, Z. Liu, L. van der Maaten, and K. Q. Weinberger · 2017
Closest in time.
Learning efficient convolutional networks through network slimming
Z. Liu, J. Li, Z. Shen, G. Huang, S. Yan, and C. Zhang · 2017
Closest in time.
SGDR: stochastic gradient descent with restarts
I. Loshchilov and F. Hutter · 2017
Closest in time.
Interleaved group convolutions for deep neural networks
T. Zhang, G.-J. Qi, B. Xiao, and J. Wang · 2017
Closest in time.
Shufflenet: An extremely efficient convolutional neural network for mobile devices
X. Zhang, X. Zhou, M. Lin, and J. Sun · 2017
Closest in time.
Learning transferable architectures for scalable image recognition
B. Zoph, V. Vasudevan, J. Shlens, and Q. V. Le · 2017
Closest in time.
Multi-scale dense networks for resource efficient image classification
G. Huang, D. Chen, T. Li, F. Wu, L. van der Maaten, and K. Q. Weinberger · 2018
Closest in time.