Fetching the paper…
Reading the bibliography…
Recent work in network quantization produced state-of-the-art results using mixed precision quantization.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky, G. Hinton, et al · 2009
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Y. Bengio, N. Léonard, and A. Courville · 2013
Earlier work this paper cites.
S. Han, H. Mao, and W. J. Dally · 2015
Earlier work this paper cites.
Learning both weights and connections for efficient neural network
S. Han, J. Pool, J. Tran, and W. Dally · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Ssd: Single shot multibox detector
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, and A. C. Berg · 2016
Earlier work this paper cites.
Convolutional neural networks using logarithmic data representation
D. Miyashita, E. H. Lee, and B. Murmann · 2016
Earlier work this paper cites.
Deep learning with low precision by half-wave gaussian quantization
Z. Cai, X. He, J. Sun, and N. Vasconcelos · 2017
Earlier work this paper cites.
Learning efficient object detection models with knowledge distillation
G. Chen, W. Choi, X. Yu, T. Han, and M. Chandraker · 2017
Earlier work this paper cites.
Learning accurate low-bit deep neural networks with stochastic quantization
Y. Dong, R. Ni, J. Li, Y. Chen, J. Zhu, and H. Su · 2017
Earlier work this paper cites.
Channel pruning for accelerating very deep neural networks
Y. He, X. Zhang, and J. Sun · 2017
Earlier work this paper cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko, W. Wang, T. Weyand, M. Andreetto, and H. Adam · 2017
Earlier work this paper cites.
Categorical reparametrization with gumble-softmax
E. Jang, S. Gu, and B. Poole · 2017
Earlier work this paper cites.
Learning efficient convolutional networks through network slimming
Z. Liu, J. Li, Z. Shen, G. Huang, S. Yan, and C. Zhang · 2017
Cited alongside, same era.
The concrete distribution: A continuous relaxation of discrete random variables
C. J. Maddison, A. Mnih, and Y. W. Teh · 2017
Cited alongside, same era.
Pruning convolutional neural networks for resource efficient inference
P. Molchanov, S. Tyree, T. Karras, T. Aila, and J. Kautz · 2017
Cited alongside, same era.
Incremental network quantization: Towards lossless cnns with low-precision weights
A. Zhou, A. Yao, Y. Guo, L. Xu, and Y. Chen · 2017
Cited alongside, same era.
Nice: Noise injection and clamping estimation for neural network quantization
C. Baskin, N. Liss, Y. Chai, E. Zheltonozhskii, E. Schwartz, R. Giryes, A. Mendelson, and A. M. Bronstein · 2018
Cited alongside, same era.
Shufflenet: An extremely efficient convolutional neural network for mobile devices
X. Zhang, X. Zhou, M. Lin, and J. Sun · 2018
Later among the works it cites.
Post training 4-bit quantization of convolutional networks for rapid-deployment
R. Banner, Y. Nahshan, and D. Soudry · 2019
Later among the works it cites.
Hawq-v2: Hessian aware trace-weighted quantization of neural networks
Z. Dong, Z. Yao, Y. Cai, D. Arfeen, A. Gholami, M. W. Mahoney, and K. Keutzer · 2019
Later among the works it cites.
Hawq: Hessian aware quantization of neural networks with mixed-precision
Z. Dong, Z. Yao, A. Gholami, M. W. Mahoney, and K. Keutzer · 2019
Later among the works it cites.
Learned step size quantization
S. K. Esser, J. L. McKinstry, D. Bablani, R. Appuswamy, and D. S. Modha · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Baskin, E. Schwartz, E. Zheltonozhskii, N. Liss, R. Giryes, A. M. Bronstein, and A. Mendelson · 2018
Cited alongside, same era.
Pact: Parameterized clipping activation for quantized neural networks
J. Choi, Z. Wang, S. Venkataramani, P. I.-J. Chuang, V. Srinivasan, and K. Gopalakrishnan · 2018
Cited alongside, same era.
Squeezenext: Hardware-aware neural network design
A. Gholami, K. Kwon, B. Wu, Z. Tai, X. Yue, P. Jin, S. Zhao, and K. Keutzer · 2018
Cited alongside, same era.
Squeeze-and-excitation networks
J. Hu, L. Shen, and G. Sun · 2018
Cited alongside, same era.
Quantization and training of neural networks for efficient integer-arithmetic-only inference
B. Jacob, S. Kligys, B. Chen, M. Zhu, M. Tang, A. Howard, H. Adam, and D. Kalenichenko · 2018
Cited alongside, same era.
Model compression via distillation and quantization
A. Polino, R. Pascanu, and D. Alistarh · 2018
Cited alongside, same era.
Mobilenetv2: Inverted residuals and linear bottlenecks
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen · 2018
Cited alongside, same era.
Filter pruning via geometric median for deep convolutional neural networks acceleration
Y. He, P. Liu, Z. Wang, Z. Hu, and Y. Yang · 2019
Later among the works it cites.
Searching for mobilenetv3
A. Howard, M. Sandler, G. Chu, L.-C. Chen, B. Chen, M. Tan, W. Wang, Y. Zhu, R. Pang, V. Vasudevan, et al · 2019
Later among the works it cites.
S. R. Jain, A. Gural, M. Wu, and C. Dick · 2019
Later among the works it cites.
DARTS: Differentiable architecture search
H. Liu, K. Simonyan, and Y. Yang · 2019
Later among the works it cites.
Efficientnet: Rethinking model scaling for convolutional neural networks
M. Tan and Q. Le · 2019
Later among the works it cites.
Efficientdet: Scalable and efficient object detection
M. Tan, R. Pang, and Q. V. Le · 2019
Later among the works it cites.
Haq: Hardware-aware automated quantization with mixed precision
K. Wang, Z. Liu, Y. Lin, J. Lin, and S. Han · 2019
Later among the works it cites.
Fbnet: Hardware-aware efficient convnet design via differentiable neural architecture search
B. Wu, X. Dai, P. Zhang, Y. Wang, F. Sun, Y. Wu, Y. Tian, P. Vajda, Y. Jia, and K. Keutzer · 2019
Later among the works it cites.
Zeroq: A novel zero shot quantization framework
Y. Cai, Z. Yao, Z. Dong, A. Gholami, M. W. Mahoney, and K. Keutzer · 2020
Closest in time.
On the variance of the adaptive learning rate and beyond
L. Liu, H. Jiang, P. He, W. Chen, X. Liu, J. Gao, and J. Han · 2020
Closest in time.
Mixed precision dnns: All you need is a good parametrization
S. Uhlich, L. Mauch, F. Cardinaux, K. Yoshiyama, J. A. Garcia, S. Tiedemann, T. Kemp, and A. Nakamura · 2020
Closest in time.