Fetching the paper…
Reading the bibliography…
Large DNNs with mixed-precision quantization can achieve ultra-high compression while retaining high classification performance.
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
I. Hubara, M. Courbariaux, D. Soudry, R. El-Yaniv, and Y. Bengio, “Binarized neural networks,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi, “XNOR-Net: Imagenet classification using binary convolutional neural networks,” in European conference on computer vision . Springer, 2016, pp. 525–542
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
F. Li, B. Zhang, and B. Liu, “Ternary weight networks,” arXiv preprint arXiv:1605.04711 , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
A. Inc., “An on-device deep neural network for face detection,” Nov 2017. [Online]. Available: https://machinelearning.apple.com/research/face-detection
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
2018
Cited alongside, same era.
D. Zhang, J. Yang, D. Ye, and G. Hua, “LQ-Nets: Learned quantization for highly accurate and compact deep neural networks,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 365–382
2018
Cited alongside, same era.
2018
Cited alongside, same era.
A. S. Rakin, Z. He, and D. Fan, “Bit-flip attack: Crushing neural network with progressive bit search,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 1211–1220
2019
Later among the works it cites.
L. Deng, G. Li, S. Han, L. Shi, and Y. Xie, “Model compression and hardware acceleration for neural networks: A comprehensive survey,” Proceedings of the IEEE , vol. 108, no. 4, pp. 485–532, 2020
2020
Later among the works it cites.
S. Kundu, M. Nazemi, M. Pedram, K. M. Chugg, and P. A. Beerel, “Pre-defined sparsity for low-complexity convolutional neural networks,” IEEE Transactions on Computers , vol. 69, no. 7, pp. 1045–1058, 2020
2020
Later among the works it cites.
S. Kundu, M. Nazemi, P. A. Beerel, and M. Pedram, “DNR: A tunable robust pruning framework through dynamic network rewiring of DNNs,” in Proceedings of the 26th Asia and South Pacific Design Automation Conference , 2021, pp. 344–350
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
K. Wang, Z. Liu, Y. Lin, J. Lin, and S. Han, “HAQ: Hardware-aware automated quantization with mixed precision,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 8612–8620
2019
Cited alongside, same era.
Z. Dong, Z. Yao, A. Gholami, M. W. Mahoney, and K. Keutzer, “HAWQ: Hessian aware quantization of neural networks with mixed-precision,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 293–302
2019
Cited alongside, same era.
J. Choi, S. Venkataramani, V. Srinivasan, K. Gopalakrishnan, Z. Wang, and P. Chuang, “Accurate and efficient 2-bit quantized neural networks.” in MLSys , 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
S. Kundu, Q. Sun, Y. Fu, M. Pedram, and P. A. Beerel, “Analyzing the confidentiality of undistillable teachers in knowledge distillation,” Advances in Neural Information Processing Systems (NeurIPS 2021) , vol. 34, 2021
2021
Closest in time.
2021
Closest in time.
Z. Yao, Z. Dong, Z. Zheng, A. Gholami, J. Yu, E. Tan, L. Wang, Q. Huang, Y. Wang, M. Mahoney et al. , “HAWQ-V3: Dyadic neural network quantization,” in International Conference on Machine Learning . PMLR, 2021, pp. 11 875–11 886
2021
Closest in time.
K. Vasquez, Y. Venkatesha, A. Bhattacharjee, A. Moitra, and P. Panda, “Activation density based mixed-precision quantization for energy efficient neural networks,” DATE , 2021
2021
Closest in time.