Fetching the paper…
Reading the bibliography…
Network quantization is a powerful technique to compress convolutional neural networks.
MMDetection: Open MMLab Detection Toolbox and Benchmark
Chen, K.; Wang, J.; Pang, J.; Cao, Y.; Xiong, Y.; Li, X.; Sun, S.; Feng, W.; Liu, Z.; Xu, J.; Zhang, Z.; Cheng, D.; Zhu, C.; Cheng, T.; Zhao, Q.; Li, B.; Lu, X.; Zhu, R.; Wu, Y.; Dai, J.; Wang, J.; Shi, J.; Ouyang, W.; Loy, C. C.; and Lin, D. 2019 · 1906
Earlier work this paper cites.
Improving Post Training Neural Quantization: Layer-wise Calibration and Integer Programming
Hubara, I.; Nahshan, Y.; Hanani, Y.; Banner, R.; and Soudry, D. 2020 · 2006
Earlier work this paper cites.
EasyQuant: Post-training Quantization via Scale Optimization
Wu, D.; Tang, Q.; Zhao, Y.; Zhang, M.; Fu, Y.; and Zhang, D. 2020 · 2006
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
Lin, T.; Maire, M.; Belongie, S. J.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Russakovsky, O.; Deng, J.; Su, H.; Krause, J.; Satheesh, S.; Ma, S.; Huang, Z.; Karpathy, A.; Khosla, A.; Bernstein, M. S.; Berg, A. C.; and Li, F. 2015 · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
PACT: Parameterized Clipping Activation for Quantized Neural Networks
Choi, J.; Wang, Z.; Venkataramani, S.; Chuang, P. I.; Srinivasan, V.; and Gopalakrishnan, K. 2018 · 2018
Earlier work this paper cites.
Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference
Jacob, B.; Kligys, S.; Chen, B.; Zhu, M.; Tang, M.; Howard, A. G.; Adam, H.; and Kalenichenko, D. 2018 · 2018
Earlier work this paper cites.
Quantizing deep convolutional networks for efficient inference: A whitepaper
Krishnamoorthi, R. 2018 · 2018
Earlier work this paper cites.
YOLOv3: An Incremental Improvement
Redmon, J.; and Farhadi, A. 2018 · 2018
Cited alongside, same era.
LQ-Nets: Learned Quantization for Highly Accurate and Compact Deep Neural Networks
Zhang, D.; Yang, J.; Ye, D.; and Hua, G. 2018 · 2018
Cited alongside, same era.
Post training 4-bit quantization of convolutional networks for rapid-deployment
Banner, R.; Nahshan, Y.; and Soudry, D. 2019 · 2019
Cited alongside, same era.
Low-bit Quantization of Neural Networks for Efficient Inference
Choukroun, Y.; Kravchik, E.; Yang, F.; and Kisilev, P. 2019 · 2019
Cited alongside, same era.
Learning to Quantize Deep Networks by Optimizing Quantization Intervals With Task Loss
Jung, S.; Son, C.; Lee, S.; Son, J.; Han, J.; Kwak, Y.; Hwang, S. J.; and Choi, C. 2019 · 2019
Cited alongside, same era.
Model Compression and Hardware Acceleration for Neural Networks: A Comprehensive Survey
Deng, L.; Li, G.; Han, S.; Shi, L.; and Xie, Y. 2020 · 2020
Later among the works it cites.
Up or Down? Adaptive Rounding for Post-Training Quantization
Nagel, M.; Amjad, R. A.; van Baalen, M.; Louizos, C.; and Blankevoort, T. 2020 · 2020
Later among the works it cites.
Q-BERT: Hessian Based Ultra Low Precision Quantization of BERT
Shen, S.; Dong, Z.; Ye, J.; Ma, L.; Yao, Z.; Gholami, A.; Mahoney, M. W.; and Keutzer, K. 2020 · 2020
Later among the works it cites.
And the Bit Goes Down: Revisiting the Quantization of Neural Networks
Stock, P.; Joulin, A.; Gribonval, R.; Graham, B.; and Jégou, H. 2020 · 2020
Later among the works it cites.
A Survey of Quantization Methods for Efficient Neural Network Inference
Gholami, A.; Kim, S.; Dong, Z.; Yao, Z.; Mahoney, M. W.; and Keutzer, K. 2021 · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A Survey of ReRAM-Based Architectures for Processing-In-Memory and Neural Networks
Mittal, S. 2019 · 2019
Cited alongside, same era.
Data-Free Quantization Through Weight Equalization and Bias Correction
Nagel, M.; van Baalen, M.; Blankevoort, T.; and Welling, M. 2019 · 2019
Cited alongside, same era.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Paszke, A.; Gross, S.; Massa, F.; Lerer, A.; Bradbury, J.; Chanan, G.; Killeen, T.; Lin, Z.; Gimelshein, N.; Antiga, L.; Desmaison, A.; Kopf, A.; Yang, E.; DeVito, Z.; Raison, M.; Tejani, A.; Chilamkurthy, S.; Steiner, B.; Fang, L.; Bai, J.; and Chintala, S. 2019 · 2019
Cited alongside, same era.
Low Bit-Width Convolutional Neural Network on RRAM
Cai, Y.; Tang, T.; Xia, L.; Li, B.; Wang, Y.; and Yang, H. 2020 · 2020
Cited alongside, same era.
Proper ResNet Implementation for CIFAR10/CIFAR100 in PyTorch
Idelbayev, Y. 2021 · 2021
Closest in time.
QPP: Real-Time Quantization Parameter Prediction for Deep Neural Networks
Kryzhanovskiy, V.; Balitskiy, G.; Kozyrskiy, N.; and Zuruev, A. 2021 · 2021
Closest in time.
BRECQ: Pushing the Limit of Post-Training Quantization by Block Reconstruction
Li, Y.; Gong, R.; Tan, X.; Yang, Y.; Hu, P.; Zhang, Q.; Yu, F.; Wang, W.; and Gu, S. 2021 · 2021
Closest in time.