Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference
Jacob, B., Kligys, S., Chen, B., Zhu, M., Tang, M., Howard, A., Adam, H., and Kalenichenko, D · 2018
Cited alongside, same era.
Discovering Low-Precision Networks Close to Full-Precision Networks for Efficient Embedded Inference
Original
McKinstry, J. L., Esser, S. K., Appuswamy, R., Bablani, D., Arthur, J. V., Yildiz, I. B., and Modha, D. S · 2018
Cited alongside, same era.
Scaling for Edge Inference of Deep Neural Networks
Xu, X., Ding, Y., Hu, S. X., Niemier, M., Cong, J., Hu, Y., and Shi, Y · 2018
Cited alongside, same era.
Seernet: Predicting convolutional neural network feature-map sparsity through low-bit quantization
Cao, S., Ma, L., Xiao, W., Zhang, C., Liu, Y., Zhang, L., Nie, L., and Yang, Z · 2019
Cited alongside, same era.
Same, Same but Different - Recovering Neural Network Quantization Error through Weight Factorization
Meller, E., Finkelstein, A., Almog, U., and Grobman, M · 2019
Cited alongside, same era.
Data-Free Quantization Through Weight Equalization and Bias Correction
Nagel, M., van Baalen, M., Blankevoort, T., and Welling, M · 2019
Cited alongside, same era.
Cell Division: Weight Bit-Width Reduction Technique for Convolutional Neural Network Hardware Accelerators
Park, H. and Choi, K · 2019
Cited alongside, same era.
Energy-Efficient Neural Network Accelerator Based on Outlier-Aware Low-Precision Computation
Park, E., Kim, D., and Yoo, S
Cited in the paper.
Value-aware Quantization for Training and Inference of Neural Networks
Original
Park, E., Yoo, S., and Vajda, P
Cited in the paper.
DNN Dataflow Choice Is Overrated
Original
Yang, X., Gao, M., Pu, J., Nayak, A., Liu, Q., Bell, S. E., Setter, J. O., Cao, K., Ha, H., Kozyrakis, C., and Horowitz, M
Cited in the paper.
Synetgy: Algorithm-Hardware Co-Design for Convnet Accelerators on Embedded FPGAs
Original
Yang, Y., Huang, Q., Wu, B., Zhang, T., Ma, L., Gambardella, G., Blott, M., Lavagno, L., Vissers, K., Wawrzynek, J., and Keutzer, K
Cited in the paper.