Fetching the paper…
Reading the bibliography…
The biggest challenge for the deployment of Deep Neural Networks (DNNs) close to the generated data on edge devices is their size, i.e., memory footprint and computational complexity.
1902
Earlier work this paper cites.
1905
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, J. L. Li, K. Li, and F. F. Li, “Imagenet: A large-scale hierarchical image database,” In Proceedings of 2009 IEEE conference on Computer Vision and Pattern Recognition (CVPR), 2009, pp.248–255
2009
Earlier work this paper cites.
K. T. Malladi, F. A. Nothaft, K. Periyathambi, B. C. Lee, C. Kozyrakis, and M. Horowitz, “Towards energy-proportional datacenter memory with mobile DRAM,” In Proceedings of 39th Annual International Symposium on Computer Architecture (ISCA), 2012, pp. 37–-48
2012
Earlier work this paper cites.
S. Clerc, F. Abouzeid, G. Gasiot, D. Gauthier, D. Soussan, and P. Roche, “A 0.32 V, 55fJ per bit access energy, CMOS 65nm bit-interleaved SRAM with radiation Soft Error tolerance,” In 2012 IEEE International Conference on IC Design & Technology, 2012, pp. 1–4
2012
Earlier work this paper cites.
M. Horowitz, “1.1 computing’s energy problem (and what we can do about it),” In IEEE International Solid-State Circuits Conference (IEEE ISSCC), 2014, pp. 10–14
2014
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, F. F. Li, “ImageNet Large Scale Visual Recognition Challenge,” In Proceedings of International Journal of Computer Vision (IJCV), 2015, pp. 211–252
2015
Earlier work this paper cites.
M. Schaffner, F. K. Gürkaynak, A. Smolic, and L. Benini, “DRAM or no-DRAM? Exploring linear solver architectures for image domain warping in 28 nm CMOS,” In Proceedings of Design, Automation, & Test in Europe Conference Exhibition (DATE), May 2015, pp. 707-–712
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” In Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR), 2016, pp. 770–778
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
D. Lin, S. Talathi, and S. Annapureddy, “Fixed point quantization of deep convolutional networks,” In Proceedings of International Conference on Machine Learning (ICML), 2016, pp. 2849–2858
2016
Earlier work this paper cites.
X. Dong, S. Chen, and S. Pan, “Learning to prune deep neural networks via layer-wise optimal brain surgeon,” In Proceedings of 31st Conference on Neural Information Processing Systems (NIPS), 2017
2017
Cited alongside, same era.
V. Akhlaghi, A. Yazdanbakhsh, K. Samadi, R. K. Gupta, and Hadi Esmaeilzadeh, “Snapea: Predictive early activation for reducing computation in deep convolutional neural networks,” In 2018 ACM/IEEE 45th Annual International Symposium on Computer Architecture (ISCA), 2018, pp. 662–673
2018
Cited alongside, same era.
Y. Zhou, S. M. Moosavi-Dezfooli, N. M. Cheung, and P. Frossard, “Adaptive quantization for deep neural network,” In Proceedings of the 32nd AAAI Conference on Artificial Intelligence (AAAI), 2018, pp. 4596–4604
2018
Cited alongside, same era.
2018
Cited alongside, same era.
L. Deng, G. Li, S. Han, L. Shi, and Y. Xie, “Model compression and hardware acceleration for neural networks: A comprehensive survey,” In Proceedings of the IEEE 108, no. 4, 2020, pp. 485–532
2020
Later among the works it cites.
S. Gupta, S. Ullah, K. Ahuja, A. Tiwari, and A. Kumar, “ALigN: A Highly Accurate Adaptive Layerwise Log2Lead Quantization of Pre-Trained Neural Networks,” In IEEE Access, vol. 8, 2020, pp. 118899-118911
2020
Later among the works it cites.
M. Nagel, R. A. Amjad, M. Van Baalen, C. Louizos, and T. Blankevoort, “Up or Down? Adaptive Rounding for Post-Training Quantization,” In Proceedings of International Conference on Machine Learning (ICML), 2020, pp. 7197–7206
2020
Later among the works it cites.
T. Stadtmann, C. Latotzke, and T. Gemmeke, “From Quantitative Analysis to Synthesis of Efficient Binary Neural Networks,” In Proceedings of the 19th IEEE International Conference On Machine Learning And Applications (ICMLA), 2020, pp. 93–100
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Yang, X. Shen, J. Xing, X. Tian, H. Li, B. Deng, J. Huang, and X. S. Hua, “Quantization Networks,” In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2019, pp. 7308–7316
2019
Cited alongside, same era.
Y. H. Chen, T. J. Yang, J. S. Emer, and V. Sze, “Eyeriss v2: A Flexible Accelerator for Emerging Deep Neural Networks on Mobile Devices,” In IEEE Journal on Emerging and Selected Topics in Circuits and Systems (IEEE J. Emerg. Sel), vol. 9, no. 2, 2019, pp. 292–308
2019
Cited alongside, same era.
W. Nogami, T. Ikegami, R. Takano, and T. Kudoh, “Optimizing Weight Value Quantization for CNN Inference,” In Proceedings of 2019 International Joint Conference on Neural Networks (IJCNN), 2019, pp. 1–8
2019
Cited alongside, same era.
M. Nagel, M. van Baalen, T. Blankevoort, and M. Welling, “Data-Free Quantization Through Weight Equalization and Bias Correction,” In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2019, pp. 1325–1334
2019
Cited alongside, same era.
N. Mitschke, M. Heizmann, K. H. Noffz, and R. Wittmann, “A Fixed-Point Quantization Technique for Convolutional Neural Networks Based on Weight Scaling,” In Proceedings of 2019 IEEE International Conference on Image Processing (ICIP), 2019, pp. 3836–3840
2019
Cited alongside, same era.
S. Mittal, “A survey of FPGA-based accelerators for convolutional neural networks,” In Neural Computing and Applications (Neural. Comput. Appl.), vol. 32, no. 4, 2020, pp. 1109–1139
2020
Cited alongside, same era.
Y. Bhalgat, J. Lee, M. Nagel, T. Blankevoort, and N. Kwak, “Lsq+: Improving low-bit quantization through learnable offsets and better initialization,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops, 2020, pp. 696–697
2020
Cited alongside, same era.
C. Latotzke, and T. Gemmeke, “Efficiency Versus Accuracy: A Review of design Techniques for DNN Hardware Accelerators,” In IEEE Access, vol. 9, 2021, pp. 9785–9799
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
D. Lee, M. Cho, S. Lee, J. Song, and C. Choi, “A Novel Sensitivity Metric For Mixed-Precision Quantization With Synthetic Data Generation,” In Proceedings of 2021 IEEE International Conference on Image Processing (ICIP), 2021, pp. 1294-1298
2021
Later among the works it cites.
I. Wallossek, “Nvidia GeForce RTX 2080 Ti im großen Effizienz-Test von 140 bis 340 Watt — igorsLAB,” https://www.igorslab.de/nvidia-geforce-rtx-2080-ti-im-grossen-effizienz-test-von-140-bis-340-watt-igorslab/, accessed 2022-Sep-06 14:19
2022
Closest in time.
S. Balaban and C. Li, “RTX 2080 Ti Deep Learning Benchmarks with TensorFlow,” In https://lambdalabs.com/blog/2080-ti-deep-learning-benchmarks/, accessed 2022-Sep-06 14:19
2022
Closest in time.