Fetching the paper…
Reading the bibliography…
In the wake of the success of convolutional neural networks in image classification, object recognition, speech recognition, etc., the demand for deploying these compute-intensive ML models on embedded and mobile systems with tight power and energy constraints at low cost, as well as for boosting throughput in data centers, is growing rapidly.
1905
Earlier work this paper cites.
J. Ziv and A. Lempel, “Compression of Individual Sequences via Variable-Rate Coding,” IEEE Transactions on Information Theory , vol. 24, no. 5, pp. 530–536, 1978
1978
Earlier work this paper cites.
T. A. Welch, “A Technique for High-Performance Data Compression,” Computer , vol. 17, no. 6, pp. 8–19, 1984
1984
Earlier work this paper cites.
M. B. Lin, J. F. Lee, and G. E. Jan, “A lossless data compression and decompression algorithm and its hardware architecture,” IEEE Transactions on Very Large Scale Integration Systems , vol. 14, no. 9, pp. 925–936, 2006
2006
Earlier work this paper cites.
F. Iandola, M. Moskewicz, S. Karayev, R. Girshick, T. Darrell, and K. Keutzer, “DenseNet: Implementing Efficient ConvNet Descriptor Pyramids,” UC Berkeley, Berkeley, CA, USA, Tech. Rep., 2014
2014
Earlier work this paper cites.
JEDEC Solid State Technology Association, “Low Power Double Data Rate 4 (LPDDR4),” JEDEC Solid State Technology Association, Tech. Rep. August, 2014
2014
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving Deep into Rectifiers: Surpassing Human-Level Performance on ImageNet Classification,” in Proc. IEEE ICCV . IEEE, dec 2015, pp. 1026–1034. [Online]. Available: http://ieeexplore.ieee.org/document/7410480/
2015
Earlier work this paper cites.
L. Cavigelli, D. Gschwend, C. Mayer, S. Willi, B. Muheim, and L. Benini, “Origamii: A Convolutional Network Accelerator,” in Proc. IEEE GLSVLSI . New York, New York, USA: ACM Press, 2015, pp. 199–204. [Online]. Available: http://dl.acm.org/citation.cfm?doid=2742060.2743766
2015
Earlier work this paper cites.
L. Cavigelli, M. Magno, and L. Benini, “Accelerating Real-Time Embedded Scene Labeling with Convolutional Networks,” in Proc. ACM/IEEE DAC , 2015, pp. 108:1—-108:6
2015
Earlier work this paper cites.
M. Courbariaux, Y. Bengio, and J.-P. David, “BinaryConnect: Training Deep Neural Networks with binary weights during propagations,” in Adv. NIPS , 2015, pp. 3123–3131
2015
Earlier work this paper cites.
W. Chen, J. T. Wilson, S. Tyree, K. Q. Weinberger, and Y. Chen, “Compressing Neural Networks with the Hashing Trick,” in Proc. ICML , vol. 37, 2015, pp. 2285–2294
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep Residual Learning for Image Recognition,” Proc. IEEE CVPR , pp. 770–778, 2015
2015
Earlier work this paper cites.
X. Bian, S. N. Lim, and N. Zhou, “Multiscale fully convolutional network with application to industrial inspection,” in Proc. IEEE WACV , mar 2016, pp. 1–8
2016
Earlier work this paper cites.
L. Cavigelli, D. Bernath, M. Magno, and L. Benini, “Computationally efficient target classification in multispectral image data with Deep Neural Networks,” in Proc. SPIE Security + Defence , vol. 9997, 2016
2016
Earlier work this paper cites.
J. Albericio, P. Judd, T. Hetherington, T. Aamodt, N. E. Jerger, and A. Moshovos, “Cnvlutin: Ineffectual-Neuron-Free Deep Neural Network Computing,” in Proc. ACM/IEEE ISCA , 2016, pp. 1–13
2016
Earlier work this paper cites.
S. Zhang, Z. Du, L. Zhang, H. Lan, S. Liu, L. Li, Q. Guo, T. Chen, and Y. Chen, “Cambricon-X: An accelerator for sparse neural networks,” in Proc. IEEE/ACM MICRO , 2016, pp. 20:1–20:12
2016
Earlier work this paper cites.
Y. H. Chen, J. Emer, and V. Sze, “Eyeriss: A Spatial Architecture for Energy-Efficient Dataflow for Convolutional Neural Networks,” in Proc. ACM/IEEE ISCA , 2016, pp. 367–379
2016
Earlier work this paper cites.
S. Han, X. Liu, H. Mao, J. Pu, A. Pedram, M. A. Horowitz, and W. J. Dally, “EIE: Efficient Inference Engine on Compressed Deep Neural Network,” in Proc. ACM/IEEE ISCA , 2016, pp. 243–254
2016
Earlier work this paper cites.
——, “YodaNN: An Ultra-Low Power Convolutional Neural Network Accelerator Based on Binary Weights,” in Proc. IEEE ISVLSI , 2016, pp. 236–241
2016
Earlier work this paper cites.
S. Han, H. Mao, and W. J. Dally, “Deep Compression - Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding,” in ICLR , 2016
2016
Cited alongside, same era.
F. N. Iandola, S. Han, M. W. Moskewicz, K. Ashraf, W. J. Dally, and K. Keutzer, “SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size,” UC Berkeley, Berkeley, CA, USA, Tech. Rep., 2016
2016
Cited alongside, same era.
X. Zhou, Y. Ito, and K. Nakano, “An efficient implementation of LZW decompression in the FPGA,” Proc. IEEE IPDPS , pp. 599–607, 2016
2016
Cited alongside, same era.
J. Kim, M. Sullivan, E. Choukse, and M. Erez, “Bit-Plane Compression: Transforming Data for Better Compression in Many-Core Architectures,” in Proc. IEEE ISCA , 2016, pp. 329–340
2016
Cited alongside, same era.
——, “Graphics Double Data Rate (GDDR5) SGRAM,” JEDEC Solid State Technology Association, Tech. Rep. Feburary, 2016
R. Andri, L. Cavigelli, D. Rossi, and L. Benini, “YodaNN: An Architecture for Ultralow Power Binary-Weight CNN Acceleration,” IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems , vol. 37, no. 1, pp. 48–60, 2018
2018
Later among the works it cites.
M. Rhu, M. O’Connor, N. Chatterjee, J. Pool, Y. Kwon, and S. W. Keckler, “Compressing DMA Engine: Leveraging Activation Sparsity for Training Deep Neural Networks,” in Proc. IEEE HPCA , 2018, pp. 78–91
2018
Later among the works it cites.
D. Gudovskiy, A. Hodgkinson, and L. Rigazio, “DNN feature map compression using learned representation over GF(2),” in Proc. ECCV Workshops , 2018, pp. 502–516
2018
Later among the works it cites.
N. Ma, X. Zhang, H.-T. Zheng, and J. Sun, “ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design,” in Proc. ECCV , 2018, pp. 116–131
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
G. Litjens, T. Kooi, B. E. Bejnordi, A. A. A. Setio, F. Ciompi, M. Ghafoorian, J. A. van der Laak, B. van Ginneken, and C. I. Sánchez, “A survey on deep learning in medical image analysis,” Medical Image Analysis , vol. 42, no. December, pp. 60–88, 2017
2017
Cited alongside, same era.
B. Wu, F. Iandola, P. H. Jin, and K. Keutzer, “SqueezeDet: Unified, Small, Low Power Fully Convolutional Neural Networks for Real-Time Object Detection for Autonomous Driving,” in Proc. IEEE CVPRW , 2017, pp. 129–137
2017
Cited alongside, same era.
L. Cavigelli and L. Benini, “Origami: A 803-GOp/s/W Convolutional Network Accelerator,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 27, no. 11, pp. 2461–2475, nov 2017
2017
Cited alongside, same era.
A. Parashar, M. Rhu, A. Mukkara, A. Puglielli, R. Venkatesan, B. Khailany, J. Emer, S. W. Keckler, and W. J. Dally, “SCNN: An Accelerator for Compressed-sparse Convolutional Neural Networks,” in Proc. ACM/IEEE ISCA , 2017, pp. 27–40
2017
Cited alongside, same era.
A. Zhou, A. Yao, Y. Guo, L. Xu, and Y. Chen, “Incremental Network Quantization: Towards Lossless CNNs with Low-Precision Weights,” in Proc. ICLR , 2017
2017
Cited alongside, same era.
E. Agustsson, F. Mentzer, M. Tschannen, L. Cavigelli, R. Timofte, L. Benini, and L. Van Gool, “Soft-to-Hard Vector Quantization for End-to-End Learning Compressible Representations,” in Adv. NIPS , 2017, pp. 1141–1151
2017
Cited alongside, same era.
Y. Wang, C. Xu, C. Xu, and D. Tao, “Beyond Filters: Compact Feature Map for Portable Deep Model,” in Proc. ICML , vol. 70, 2017, pp. 3703–3711
2017
Cited alongside, same era.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “MobileNetV2: Inverted Residuals and Linear Bottlenecks,” in Proc. IEEE CVPR . IEEE, jun 2018, pp. 4510–4520
2018
Later among the works it cites.
J. Redmon and A. Farhadi, “YOLOv3: An Incremental Improvement,” University of Washington, Seattle, WA, USA, Tech. Rep., 2018
2018
Later among the works it cites.
M. Kocabas, S. Karagoz, and E. Akbas, “MultiPoseNet: Fast Multi-Person Pose Estimation Using Pose Residual Network,” in Proc. ECCV , 2018, pp. 417–433
2018
Later among the works it cites.
L. Cavigelli and L. Benini, “Extended Bit-Plane Compression for Convolutional Neural Network Accelerators,” in Proc. IEEE AICAS , 2018
2018
Later among the works it cites.
C. Gao, D. Neil, E. Ceolini, S.-C. Liu, and T. Delbruck, “DeltaRNN: A Power-efficient Recurrent Neural Network Accelerator,” Proceedings of the 2018 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays - FPGA ’18 , pp. 21–30, 2018. [Online]. Available: http://dl.acm.org/citation.cfm?doid=3174243.3174261
2018
Later among the works it cites.
E. Flamand, D. Rossi, F. Conti, I. Loi, A. Pullini, F. Rotenberg, and L. Benini, “GAP-8: A RISC-V SoC for AI at the Edge of the IoT,” in Proc. IEEE ASAP . IEEE, jul 2018, pp. 1–4
2018
Later among the works it cites.
A. Aimar, H. Mostafa, E. Calabrese, A. Rios-Navarro, R. Tapiador-Morales, I.-A. Lungu, M. B. Milde, F. Corradi, A. Linares-Barranco, S.-C. Liu, and T. Delbruck, “NullHop: A Flexible Convolutional Neural Network Accelerator Based on Sparse Representations of Feature Maps,” IEEE Transactions on Neural Networks and Learning Systems , vol. 30, no. 3, pp. 644–656, mar 2019
2019
Closest in time.
L. Cavigelli and L. Benini, “CBinfer: Exploiting Frame-to-Frame Locality for Faster Convolutional Network Inference on Video Streams,” IEEE Transactions on Circuits and Systems for Video Technology , 2019
2019
Closest in time.
P. Stock, A. Joulin, R. Gribonval, B. Graham, and H. Jégou, “And the Bit Goes Down: Revisiting the Quantization of Neural Networks,” Facebook AI Research, Tech. Rep., 2019
2019
Closest in time.
R. Andri, L. Cavigelli, D. Rossi, and L. Benini, “Hyperdrive: A Multi-Chip Systolically Scalable Binary-Weight CNN Inference Engine,” IEEE Journal on Emerging and Selected Topics in Circuits and Systems , vol. 9, no. 2, pp. 309–322, jun 2019
2019
Closest in time.
D. Palossi, A. Loquercio, F. Conti, F. Conti, E. Flamand, E. Flamand, D. Scaramuzza, and L. Benini, “A 64mW DNN-based Visual Navigation Engine for Autonomous Nano-Drones,” IEEE Internet of Things Journal , vol. PP, no. May, pp. 1–1, 2019
2019
Closest in time.
M. Mahmoud, K. Siu, and A. Moshovos, “Diffy: a Déjà vu-Free Differential Deep Neural Network Accelerator,” in Proc. IEEE/ACM MICRO . IEEE, oct 2018, pp. 134–147
2019
Closest in time.
L. Cavigelli and L. Benini, “Random Partition Relaxation for Training Binary and Ternary Weight Neural Networks,” in Submitted to International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=S1lvWeBFwB
2020
Closest in time.
M. Tan, B. Chen, R. Pang, V. Vasudevan, M. Sandler, A. Howard, and Q. V. Le, “MnasNet: Platform-Aware Neural Architecture Search for Mobile,” in Proc. IEEE CVPR , 2019, pp. 2820–2028
2028
Closest in time.