Fetching the paper…
Reading the bibliography…
Deep convolutional neural network (CNN) inference requires significant amount of memory and computation, which limits its deployment on embedded devices.
A vlsi architecture for high-performance, low-cost, on-chip learning
Hammerstrom, Dan · 1990
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Earlier work this paper cites.
Training deep neural networks with low precision multiplications
Courbariaux, Matthieu, David, Jean-Pierre, and Bengio, Yoshua · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding
Jia, Yangqing, Shelhamer, Evan, Donahue, Jeff, Karayev, Sergey, Long, Jonathan, Girshick, Ross, Guadarrama, Sergio, and Darrell, Trevor · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Earlier work this paper cites.
Origami: A convolutional network accelerator
Cavigelli, Lukas, Gschwend, David, Mayer, Christoph, Willi, Samuel, Muheim, Beat, and Benini, Luca · 2015
Earlier work this paper cites.
Reduced-precision memory value approximation for deep learning
Deng, Zhaoxia, Xu, Cong, Cai, Qiong, and Faraboschi, Paolo · 2015
Earlier work this paper cites.
Deep learning with limited numerical precision
Gupta, Suyog, Agrawal, Ankur, Gopalakrishnan, Kailash, and Narayanan, Pritish · 2015
Earlier work this paper cites.
Han, Song, Mao, Huizi, and Dally, William J · 2015
Earlier work this paper cites.
Reduced-precision strategies for bounded memory in deep neural nets
Judd, Patrick, Albericio, Jorge, Hetherington, Tayler, Aamodt, Tor, Jerger, Natalie Enright, Urtasun, Raquel, and Moshovos, Andreas · 2015
Cited alongside, same era.
Deep learning
LeCun, Yann, Bengio, Yoshua, and Hinton, Geoffrey · 2015
Cited alongside, same era.
Fixed point quantization of deep convolutional networks
Lin, Darryl D, Talathi, Sachin S, and Annapureddy, V Sreekanth · 2015
Cited alongside, same era.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2015
Cited alongside, same era.
Eyeriss: An energy-efficient reconfigurable accelerator for deep convolutional neural networks
Chen, Yu-Hsin, Krishna, Tushar, Emer, Joel S, and Sze, Vivienne · 2016
Cited alongside, same era.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and¡ 0.5 mb model size
Iandola, Forrest N, Han, Song, Moskewicz, Matthew W, Ashraf, Khalid, Dally, William J, and Keutzer, Kurt · 2016
Later among the works it cites.
An energy-efficient precision-scalable convnet processor in a 40-nm cmos
Moons, Bert and Verhelst, Marian · 2016
Later among the works it cites.
Going deeper with embedded fpga platform for convolutional neural network
Qiu, Jiantao, Wang, Jie, Yao, Song, Guo, Kaiyuan, Li, Boxun, Zhou, Erjin, Yu, Jincheng, Tang, Tianqi, Xu, Ningyi, Song, Sen, et al · 2016
Later among the works it cites.
Xnor-net: Imagenet classification using binary convolutional neural networks
Rastegari, Mohammad, Ordonez, Vicente, Redmon, Joseph, and Farhadi, Ali · 2016
Later among the works it cites.
Throughput-optimized opencl-based fpga accelerator for large-scale convolutional neural networks
Suda, Naveen, Chandra, Vikas, Dasika, Ganesh, Mohanty, Abinash, Ma, Yufei, Vrudhula, Sarma, Seo, Jae-sun, and Cao, Yu · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Binarynet: Training deep neural networks with weights and activations constrained to +1 or -1
Courbariaux, Matthieu and Bengio, Yoshua · 2016
Cited alongside, same era.
Ristretto: Hardware-oriented approximation of convolutional neural networks
Gysel, Philipp · 2016
Cited alongside, same era.
Hardware-oriented approximation of convolutional neural networks
Gysel, Philipp, Motamedi, Mohammad, and Ghiasi, Soheil · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2016
Cited alongside, same era.
Later among the works it cites.
Accelerating deep convolutional networks using low-precision and sparsity
Venkatesh, Ganesh, Nurvitadhi, Eriko, and Marr, Debbie · 2016
Later among the works it cites.
The microsoft 2016 conversational speech recognition system
Xiong, W, Droppo, J, Huang, X, Seide, F, Seltzer, M, Stolcke, A, Yu, D, and Zweig, G · 2016
Later among the works it cites.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients
Zhou, Shuchang, Ni, Zekun, Zhou, Xinyu, Wen, He, Wu, Yuxin, and Zou, Yuheng · 2016
Later among the works it cites.
Dnpu: An 8.1tops/w reconfigurable cnn-rnn processor for general-purpose deep neural networks
Shin, Dongjoo, Lee, Jinmook, Lee, Jinsu, and Yoo, Hoi-Jun · 2017
Closest in time.