A fast learning algorithm for deep belief nets
Geoffrey E Hinton, Simon Osindero, and Yee-Whye Teh · 2006
Earlier work this paper cites.
Optimizing NUCA organizations and wiring alternatives for large caches with CACTI 6.0
Naveen Muralimanohar, Rajeev Balasubramonian, and Norm Jouppi · 2007
Earlier work this paper cites.
Neuflow: A runtime reconfigurable dataflow processor for vision
Clément Farabet, Berin Martini, Benoit Corda, Polina Akselrod, Eugenio Culurciello, and Yann LeCun · 2011
Earlier work this paper cites.
Halide: A language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines
Jonathan Ragan-Kelley, Connelly Barnes, Andrew Adams, Sylvain Paris, Frédo Durand, and Saman Amarasinghe · 2013
Earlier work this paper cites.
Improving high level synthesis optimization opportunity through polyhedral transformations
Wei Zuo, Yun Liang, Peng Li, Kyle Rupnow, Deming Chen, and Jason Cong · 2013
Earlier work this paper cites.
DianNao: A small-footprint high-throughput accelerator for ubiquitous machine-learning
Tianshi Chen, Zidong Du, Ninghui Sun, Jia Wang, Chengyong Wu, Yunji Chen, and Olivier Temam · 2014
Earlier work this paper cites.
DaDianNao: A machine-learning supercomputer
Yunji Chen, Tao Luo, Shaoli Liu, Shijin Zhang, Liqiang He, Jia Wang, Ling Li, Tianshi Chen, Zhiwei Xu, Ninghui Sun, and Olivier Temam · 2014
Earlier work this paper cites.
A 240 G-ops/s mobile coprocessor for deep neural networks
Vinayak Gokhale, Jonghoon Jin, Aysegul Dundar, Berin Martini, and Eugenio Culurciello · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Original
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Original
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Earlier work this paper cites.
Going deeper with convolutions
Original
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2014
Earlier work this paper cites.
ShiDianNao: Shifting vision processing closer to the sensor
Zidong Du, Robert Fasthuber, Tianshi Chen, Paolo Ienne, Ling Li, Tao Luo, Xiaobing Feng, Yunji Chen, and Olivier Temam · 2015
Earlier work this paper cites.
Polymage: Automatic optimization for image processing pipelines
Ravi Teja Mullapudi, Vinay Vasista, and Uday Bondhugula · 2015
Earlier work this paper cites.
Optimizing FPGA-based accelerator design for deep convolutional neural networks
Chen Zhang, Peng Li, Guangyu Sun, Yijin Guan, Bingjun Xiao, and Jason Cong · 2015
Earlier work this paper cites.
Cnvlutin: Ineffectual-neuron-free deep neural network computing
Jorge Albericio, Patrick Judd, Tayler Hetherington, Tor Aamodt, Natalie Enright Jerger, and Andreas Moshovos · 2016
Earlier work this paper cites.
Fused-layer CNN accelerators
Manoj Alwani, Han Chen, Michael Ferdman, and Peter Milder · 2016
Earlier work this paper cites.