Fetching the paper…
Reading the bibliography…
Using FPGAs to accelerate ConvNets has attracted significant attention in recent years.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Roofline: An insightful visual performance model for floating-point programs and multicore architectures
Samuel Williams, Andrew Waterman, and David Patterson · 2009
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Multi-scale context aggregation by dilated convolutions
Fisher Yu and Vladlen Koltun · 2015
Earlier work this paper cites.
Song Han, Huizi Mao, and William J Dally · 2015
Earlier work this paper cites.
Optimizing fpga-based accelerator design for deep convolutional neural networks
Chen Zhang, Peng Li, Guangyu Sun, Yijin Guan, Bingjun Xiao, and Jason Cong · 2015
Earlier work this paper cites.
Identity mappings in deep residual networks
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Squeezenet: Alexnet-level accuracy with 50x fewer parameters and< 0.5 mb model size
Forrest N Iandola, Song Han, Matthew W Moskewicz, Khalid Ashraf, William J Dally, and Kurt Keutzer · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Chenzhuo Zhu, Song Han, Huizi Mao, and William J Dally · 2016
Earlier work this paper cites.
Xnor-net: Imagenet classification using binary convolutional neural networks
Mohammad Rastegari, Vicente Ordonez, Joseph Redmon, and Ali Farhadi · 2016
Earlier work this paper cites.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients
Shuchang Zhou, Yuxin Wu, Zekun Ni, Xinyu Zhou, He Wen, and Yuheng Zou · 2016
Earlier work this paper cites.
A high performance fpga-based accelerator for large-scale convolutional neural networks
Huimin Li, Xitian Fan, Li Jiao, Wei Cao, Xuegong Zhou, and Lingli Wang · 2016
Earlier work this paper cites.
Going deeper with embedded fpga platform for convolutional neural network
Jiantao Qiu, Jie Wang, Song Yao, Kaiyuan Guo, Boxun Li, Erjin Zhou, Jincheng Yu, Tianqi Tang, Ningyi Xu, Sen Song, et al · 2016
Earlier work this paper cites.
High level synthesis with a dataflow architectural template, 2016
Shaoyi Cheng and John Wawrzynek · 2016
Cited alongside, same era.
Throughput-optimized opencl-based fpga accelerator for large-scale convolutional neural networks
Naveen Suda, Vikas Chandra, Ganesh Dasika, Abinash Mohanty, Yufei Ma, Sarma Vrudhula, Jae-sun Seo, and Yu Cao · 2016
Cited alongside, same era.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger · 2017
Cited alongside, same era.
Learning transferable architectures for scalable image recognition
Barret Zoph, Vijay Vasudevan, Jonathon Shlens, and Quoc V Le · 2017
Cited alongside, same era.
Shift: A zero flop, zero parameter alternative to spatial convolutions
Bichen Wu, Alvin Wan, Xiangyu Yue, Peter Jin, Sicheng Zhao, Noah Golmant, Amir Gholaminejad, Joseph Gonzalez, and Kurt Keutzer · 2017
Accelerating binarized convolutional neural networks with software-programmable fpgas
Ritchie Zhao, Weinan Song, Wentao Zhang, Tianwei Xing, Jeng-Hau Lin, Mani Srivastava, Rajesh Gupta, and Zhiru Zhang · 2017
Later among the works it cites.
A fully connected layer elimination for a binarizec convolutional neural network on an fpga
Hiroki Nakahara, Tomoya Fujii, and Shimpei Sato · 2017
Later among the works it cites.
Software-hardware codesign for efficient neural network acceleration
Kaiyuan Guo, Song Han, Song Yao, Yu Wang, Yuan Xie, and Huazhong Yang · 2017
Later among the works it cites.
Accelerating low bit-width convolutional neural networks with embedded fpga
Li Jiao, Cheng Luo, Wei Cao, Xuegong Zhou, and Lingli Wang · 2017
Later among the works it cites.
Co-design of deep neural nets and neural net accelerators for embedded vision applications
Kiseok Kwon, Alon Amid, Amir Gholami, Bichen Wu, Krste Asanovic, and Kurt Keutzer · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Squeezedet: Unified, small, low power fully convolutional neural networks for real-time object detection for autonomous driving
Bichen Wu, Forrest N Iandola, Peter H Jin, and Kurt Keutzer · 2017
Cited alongside, same era.
Bichen Wu, Alvin Wan, Xiangyu Yue, and Kurt Keutzer · 2017
Cited alongside, same era.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam · 2017
Cited alongside, same era.
Towards Effective Low-bitwidth Convolutional Neural Networks
B. Zhuang, C. Shen, M. Tan, L. Liu, and I. Reid · 2017
Cited alongside, same era.
Improving the performance of opencl-based fpga accelerator for convolutional neural network
Jialiang Zhang and Jing Li · 2017
Cited alongside, same era.
Optimizing loop operation and dataflow in fpga acceleration of deep convolutional neural networks
Yufei Ma, Yu Cao, Sarma Vrudhula, and Jae-sun Seo · 2017
Cited alongside, same era.
Frequency domain acceleration of convolutional neural networks on cpu-fpga shared memory system
Chi Zhang and Viktor Prasanna · 2017
Cited alongside, same era.
Squeezenext: Hardware-aware neural network design
Amir Gholami, Kiseok Kwon, Bichen Wu, Zizheng Tai, Xiangyu Yue, Peter Jin, Sicheng Zhao, and Kurt Keutzer · 2018
Closest in time.
Shufflenet v2: Practical guidelines for efficient cnn architecture design
Ningning Ma, Xiangyu Zhang, Hai-Tao Zheng, and Jian Sun · 2018
Closest in time.
Bichen Wu, Xuanyu Zhou, Sicheng Zhao, Xiangyu Yue, and Kurt Keutzer · 2018
Closest in time.
Mobilenetv2: Inverted residuals and linear bottlenecks
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen · 2018
Closest in time.
Shift-based Primitives for Efficient Convolutional Neural Networks
H. Zhong, X. Liu, Y. He, and Y. Ma · 2018
Closest in time.
Pact: Parameterized clipping activation for quantized neural networks
Jungwook Choi, Zhuo Wang, Swagath Venkataramani, Pierce I-Jen Chuang, Vijayalakshmi Srinivasan, and Kailash Gopalakrishnan · 2018
Closest in time.
Finn-r: An end-to-end deep-learning framework for fast exploration of quantized neural networks, 2018
Michaela Blott, Thomas Preusser, Nicholas Fraser, Giulio Gambardella, Kenneth O’Brien, and Yaman Umuroglu · 2018
Closest in time.
Label refinery: Improving imagenet classification through label progression
Hessam Bagherinezhad, Maxwell Horton, Mohammad Rastegari, and Ali Farhadi · 2018
Closest in time.
Vivado Design Suite User Guide - High-Level Synthesis (UG902), 2018
Xilinx · 2018
Closest in time.
PYNQ Introduction
Xilinx · 2018
Closest in time.
Fp-bnn: Binarized neural network on fpga
Shuang Liang, Shouyi Yin, Leibo Liu, Wayne Luk, and Shaojun Wei · 2018
Closest in time.