Fetching the paper…
Reading the bibliography…
The convolutional neural network (CNN) has become a state-of-the-art method for several artificial intelligence domains in recent years.
An algorithm for subgraph isomorphism
J. R. Ullmann · 1976
Earlier work this paper cites.
A (sub)graph isomorphism algorithm for matching large graphs
L. P. Cordella and et al · 2004
Earlier work this paper cites.
Alpha blending two data streams using a dsp48 ddr technique
Reed P Tidwell · 2005
Earlier work this paper cites.
Opencl: A parallel programming standard for heterogeneous computing systems
Stone and et al · 2010
Earlier work this paper cites.
Halide:a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines
Jonathan Ragan-Kelley, Connelly Barnes, Andrew Adams, Sylvain Paris, and Saman Amarasinghe · 2013
Earlier work this paper cites.
An in-depth comparison of subgraph isomorphism algorithms in graph databases
Jinsoo Lee and et al · 2013
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Going deeper with convolutions
Szegedy and et al · 2015
Earlier work this paper cites.
Polymage: Automatic optimization for image processing pipelines
Mullapudi and et al · 2015
Earlier work this paper cites.
Optimizing fpga-based accelerator design for deep convolutional neural networks
Chen Zhang and et al · 2015
Earlier work this paper cites.
Exploiting vertex relationships in speeding up subgraph isomorphism over large graphs
Xuguang Ren and et al · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He and et al · 2016
Earlier work this paper cites.
Angel-eye: A complete design flow for mapping cnn onto customized hardware
Kaiyuan Guo and et al · 2016
Earlier work this paper cites.
From high-level deep neural models to fpgas
Hardik Sharma and et al · 2016
Earlier work this paper cites.
Cambricon-x: An accelerator for sparse neural networks
Shijin Zhang and et al · 2016
Earlier work this paper cites.
Deepburning: Automatic generation of fpga-based learning accelerators for the neural network family
Ying Wang and et al · 2016
Earlier work this paper cites.
Automatically scheduling halide image processing pipelines
Dillon Sharlet and et al · 2016
Earlier work this paper cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Sanjay Surendranath Girija · 2016
Cited alongside, same era.
Fused-layer cnn accelerators
Manoj Alwani and et al · 2016
Cited alongside, same era.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Andrew G Howard and et al · 2017
Cited alongside, same era.
Platform choices and design demands for iot platforms: cost, power, and performance tradeoffs
Deming Chen and et al · 2017
Cited alongside, same era.
In-datacenter performance analysis of a tensor processing unit
Norman P Jouppi and et al · 2017
Cited alongside, same era.
fpgaconvnet: Automated mapping of convolutional neural networks on fpgas
Stylianos I. Venieris and et al · 2017
An effective fusion and tile size model for optimizing image processing pipelines
Abhinav Jangda and Uday Bondhugula · 2018
Later among the works it cites.
Tvm: end-to-end optimization stack for deep learning
Tianqi Chen and et al · 2018
Later among the works it cites.
Tensor comprehensions: Framework-agnostic high-performance machine learning abstractions
Nicolas Vasilache and et al · 2018
Later among the works it cites.
Yolov3: An incremental improvement
Joseph Redmon and Ali Farhadi · 2018
Later among the works it cites.
Intel ngraph: An intermediate representation, compiler, and executor for deep learning
Scott Cyphers et al · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Compiling deep learning models for custom hardware accelerators
Andre Xian Ming Chang, Aliasger Zaidy, Vinayak Gokhale, and Eugenio Culurciello · 2017
Cited alongside, same era.
Dlvm: A modern compiler infrastructure for deep learning systems
Richard Wei, Lane Schwartz, and Vikram Adve · 2017
Cited alongside, same era.
Exploring heterogeneous algorithms for accelerating deep convolutional neural networks on fpgas
Qingcheng Xiao and et al · 2017
Cited alongside, same era.
Incremental network quantization: Towards lossless cnns with low-precision weights
Aojun Zhou, Anbang Yao, Yiwen Guo, Lin Xu, and Yurong Chen · 2017
Cited alongside, same era.
Embedded vision with int8 optimization on xilinx devices
VSKDKK Yao Fu, Wu Ephrem, and V Kathail · 2017
Cited alongside, same era.
FP-DNN: An automated framework for mapping deep neural networks onto fpgas with rtl-hls hybrid templates
Yijin Guan and et al · 2017
Cited alongside, same era.
Later among the works it cites.
An effective fusion and tile size model for optimizing image processing pipelines
Abhinav Jangda and et al · 2018
Later among the works it cites.
Toolflows for mapping convolutional neural networks on fpgas: A survey and future directions
Stylianos and et al · 2018
Later among the works it cites.
Nervana, 2019
Intel · 2019
Closest in time.
Next-level computing powered by intel ai, 2019
Intel · 2019
Closest in time.
Volta, 2019
Nvidia · 2019
Closest in time.
Adaptive machine learning acceleration, 2019
Xilinx · 2019
Closest in time.
Llvm, 2019
OpenSource · 2019
Closest in time.
Caffe, 2019
Yangqing Jia · 2019
Closest in time.
Darknet, 2019
Joseph Redmon · 2019
Closest in time.
paddlepaddle, 2019
Baidu · 2019
Closest in time.
Xla overview, 2019
Google · 2019
Closest in time.