Fetching the paper…
Reading the bibliography…
We propose a novel hardware and software co-exploration framework for efficient neural architecture search (NAS).
J. D. Schaffer, D. Whitley, and L. J. Eshelman, “Combinations of genetic algorithms and neural networks: A survey of the state of the art,” in International Workshop on Combinations of Genetic Algorithms and Neural Networks (COGANN) . IEEE, 1992, pp. 1–37
1992
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted boltzmann machines,” in International Conference on Machine Learning (ICML) , 2010, pp. 807–814
2010
Earlier work this paper cites.
2014
Earlier work this paper cites.
C. Zhang, P. Li, G. Sun, Y. Guan, B. Xiao, and J. Cong, “Optimizing fpga-based accelerator design for deep convolutional neural networks,” in International Symposium on Field-Programmable Gate Arrays (FPGA) . ACM, 2015, pp. 161–170
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
C. Zhang, D. Wu, J. Sun, G. Sun, G. Luo, and J. Cong, “Energy-efficient cnn implementation on a deeply pipelined fpga cluster,” in International Symposium on Low Power Electronics and Design (ISLPED) . ACM, 2016, pp. 326–331
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Venkataramani, A. Ranjan, S. Banerjee, D. Das, S. Avancha, A. Jagannathan, A. Durg, D. Nagaraj, B. Kaul, P. Dubey et al. , “Scaledeep: A scalable compute architecture for learning and evaluating deep networks,” in ACM SIGARCH Computer Architecture News , vol. 45, no. 2. ACM, 2017, pp. 13–26
2017
Earlier work this paper cites.
P. Whatmough, S. Lee, N. Mulholland, P. Hansen, S. Kodali, D. Brooks, and G. Wei, “Dnn engine: A 16nm sub-uj deep neural network inference accelerator for the embedded masses,” in 2017 IEEE Hot Chips 29 Symposium , 2017
2017
Earlier work this paper cites.
B. Zoph and Q. V. Le, “Neural architecture search with reinforcement learning,” in International Conference on Learning Representations (ICLR) , 2017
2017
Earlier work this paper cites.
L. Xie and A. Yuille, “Genetic cnn,” in International Conference on Computer Vision (ICCV) . IEEE, 2017, pp. 1388–1397
2017
Earlier work this paper cites.
Y.-H. Kim, B. Reddy, S. Yun, and C. Seo, “Nemo: Neuro-evolution with multiobjective optimization of deep neural network for speed and accuracy,” in ICML 2017 AutoML Workshop , 2017
2017
Cited alongside, same era.
Y. Shen, M. Ferdman, and P. Milder, “Maximizing cnn accelerator efficiency through resource partitioning,” in International Symposium on Computer Architecture (ISCA) . IEEE, 2017, pp. 535–547
2017
Cited alongside, same era.
H. Cai, T. Chen, W. Zhang, Y. Yu, and J. Wang, “Efficient architecture search by network transformation.” AAAI, 2018
2018
Cited alongside, same era.
B. Zoph, V. Vasudevan, J. Shlens, and Q. V. Le, “Learning transferable architectures for scalable image recognition,” in IEEE conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 8697–8710
2018
Cited alongside, same era.
T. Geng, T. Wang, A. Sanaullah, C. Yang, R. Xu, R. Patel, and M. Herbordt, “Fpdeep: Acceleration and load balancing of cnn training on fpga clusters,” in International Symposium on Field-Programmable Custom Computing Machines (FCCM) . IEEE, 2018, pp. 81–84
2018
Later among the works it cites.
X. Zhang, J. Wang, C. Zhu, Y. Lin, J. Xiong, W.-m. Hwu, and D. Chen, “Dnnbuilder: An automated tool for building high-performance dnn hardware accelerators for fpgas,” in International Conference on Computer-Aided Design (ICCAD) . ACM, 2018, p. 56
2018
Later among the works it cites.
X. Wei, Y. Liang, X. Li, C. H. Yu, P. Zhang, and J. Cong, “Tgpa: tile-grained pipeline architecture for low latency cnn inference,” in International Conference on Computer-Aided Design (ICCAD) . IEEE, 2018, pp. 1–8
2018
Later among the works it cites.
W. Jiang, E. H.-M. Sha, Q. Zhuge, L. Yang, H. Dong, and X. Chen, “On the design of minimal-cost pipeline systems satisfying hard/soft real-time constraints,” IEEE Transactions on Emerging Topics in Computing , 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
J. Wang, Q. Lou, X. Zhang, C. Zhu, Y. Lin, and D. Chen, “Design flow of accelerating hybrid extremely low bit-width neural network in embedded fpga,” in 2018 28th International Conference on Field Programmable Logic and Applications (FPL) . IEEE, 2018, pp. 163–1636
2018
Cited alongside, same era.
2018
Cited alongside, same era.
G. Bender, P.-J. Kindermans, B. Zoph, V. Vasudevan, and Q. Le, “Understanding and simplifying one-shot architecture search,” in International Conference on Machine Learning , 2018, pp. 549–558
2018
Cited alongside, same era.
E. Chung, J. Fowers, K. Ovtcharov, M. Papamichael, A. Caulfield, T. Massengill, M. Liu, D. Lo, S. Alkalay, M. Haselman et al. , “Serving dnns in real time at datacenter scale with project brainwave,” IEEE Micro , vol. 38, no. 2, pp. 8–20, 2018
2018
Cited alongside, same era.
J. Fowers, K. Ovtcharov, M. Papamichael, T. Massengill, M. Liu, D. Lo, S. Alkalay, M. Haselman, L. Adams, M. Ghandi et al. , “A configurable cloud-scale dnn processor for real-time ai,” in International Symposium on Computer Architecture (ISCA) . IEEE, 2018, pp. 1–14
2018
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
W.-H. Chen, K.-X. Li, W.-Y. Lin, K.-H. Hsu, P.-Y. Li, C.-H. Yang, C.-X. Xue, E.-Y. Yang, Y.-K. Chen, Y.-S. Chang et al. , “A 65nm 1Mb nonvolatile computing-in-memory ReRAM macro with sub-16ns multiply-and-accumulate for binary DNN AI edge processors,” in 2018 IEEE International Solid-State Circuits Conference-(ISSCC) . IEEE, 2018, pp. 494–496
2018
Later among the works it cites.
2019
Closest in time.
Amazon, “Ec2 f1 instances,” https://aws.amazon.com/ ec2/instance-types/f1 , 2017, accessed: 2019-01-20
2019
Closest in time.
Microsoft, “Real-time ai: Microsoft announces preview of project brainwave,” https://blogs.microsoft.com/ai/build-2018-project-brainwave/ , 2018, accessed: 2019-01-20
2019
Closest in time.
W. Zhang, J. Zhang, M. Shen, G. Luo, and N. Xiao, “An efficient mapping approach to large-scale dnns on multi-fpga architectures,” in Design, Automation & Test in Europe Conference & Exhibition (DATE), 2019 . IEEE, 2019, pp. 1–4
2019
Closest in time.
C. Hao, X. Zhang, Y. Li, S. Huang, J. Xiong, K. Rupnow, W.-m. Hwu, and D. Chen, “FPGA/DNN Co-Design: An Efficient Design Methodology for IoT Intelligence on the Edge,” in Proceedings of the 56th Annual Design Automation Conference 2019 . ACM, 2019, p. 206
2019
Closest in time.
Y. Ma, Y. Cao, S. Vrudhula, and J.-s. Seo, “Performance modeling for cnn inference accelerators on fpga,” IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems , 2019
2019
Closest in time.
K. Wang, Z. Liu, Y. Lin, J. Lin, and S. Han, “HAQ: Hardware-Aware Automated Quantization with Mixed Precision,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 8612–8620
2019
Closest in time.