Fetching the paper…
Reading the bibliography…
Conventional multiply-accumulate (MAC) operations have long dominated computation time for deep neural networks (DNNs), espcially convolutional neural networks (CNNs).
EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks
Mingxing Tan and Quoc V. Le. 2019 · 1905
Earlier work this paper cites.
Comparison of parametric representations for monosyllabic word recognition in continuously spoken sentences
S. Davis and P. Mermelstein. 1980 · 1980
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
Product Quantization for Nearest Neighbor Search
Herve Jégou, Matthijs Douze, and Cordelia Schmid. 2011 · 2010
Earlier work this paper cites.
Optimized Product Quantization
Tiezheng Ge, Kaiming He, Qifa Ke, and Jian Sun. 2014 · 2013
Earlier work this paper cites.
Learning Semantic Image Representations at a Large Scale
Yangqing Jia. 2014 · 2014
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Comparing performance, productivity and scalability of the TILT overlay processor to OpenCL HLS. In 2014 International Conference on Field-Programmable Technology (FPT) . IEEE, 20–27
Rafat Rashid, J Gregory Steffan, and Vaughn Betz. 2014 · 2014
Earlier work this paper cites.
Neural Photo Editing with Introspective Adversarial Networks
Andrew Brock, Theodore Lim, J. M. Ritchie, and Nick Weston. 2016 · 2016
Earlier work this paper cites.
OpenCL Caffe: Accelerating and Enabling a Cross Platform Machine Learning Framework. In Proceedings of the 4th International Workshop on OpenCL (Vienna, Austria) (IWOCL ’16) . ACM, New York, NY, USA, Article 8, 5 pages
Junli Gu, Yibing Liu, Yuan Gao, and Maohua Zhu. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016a · 2016
Earlier work this paper cites.
Categorical Reparameterization with Gumbel-Softmax
Eric Jang, Shixiang Gu, and Ben Poole. 2016 · 2016
Earlier work this paper cites.
Fast Algorithms for Convolutional Neural Networks
Andrew Lavin and Scott Gray. 2016 · 2016
Earlier work this paper cites.
XNOR-Net: ImageNet Classification Using Binary Convolutional Neural Networks
Mohammad Rastegari, Vicente Ordonez, Joseph Redmon, and Ali Farhadi. 2016 · 2016
Earlier work this paper cites.
An opencl™ deep learning accelerator on arria 10. In Proceedings of the 2017 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays . 55–64
Utku Aydonat, Shane O’Connell, Davor Capalija, Andrew C Ling, and Gordon R Chiu. 2017 · 2017
Earlier work this paper cites.
EMNIST: Extending MNIST to handwritten letters
Gregory Cohen, Saeed Afshar, Jonathan Tapson, and Andre Van Schaik. 2017 · 2017
Earlier work this paper cites.
Channel Pruning for Accelerating Very Deep Neural Networks
Yihui He, Xiangyu Zhang, and Jian Sun. 2017 · 2017
Earlier work this paper cites.
In-Datacenter Performance Analysis of a Tensor Processing Unit
Norman P. Jouppi, Cliff Young, Nishant Patil, David Patterson, Gaurav Agrawal, Raminder Bajwa, Sarah Bates, Suresh Bhatia, Nan Boden, Al Borchers, Rick Boyle, Pierre-luc Cantin, Clifford Chao, Chris Clark, Jeremy Coriell, Mike Daley, Matt Dau, Jeffrey Dean, Ben Gelb, Tara Vazir Ghaemmaghami, Rajendra Gottipati, William Gulland, Robert Hagmann, C. Richard Ho, Doug Hogberg, John Hu, Robert Hundt, Dan Hurt, Julian Ibarz, Aaron Jaffey, Alek Jaworski, Alexander Kaplan, Harshit Khaitan, Daniel Killebrew, Andy Koch, Naveen Kumar, Steve Lacy, James Laudon, James Law, Diemthu Le, Chris Leary, Zhuyuan Liu, Kyle Lucke, Alan Lundin, Gordon MacKean, Adriana Maggiore, Maire Mahony, Kieran Miller, Rahul Nagarajan, Ravi Narayanaswami, Ray Ni, Kathy Nix, Thomas Norrie, Mark Omernick, Narayana Penukonda, Andy Phelps, Jonathan Ross, Matt Ross, Amir Salek, Emad Samadiani, Chris Severn, Gregory Sizikov, Matthew Snelham, Jed Souter, Dan Steinberg, Andy Swing, Mercedes Tan, Gregory Thorson, Bo Tian, Horia Toma, Erick Tuttle, Vijay Vasudevan, Richard Walter, Walter Wang, Eric Wilcox, and Doe Hyun Yoon. 2017 · 2017
Cited alongside, same era.
Attention is All you Need. In Advances in Neural Information Processing Systems , I. Guyon, U. Von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett (Eds.), Vol. 30. Curran Associates, Inc
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
DLA: Compiler and FPGA Overlay for Neural Network Inference Acceleration
Mohamed S. Abdelfattah, David Han, Andrew Bitar, Roberto Dicecco, Shane O’Connell, Nitika Shanker, Joseph Chu, Ian Prins, Joshua Fender, Andrew C. Ling, and Gordon R. Chiu. 2018 · 2018
What is the State of Neural Network Pruning?. In Proceedings of Machine Learning and Systems , I. Dhillon, D. Papailiopoulos, and V. Sze (Eds.), Vol. 2. 129–146
Davis Blalock, Jose Javier Gonzalez Ortiz, Jonathan Frankle, and John Guttag. 2020 · 2020
Later among the works it cites.
AdderNet: Do We Really Need Multiplications in Deep Learning?. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Hanting Chen, Yunhe Wang, Chunjing Xu, Boxin Shi, Chao Xu, Qi Tian, and Chang Xu. 2020 · 2020
Later among the works it cites.
HPIPE: Heterogeneous Layer-Pipelined and Sparse-Aware CNN Inference for FPGAs. In Proceedings of the 2020 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays (Seaside, CA, USA) (FPGA ’20) . Association for Computing Machinery, New York, NY, USA, 320
Mathew Hall and Vaughn Betz. 2020 · 2020
Later among the works it cites.
Generalized Product Quantization Network for Semi-Supervised Image Retrieval. In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Young Kyun Jang and Nam Ik Cho. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, Matthew Tang, Andrew Howard, Hartwig Adam, and Dmitry Kalenichenko. 2018 · 2018
Cited alongside, same era.
MobileNetV2: Inverted Residuals and Linear Bottlenecks
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen. 2018 · 2018
Cited alongside, same era.
StrassenNets: Deep Learning with a Multiplication Budget. In Proceedings of the 35th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 80) , Jennifer Dy and Andreas Krause (Eds.). PMLR, Stockholmsmässan, Stockholm Sweden, 4985–4994
Michael Tschannen, Aran Khanna, and Animashree Anandkumar. 2018 · 2018
Cited alongside, same era.
Speech commands: A dataset for limited-vocabulary speech recognition
Pete Warden. 2018 · 2018
Cited alongside, same era.
Product Quantization Network for Fast Image Retrieval. In Proceedings of the European Conference on Computer Vision (ECCV)
Tan Yu, Junsong Yuan, Chen Fang, and Hailin Jin. 2018 · 2018
Cited alongside, same era.
PQ-CNN: Accelerating Product Quantized Convolutional Neural Network on FPGA. In 2018 IEEE 26th Annual International Symposium on Field-Programmable Custom Computing Machines (FCCM) . 207–207
Jialiang Zhang and Jing Li. 2018 · 2018
Cited alongside, same era.
Hello Edge: Keyword Spotting on Microcontrollers
Yundong Zhang, Naveen Suda, Liangzhen Lai, and Vikas Chandra. 2018 · 2018
Cited alongside, same era.
Eyeriss v2: A Flexible Accelerator for Emerging Deep Neural Networks on Mobile Devices
Yu-Hsin Chen, Tien-Ju Yang, Joel Emer, and Vivienne Sze. 2019 · 2019
Cited alongside, same era.
End-To-End Supervised Product Quantization for Image Search and Retrieval. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Benjamin Klein and Lior Wolf. 2019 · 2019
Cited alongside, same era.
And the Bit Goes Down: Revisiting the Quantization of Neural Networks. In International Conference on Learning Representations (ICLR)
Pierre Stock, Armand Joulin, Rémi Gribonval, Benjamin Graham, and Hervé Jégou. 2020 · 2020
Later among the works it cites.
Why Gradient Clipping Accelerates Training: A Theoretical Justification for Adaptivity. In International Conference on Learning Representations
Jingzhao Zhang, Tianxing He, Suvrit Sra, and Ali Jadbabaie. 2020 · 2020
Later among the works it cites.
Micronets: Neural network architectures for deploying tinyml applications on commodity microcontrollers
Colby Banbury, Chuteng Zhou, Igor Fedorov, Ramon Matas, Urmish Thakker, Dibakar Gope, Vijay Janapa Reddi, Matthew Mattina, and Paul Whatmough. 2021b · 2021
Later among the works it cites.
Keyword Transformer: A Self-Attention Model for Keyword Spotting. In Proc. Interspeech 2021 . 4249–4253
Axel Berg, Mark O’Connor, and Miguel Tairum Cruz. 2021 · 2021
Later among the works it cites.
Multiplying Matrices Without Multiplying. In Proceedings of the 38th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 992–1004
Davis Blalock and John Guttag. 2021 · 2021
Later among the works it cites.
Image Compression with Product Quantized Masked Image Modeling
Alaaeldin El-Nouby, Matthew J. Muckley, Karen Ullrich, Ivan Laptev, Jakob Verbeek, and Hervé Jégou. 2022 · 2022
Later among the works it cites.
Adaptable Butterfly Accelerator for Attention-based NNs via Hardware and Algorithm Co-design. In 2022 55th IEEE/ACM International Symposium on Microarchitecture (MICRO) . IEEE Computer Society, Los Alamitos, CA, USA, 599–615
H. Fan, T. Chau, S. I. Venieris, R. Lee, A. Kouris, W. Luk, N. D. Lane, and M. S. Abdelfattah. 2022a · 2022
Later among the works it cites.
Adaptable Butterfly Accelerator for Attention-based NNs via Hardware and Algorithm Co-design. In 2022 55th IEEE/ACM International Symposium on Microarchitecture (MICRO) . IEEE Computer Society, Los Alamitos, CA, USA, 599–615
H. Fan, T. Chau, S. I. Venieris, R. Lee, A. Kouris, W. Luk, N. D. Lane, and M. S. Abdelfattah. 2022b · 2022
Later among the works it cites.
WaveMix: A Resource-efficient Neural Network for Image Analysis
Pranav Jeevan, Kavitha Viswanathan, Anandu A S, and Amit Sethi. 2022 · 2022
Later among the works it cites.
Look-ups are not (yet) all you need for deep learning inference
Calvin McCarter and Nicholas Dronen. 2022 · 2022
Later among the works it cites.
MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer. In International Conference on Learning Representations
Sachin Mehta and Mohammad Rastegari. 2022 · 2022
Later among the works it cites.
EdgeViTs: Competing Light-weight CNNs on Mobile Devices with Vision Transformers. In European Conference on Computer Vision
Junting Pan, Adrian Bulat, Fuwen Tan, Xiatian Zhu, Lukasz Dudziak, Hongsheng Li, Georgios Tzimiropoulos, and Brais Martinez. 2022 · 2022
Later among the works it cites.
PECAN: A Product-Quantized Content Addressable Memory Network
Jie Ran, Rui Lin, Jason Chun Lok Li, Jiajun Zhou, and Ngai Wong. 2022 · 2022
Later among the works it cites.