Fetching the paper…
Reading the bibliography…
In this research, we propose a new low-precision framework, TENT, to leverage the benefits of a tapered fixed-point numerical format in TinyML models.
Why systolic architectures?
Hsiang-Tsung Kung. 1982 · 1982
Earlier work this paper cites.
Latency lags bandwith
David A Patterson. 2004 · 2004
Earlier work this paper cites.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems
Martín Abadi et al · 2015
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Earlier work this paper cites.
Beating Floating Point at its Own Game: Posit Arithmetic
John L Gustafson and Isaac T Yonemoto. 2017 · 2017
Earlier work this paper cites.
8-bit inference with TensorRT. In GPU Technology Conference
S Migacz. 2017 · 2017
Earlier work this paper cites.
Apprentice: Using Knowledge Distillation Techniques To Improve Low-Precision Network Accuracy
Asit Mishra and Debbie Marr. 2017 · 2017
Earlier work this paper cites.
Pact: Parameterized clipping activation for quantized neural networks
Jungwook Choi, Zhuo Wang, Swagath Venkataramani, Pierce I-Jen Chuang, Vijayalakshmi Srinivasan, and Kailash Gopalakrishnan. 2018 · 2018
Earlier work this paper cites.
Ristretto: A framework for empirical study of resource-efficient inference in convolutional neural networks
Philipp Gysel, Jon Pimentel, Mohammad Motamedi, and Soheil Ghiasi. 2018 · 2018
Earlier work this paper cites.
Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, et al · 2018
Earlier work this paper cites.
Proteus: Exploiting precision variability in deep neural networks
Patrick Judd, Jorge Albericio, Tayler Hetherington, Tor Aamodt, Natalie Enright Jerger, Raquel Urtasun, and Andreas Moshovos. 2018 · 2018
Earlier work this paper cites.
Quantizing deep convolutional networks for efficient inference: A whitepaper
Raghuraman Krishnamoorthi. 2018 · 2018
Earlier work this paper cites.
CMSIS-NN: efficient neural network kernels for arm cortex-M CPUS. CoRR abs/1801.06601 (2018)
Liangzhen Lai, Naveen Suda, and Vikas Chandra. 2018 · 2018
Earlier work this paper cites.
Discovering low-precision networks close to full-precision networks for efficient embedded inference
Jeffrey L McKinstry, Steven K Esser, Rathinakumar Appuswamy, Deepika Bablani, John V Arthur, Izzet B Yildiz, and Dharmendra S Modha. 2018 · 2018
Cited alongside, same era.
An analytical method to determine minimum per-layer precision of deep neural networks. In 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 1090–1094
Charbel Sakr and Naresh Shanbhag. 2018 · 2018
Cited alongside, same era.
Scale-sim: Systolic cnn accelerator simulator
Ananda Samajdar, Yuhao Zhu, Paul Whatmough, Matthew Mattina, and Tushar Krishna. 2018 · 2018
Cited alongside, same era.
Mobilenetv2: Inverted residuals and linear bottlenecks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 4510–4520
Mark Sandler, Andrew Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen. 2018 · 2018
Cited alongside, same era.
Robust navigation with tinyML for autonomous mini-vehicles
Miguel de Prado, Romain Donze, Alessandro Capotondi, Manuele Rusci, Serge Monnerat, Luca Benini, and Nuria Pazos. 2020 · 2020
Later among the works it cites.
Hawq-v2: Hessian aware trace-weighted quantization of neural networks
Zhen Dong, Zhewei Yao, Daiyaan Arfeen, Amir Gholami, Michael W Mahoney, and Kurt Keutzer. 2020 · 2020
Later among the works it cites.
ReLeQ : A Reinforcement Learning Approach for Automatic Deep Quantization of Neural Networks
A. T. Elthakeb, P. Pilligundla, F. Mireshghallah, A. Yazdanbakhsh, and H. Esmaeilzadeh. 2020 · 2020
Later among the works it cites.
Scaling Laws for Autoregressive Generative Modeling
Tom Henighan, Jared Kaplan, Mor Katz, Mark Chen, Christopher Hesse, Jacob Jackson, Heewoo Jun, Tom B Brown, Prafulla Dhariwal, Scott Gray, et al · 2020
Later among the works it cites.
A Generalized Framework for Matching Arithmetic Format to Application Requirements. In https://posithub.org/
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural Network Distillation on IoT Platforms for Sound Event Detection.. In Interspeech . 3609–3613
Gianmarco Cerutti, Rahul Prasad, Alessio Brutti, and Elisabetta Farella. 2019 · 2019
Cited alongside, same era.
Eyeriss v2: A flexible accelerator for emerging deep neural networks on mobile devices
Yu-Hsin Chen, Tien-Ju Yang, Joel Emer, and Vivienne Sze. 2019 · 2019
Cited alongside, same era.
Aakanksha Chowdhery, Pete Warden, Jonathon Shlens, Andrew Howard, and Rocky Rhodes. 2019 · 2019
Cited alongside, same era.
Sparse: Sparse architecture search for cnns on resource-constrained microcontrollers. In Advances in Neural Information Processing Systems . 4977–4989
Igor Fedorov, Ryan P Adams, Matthew Mattina, and Paul Whatmough. 2019 · 2019
Cited alongside, same era.
Mnasnet: Platform-aware neural architecture search for mobile. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 2820–2828
Mingxing Tan, Bo Chen, Ruoming Pang, Vijay Vasudevan, Mark Sandler, Andrew Howard, and Quoc V Le. 2019 · 2019
Cited alongside, same era.
Haq: Hardware-aware automated quantization with mixed precision. In Proceedings of the IEEE conference on computer vision and pattern recognition . 8612–8620
Kuan Wang, Zhijian Liu, Yujun Lin, Ji Lin, and Song Han. 2019 · 2019
Cited alongside, same era.
Colby Banbury, Chuteng Zhou, Igor Fedorov, Ramon Matas Navarro, Urmish Thakkar, Dibakar Gope, Vijay Janapa Reddi, Matthew Mattina, and Paul N Whatmough. 2020b · 2020
Cited alongside, same era.
Benchmarking TinyML Systems: Challenges and Direction
Colby R Banbury, Vijay Janapa Reddi, Max Lam, William Fu, Amin Fazel, Jeremy Holleman, Xinyuan Huang, Robert Hurtado, David Kanter, Anton Lokhmotov, et al · 2020
Cited alongside, same era.
John L. Gustafson. 2020 · 2020
Later among the works it cites.
muNAS: Constrained Neural Architecture Search for Microcontrollers
Edgar Liberis, Łukasz Dudziak, and Nicholas D Lane. 2020 · 2020
Later among the works it cites.
Mcunet: Tiny deep learning on iot devices
Ji Lin, Wei-Ming Chen, Yujun Lin, Chuang Gan, Song Han, et al · 2020
Later among the works it cites.
Least squares binary quantization of neural networks. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops . 698–699
Hadi Pouransari, Zhucheng Tu, and Oncel Tuzel. 2020 · 2020
Later among the works it cites.
Leveraging Automated Mixed-Low-Precision Quantization for tiny edge microcontrollers
Manuele Rusci, Marco Fariselli, Alessandro Capotondi, and Luca Benini. 2020 · 2020
Later among the works it cites.
Improved protein structure prediction using potentials from deep learning
Andrew W Senior, Richard Evans, John Jumper, James Kirkpatrick, et al · 2020
Later among the works it cites.
Compressed deep networks: Goodbye svd, hello robust low-rank approximation
Murad Tukan, Alaa Maalouf, Matan Weksler, and Dan Feldman. 2020 · 2020
Later among the works it cites.
DeepWeeds: A Multiclass Weed Species Image Dataset for Deep Learning
Alex Olsen, Dmitry A Konovalov, Bronson Philippa, Peter Ridd, et al · 2058
Closest in time.