Fetching the paper…
Reading the bibliography…
Edge TPUs are a domain of accelerators for low-power, edge devices and are widely used in various Google products such as Coral and Pixel devices.
F. Hartleb and V. Mertsiotakis, “Bounds for the Mean Runtime of Parallel Programs,” in Proceedings of the Sixth International Conference on Modelling Techniques and Tools for Computer Performance Evaluation , 1992
1992
Earlier work this paper cites.
T. Fahringer and H. P. Zima, “ A Static Parameter based Performance Prediction Tool for Parallel Programs ,” in ICS , 1993
1993
Earlier work this paper cites.
C. Y. Park, “ Predicting Program Execution Times by Analyzing Static and Dynamic Program Paths ,” Real-Time Systems , 1993
1993
Earlier work this paper cites.
R. Rugina and K. E. Schauser, “ Predicting the Running Times of Parallel Programs by Simulation ,” in IPDPS , 1998
1998
Earlier work this paper cites.
C. Ferdinand, R. Heckmann, M. Langenbach, F. Martin, M. Schmidt, H. Theiling, S. Thesing, and R. Wilhelm, “ Reliable and Precise WCET Determination for a Real-life Processor ,” in International Workshop on Embedded Software , 2001
2001
Earlier work this paper cites.
V. S. Adve and M. K. Vernon, “ Parallel Program Performance Prediction using Deterministic Task Graph Analysis ,” TOCS , 2004
2004
Earlier work this paper cites.
V. Blanco, J. A. González, C. León, C. Rodrıguez, G. Rodrıguez, and M. Printista, “ Predicting the Performance of Parallel Programs ,” Parallel Computing , 2004
2004
Earlier work this paper cites.
C. Dubach, J. Cavazos, B. Franke, G. Fursin, M. F. O’Boyle, and O. Temam, “ Fast Compiler Optimisation Evaluation using Code-feature based Performance Prediction ,” in CF , 2007
2007
Earlier work this paper cites.
X. Li, Y. Liang, T. Mitra, and A. Roychoudhury, “ Chronos: A Timing Analyzer for Embedded Software ,” Science of Computer Programming , 2007
2007
Earlier work this paper cites.
T. M. Taha and S. Wills, “ An Instruction Throughput Model of Superscalar Processors ,” IEEE Transactions on Computers , 2008
2008
Earlier work this paper cites.
X. E. Chen and T. M. Aamodt, “ A First-order Fine-grained Multithreaded Throughput Model ,” in HPCA , 2009
2009
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “ Learning Multiple Layers of Features from Tiny Images ,” 2009
2009
Earlier work this paper cites.
L. Huang, J. Jia, B. Yu, B.-G. Chun, P. Maniatis, and M. Naik, “ Predicting Execution Time of Computer Programs using Sparse Polynomial Regression ,” in NeurIPS , 2010
2010
Earlier work this paper cites.
A. Patel, F. Afram, S. Chen, and K. Ghose, “ MARSS: A Full System Simulator for Multicore x86 CPUs ,” in DAC , 2011
2011
Earlier work this paper cites.
S. A. Seshia and J. Kotker, “ GameTime: A Toolkit for Timing Analysis of Software ,” in TACAS , 2011
2011
Earlier work this paper cites.
S. A. Seshia and A. Rakhlin, “ Quantitative Analysis of Systems using Game-theoretic Learning ,” TECS , 2012
2012
Earlier work this paper cites.
D. Sanchez and C. Kozyrakis, “ ZSim: Fast and Accurate Microarchitectural Simulation of Thousand-core Systems ,” ACM SIGARCH Computer architecture news , 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
W. Hamilton, Z. Ying, and J. Leskovec, “ Inductive Representation Learning on Large Graphs ,” in NeurIPS , 2017
2017
Cited alongside, same era.
N. P. Jouppi, C. Young, N. Patil, D. Patterson, G. Agrawal, R. Bajwa, S. Bates, S. Bhatia, N. Boden, A. Borchers, R. Boyle, P.-l. Cantin, C. Chao, C. Clark, J. Coriell, M. Daley, M. Dau, J. Dean, B. Gelb, T. V. Ghaemmaghami, R. Gottipati, W. Gulland, R. Hagmann, C. R. Ho, D. Hogberg, J. Hu, R. Hundt, D. Hurt, J. Ibarz, A. Jaffey, A. Jaworski, A. Kaplan, H. Khaitan, D. Killebrew, A. Koch, N. Kumar, S. Lacy, J. Laudon, J. Law, D. Le, C. Leary, Z. Liu, K. Lucke, A. Lundin, G. MacKean, A. Maggiore, M. Mahony, K. Miller, R. Nagarajan, R. Narayanaswami, R. Ni, K. Nix, T. Norrie, M. Omernick, N. Penukonda, A. Phelps, J. Ross, M. Ross, A. Salek, E. Samadiani, C. Severn, G. Sizikov, M. Snelham, J. Souter, D. Steinberg, A. Swing, M. Tan, G. Thorson, B. Tian, H. Toma, E. Tuttle, V. Vasudevan, R. Walter, W. Wang, E. Wilcox, and D. H. Yoon, “ In-data Center Performance Analysis of a Tensor Processing Unit ,” in ISCA , 2017
I. Corporation, “Intel Architecture Code Analyzer,” https://software.intel.com/content/www/us/en/develop/articles/intel-architecture-code-analyzer.html , 2019, accessed: 2020-09-10
2020
Later among the works it cites.
W. Wen, H. Liu, Y. Chen, H. Li, G. Bender, and P.-J. Kindermans, “ Neural Predictor for Neural Architecture Search ,” in ECCV , 2020
2020
Later among the works it cites.
“Edge TPU,” https://cloud.google.com/edge-tpu , accessed: 2021-01-09
2021
Closest in time.
“Edge TPU Compiler,” https://coral.ai/docs/edgetpu/compiler/#system-requirements , accessed: 2021-01-09
2021
Closest in time.
“Introducing the Next Generation of On-Device Vision Models: MobileNetV3 and MobileNetEdgeTPU,” https://ai.googleblog.com/2019/11/introducing-next-generation-on-device.html , accessed: 2021-01-09
2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2018
Cited alongside, same era.
E. Chung, J. Fowers, K. Ovtcharov, M. Papamichael, A. Caulfield, T. Massengill, M. Liu, D. Lo, S. Alkalay, M. Haselman, M. Abeydeera, L. Adams, H. Angepat, C. Boehn, D. Chiou, O. Firestein, A. Forin, K. S. Gatlin, M. Ghandi, S. Heil, K. Holohan, A. El Husseini, T. Juhasz, K. Kagi, R. K. Kovvuri, S. Lanka, F. van Megen, D. Mukhortov, P. Patel, B. Perez, A. Rapsang, S. Reinhardt, B. Rouhani, A. Sapek, R. Seera, S. Shekar, B. Sridharan, G. Weisz, L. Woods, P. Yi Xiao, D. Zhang, R. Zhao, and D. Burger, “ Serving DNNs in Real Time at Datacenter Scale with Project Brainwave ,” IEEE Micro , 2018
2018
Cited alongside, same era.
J. Laukemann, J. Hammer, J. Hofmann, G. Hager, and G. Wellein, “ Automated Instruction Stream Throughput Prediction for Intel and AMD Microarchitectures ,” in PMBS , 2018
2018
Cited alongside, same era.
A. Adams, K. Ma, L. Anderson, R. Baghdadi, T.-M. Li, M. Gharbi, B. Steiner, S. Johnson, K. Fatahalian, F. Durand, and J. Ragan-Kelley, “ Learning to Optimize Halide with Tree Search and Random Programs ,” TOG , 2019
2019
Cited alongside, same era.
J. Chang, X. Zhang, Y. Guo, G. Meng, S. XIiang, and C. Pan, “ DATA: Differentiable ArchiTecture Approximation ,” in NeurIPS , 2019
2019
Cited alongside, same era.
H. Hu, J. Langford, R. Caruana, S. Mukherjee, E. J. Horvitz, and D. Dey, “ Efficient Forward Architecture Search ,” in NeurIPS , 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
C. Mendis, A. Renda, S. P. Amarasinghe, and M. Carbin, “ Ithemal: Accurate, Portable and Fast Basic Block Throughput Estimation using Deep Neural Networks ,” in ICML , 2019
2019
Cited alongside, same era.
“Run Inference on the Edge TPU with Python,” https://coral.ai/docs/edgetpu/tflite-python/#overview , accessed: 2021-01-09
2021
Closest in time.
Apple, “M1 Processor,” https://www.apple.com/mac/m1/ , 2021
2021
Closest in time.
DeepMind, “Sonnet,” https://github.com/deepmind/sonnet , 2021
2021
Closest in time.
EETimes, “AWS Rolls Out AI Inference Chip,” https://www.eetimes.com/aws-rolls-out-ai-inference-chip/ , 2021
2021
Closest in time.
Facebook, “Accelerating Facebook’s Infrastructure with Application-specific Hardware,” https://engineering.fb.com/2019/03/14/data-center-engineering/accelerating-infrastructure/ , 2021
2021
Closest in time.
Google, “ML for Mobile and Edge Devices - TensorFlow Lite,” https://www.tensorflow.org/lite , 2021
2021
Closest in time.
K. Hegde, P.-A. Tsai, S. Huang, V. Chandra, A. Parashar, and C. W. Fletcher, “ Mind Mappings: Enabling Efficient Algorithm-Accelerator Mapping Space Search ,” in ASPLOS , 2021
2021
Closest in time.
S. Kaufman, P. Phothilimthana, Y. Zhou, C. Mendis, S. Roy, A. Sabne, and M. Burrows, “ A Learned Performance Model for Tensor Processing Units ,” MLSys , 2021
2021
Closest in time.
2021
Closest in time.
“Coral,” https://coral.ai/ , accessed: 2022-09-30
2022
Closest in time.
2022
Closest in time.
O. Sýkora, P. M. Phothilimthana, C. Mendis, and A. Yazdanbakhsh, “GRANITE: A Graph Neural Network Model for Basic Block Throughput Estimation,” in IISWC , 2022
2022
Closest in time.