Fetching the paper…
Reading the bibliography…
Deep learning compiler frameworks are gaining ground as a more portable back-end for deep learning applications on increasingly diverse hardware.
Simulating Execution Time of Tensor Programs using Graph Neural Networks
Jakub M. Tomczak, Romain Lepert, and Auke Wiggers. 2019 · 1904
Earlier work this paper cites.
Simulated annealing
Dimitris Bertsimas, John Tsitsiklis, et al · 1993
Earlier work this paper cites.
Chameleon: Adaptive Code Optimization for Expedited Deep Neural Network Compilation
Byung Hoon Ahn, Prannoy Pilligundla, Amir Yazdanbakhsh, and Hadi Esmaeilzadeh. 2020 · 2001
Earlier work this paper cites.
Learning Fast Adaptation on Cross-Accented Speech Recognition
Genta Indra Winata, Samuel Cahyawijaya, Zihan Liu, Zhaojiang Lin, Andrea Madotto, Peng Xu, and Pascale Fung. 2020 · 2003
Earlier work this paper cites.
Spatial Sharing of GPU for Autotuning DNN models
Aditya Dhakal, Junguk Cho, Sameer G. Kulkarni, K. K. Ramakrishnan, and Puneet Sharma. 2020 · 2008
Earlier work this paper cites.
A Learned Performance Model for the Tensor Processing Unit
Samuel J. Kaufman, Phitchaya Mangpo Phothilimthana, Yanqi Zhou, and Mike Burrows. 2020 · 2008
Earlier work this paper cites.
Milepost GCC: Machine Learning Enabled Self-tuning Compiler
Grigori Fursin, Yuriy Kashnikov, Abdul Wahid Memon, Zbigniew Chamski, Olivier Temam, Mircea Namolaru, Elad Yom-Tov, Bilha Mendelson, Ayal Zaks, Eric Courtois, François Bodin, Phil Barnard, Elton Ashton, Edwin V. Bonilla, John Thomson, Christopher K. I. Williams, and Michael F. P. O’Boyle. 2011 · 2011
Earlier work this paper cites.
Halide: A Language and Compiler for Optimizing Parallelism, Locality, and Recomputation in Image Processing Pipelines. In Proceedings of the 34th ACM SIGPLAN Conference on Programming Language Design and Implementation (Seattle, Washington, USA) (PLDI ’13) . Association for Computing Machinery, New York, NY, USA, 519–530
Jonathan Ragan-Kelley, Connelly Barnes, Andrew Adams, Sylvain Paris, Frédo Durand, and Saman Amarasinghe. 2013 · 2013
Earlier work this paper cites.
Parallelizing Exploration-Exploitation Tradeoffs in Gaussian Process Bandit Optimization
Thomas Desautels, Andreas Krause, and Joel W. Burdick. 2014 · 2014
Earlier work this paper cites.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2015 · 2015
Earlier work this paper cites.
MXNet: A Flexible and Efficient Machine Learning Library for Heterogeneous Distributed Systems
Tianqi Chen, Mu Li, Yutian Li, Min Lin, Naiyan Wang, Minjie Wang, Tianjun Xiao, Bing Xu, Chiyuan Zhang, and Zheng Zhang. 2015 · 2015
Cited alongside, same era.
Autotuning Algorithmic Choice for Input Sensitivity
Yufei Ding, Jason Ansel, Kalyan Veeramachaneni, Xipeng Shen, Una-May O’Reilly, and Saman Amarasinghe. 2015b · 2015
Cited alongside, same era.
Siamese neural networks for one-shot image recognition
Gregory Koch. 2015 · 2015
Cited alongside, same era.
XGBoost: A Scalable Tree Boosting System. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (San Francisco, California, USA) (KDD ’16) . Association for Computing Machinery, New York, NY, USA, 785–794
Tianqi Chen and Carlos Guestrin. 2016 · 2016
Cited alongside, same era.
The reparameterization trick for acquisition functions
James T. Wilson, Riccardo Moriconi, Frank Hutter, and Marc Peter Deisenroth. 2017 · 2017
Later among the works it cites.
Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions
Nicolas Vasilache, Oleksandr Zinenko, Theodoros Theodoridis, Priya Goyal, Zachary DeVito, William S. Moses, Sven Verdoolaege, Andrew Adams, and Albert Cohen. 2018 · 2018
Later among the works it cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Later among the works it cites.
Glow: Graph Lowering Compiler Techniques for Neural Networks
Nadav Rotem, Jordan Fix, Saleem Abdulrasool, Garret Catron, Summer Deng, Roman Dzhabarov, Nick Gibson, James Hegeman, Meghan Lele, Roman Levenstein, Jack Montgomery, Bert Maher, Satish Nadathur, Jakob Olesen, Jongsoo Park, Artem Rakhov, Misha Smelyanskiy, and Man Wang. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The Parallel Knowledge Gradient Method for Batch Bayesian Optimization. In Proceedings of the 30th International Conference on Neural Information Processing Systems (Barcelona, Spain) (NIPS’16) . Curran Associates Inc., Red Hook, NY, USA, 3134–3142
Jian Wu and Peter I. Frazier. 2016 · 2016
Cited alongside, same era.
End-to-End Deep Learning of Optimization Heuristics. In 2017 26th International Conference on Parallel Architectures and Compilation Techniques (PACT) . 219–232
C. Cummins, P. Petoumenos, Z. Wang, and H. Leather. 2017 · 2017
Cited alongside, same era.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Semi-Supervised Classification with Graph Convolutional Networks. In Proceedings of the 5th International Conference on Learning Representations (Palais des Congrès Neptune, Toulon, France) (ICLR ’17)
Thomas N. Kipf and Max Welling. 2017 · 2017
Cited alongside, same era.
XLA: TensorFlow, compiled
Chris Leary and Todd Wang. 2017 · 2017
Cited alongside, same era.
TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In Proceedings of the 13th USENIX Conference on Operating Systems Design and Implementation (Carlsbad, CA, USA) (OSDI’18) . USENIX Association, USA, 579–594
Tianqi Chen, Thierry Moreau, Ziheng Jiang, Lianmin Zheng, Eddie Yan, Meghan Cowan, Haichen Shen, Leyuan Wang, Yuwei Hu, Luis Ceze, Carlos Guestrin, and Arvind Krishnamurthy. 2018a
Cited in the paper.
Learning to Optimize Tensor Programs. In Proceedings of the 32nd International Conference on Neural Information Processing Systems (Montréal, Canada) (NIPS’18) . Curran Associates Inc., Red Hook, NY, USA, 3393–3404
Tianqi Chen, Lianmin Zheng, Eddie Yan, Ziheng Jiang, Thierry Moreau, Luis Ceze, Carlos Guestrin, and Arvind Krishnamurthy. 2018b
Cited in the paper.
Autotuning Algorithmic Choice for Input Sensitivity. In Proceedings of the 36th ACM SIGPLAN Conference on Programming Language Design and Implementation (Portland, OR, USA) (PLDI ’15) . Association for Computing Machinery, New York, NY, USA, 379–390
Yufei Ding, Jason Ansel, Kalyan Veeramachaneni, Xipeng Shen, Una-May O’Reilly, and Saman Amarasinghe. 2015a
Cited in the paper.
Later among the works it cites.
POSTER: CogR: Exploiting Program Structures for Machine-Learning Based Runtime Solutions. In 2019 28th International Conference on Parallel Architectures and Compilation Techniques (PACT) . 485–486
H. Sung, T. Chen, A. Eichenberger, and K. K. O’Brien. 2019 · 2019
Later among the works it cites.
AdaTune: Adaptive Tensor Program Compilation Made Efficient. In 34th Conference on Neural Information Processing Systems (NeurIPS 2020)
Menghao Li, Minjia Zhang, Chi Wang, and Mingqin Li. 2020 · 2020
Later among the works it cites.
Towards Fast Adaptation of Neural Architectures with Meta Learning
Dongze Lian, Yintao Xu Yin Zheng, Yanxiong Lu, Leyu Lin, Peilin Zhao, Junzhou Huang, and Shenghua Gao. 2020 · 2020
Later among the works it cites.
Meta-Transfer Learning for Zero-Shot Super-Resolution. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 3516–3525
Jae Woong Soh, Sunwoo Cho, and Nam Ik Cho. 2020 · 2020
Later among the works it cites.
ALT: Optimizing Tensor Compilation in Deep Learning Compilers with Active Learning. In 2020 IEEE 38th International Conference on Computer Design (ICCD) . 623–630
X. Zeng, T. Zhi, Z. Du, Q. Guo, N. Sun, and Y. Chen. 2020 · 2020
Later among the works it cites.