Fetching the paper…
Reading the bibliography…
The effectiveness of Machine Learning (ML) methods depend on access to large suitable datasets.
Puschel, M. and Moura, J.M.F. and Johnson, J.R. and Padua, D. and Veloso, M.M. and Singer, B.W. and Jianxin Xiong and Franchetti, F. and Gacic, A. and Voronenko, Y. and Chen, K. and Johnson, R.W. and Rizzolo, N., “SPIRAL: Code Generation for DSP Transforms,” Proceedings of the IEEE, 232–275, February 2005
2005
Earlier work this paper cites.
Hartono, Albert and Norris, Boyana and Sadayappan, P., Annotation-based empirical performance tuning using Orio, IEEE International Symposium on Parallel Distributed Processing, 2009, pp.1–11
2009
Earlier work this paper cites.
Falch, Thomas L., and Anne C. Elster. “Machine Learning-Based Auto-Tuning for Enhanced Performance Portability of OpenCL Applications.” Concurrency and Computation: Practice and Experience 29, no. 8 (2017): e4029. https://doi.org/10.1002/cpe.4029
2017
Earlier work this paper cites.
Zhang, W., Hao, M. & Snir, Marc “Predicting HPC parallel program performance based on LLVM compiler.“ Cluster Comput 20, 1179–1192 2017
2017
Earlier work this paper cites.
Cummins, Chris and Petoumenos, Pavlos and Wang, Zheng and Leather, Hugh, “End-to-End Deep Learning of Optimization Heuristics,”, IEEE, September 2017, p219–232
2017
Earlier work this paper cites.
Lim, Robert V. and Norris, Boyana and Malony, Allen D. “Autotuning GPU Kernels via Static and Predictive Analysis“, 2017
2017
Cited alongside, same era.
Amir H. Ashouri, William Killian, John Cavazos, Gianluca Palermo, Cristina Silvano “A Survey on Compiler Autotuning using Machine Learning,“ ACM Computing Surveys 2018
2018
Cited alongside, same era.
Ben-Nun, Tal and Jakobovits, Alice Shoshana and Hoefler, Torsten, “Neural Code Comprehension: A Learnable Representation of Code Semantics,”, IEEE, November 2018, p219–232
2018
Cited alongside, same era.
Mishra, Alok and Malik, Abid M and Chapman, Barbara, “Using Machine Learning for OpenMP GPU Offloading in LLVM,”, ACM SRC to be held at SC20, 2020
2020
Cited alongside, same era.
Cummins, Chris and Fisches, Zacharias V. and Ben-Nun, Tal and Hoefler, Torsten and Leather, Hugh, “ProGraML: Graph-based Deep Learning for Program Optimization and Analysis,”, March 2020
Brauckmann, Alexander and Goens, Andrés and Ertel, Sebastian and Castrillon, Jeronimo, “Compiler-based graph representations for deep learning models of code,”, ACM, February 2020, p201–211
2020
Later among the works it cites.
Kuznetsov, Evgeny & Kondratyuk, Nikolay & Logunov, Mikhail & Nikolskiy, Vsevolod & Stegailov, Vladimir. “Performance and Portability of State-of-Art Molecular Dynamics Software on Modern GPUs.“ 2020
2020
Later among the works it cites.
Bill Fiser, Sebastian Jodłowski “BEST PRACTICES WHEN BENCHMARKING CUDA APPLICATIONS“, accessed: 01.10.2020
2020
Later among the works it cites.
Ying Hu, Shane A Story, “Tips to Measure the Performance of Matrix Multiplication Using Intel® MKL“, accessed: 07.011.2020
2020
Later among the works it cites.
Ilya Grigorik “GH Archive.“, Accessed January 31, 2021. https://www.gharchive.org/
2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Closest in time.