Fetching the paper…
Reading the bibliography…
While CUDA has become a major parallel computing platform and programming model for general-purpose GPU computing, CUDA-induced bug patterns have not yet been well explored.
zhxfl. 2019 · 1905
Earlier work this paper cites.
An Empirical Comparison of Monitoring Algorithms for Access Anomaly Detection. In Proceedings of the Second ACM SIGPLAN Symposium on Principles &Amp; Practice of Parallel Programming (PPOPP ’90) . ACM, New York, NY, USA, 1–10
A. Dinning and E. Schonberg. 1990 · 1990
Earlier work this paper cites.
Improving the Accuracy of Data Race Detection
Robert H. B. Netzer and Barton P. Miller. 1991 · 1991
Earlier work this paper cites.
Eraser: A Dynamic Data Race Detector for Multithreaded Programs
Stefan Savage, Michael Burrows, Greg Nelson, Patrick Sobalvarro, and Thomas Anderson. 1997 · 1997
Earlier work this paper cites.
Intelligence Through Simulated Evolution: Forty Years of Evolutionary Programming
Lawrence J. Fogel. 1999 · 1999
Earlier work this paper cites.
Efficient and Precise Datarace Detection for Multithreaded Object-oriented Programs
Jong-Deok Choi, Keunwoo Lee, Alexey Loginov, Robert O’Callahan, Vivek Sarkar, and Manu Sridharan. 2002 · 2002
Earlier work this paper cites.
Empirical Evaluation of the Tarantula Automatic Fault-localization Technique. In Proceedings of the 20th IEEE/ACM International Conference on Automated Software Engineering (ASE ’05) . ACM, New York, NY, USA, 273–282
James A. Jones and Mary Jean Harrold. 2005 · 2005
Earlier work this paper cites.
Automated Dynamic Analysis of CUDA Programs. In Third Workshop on Software Tools for MultiCore Systems
M. Boyer, K. Skadron, and W. Weimer. 2008 · 2008
Earlier work this paper cites.
A performance study of general-purpose applications on graphics processors using CUDA
Shuai Che, Michael Boyer, Jiayuan Meng, David Tarjan, Jeremy W. Sheaffer, and Kevin Skadron. 2008 · 2008
Earlier work this paper cites.
Scalable SMT-based Verification of GPU Kernel Functions. In Proceedings of the Eighteenth ACM SIGSOFT International Symposium on Foundations of Software Engineering (FSE ’10) . ACM, New York, NY, USA, 187–196
Guodong Li and Ganesh Gopalakrishnan. 2010 · 2010
Earlier work this paper cites.
From CUDA to OpenCL: Towards a performance-portable solution for multi-platform GPU programming
Peng Du, Rick Weber, Piotr Luszczek, Stanimire Tomov, Gregory Peterson, and Jack Dongarra. 2012 · 2011
Earlier work this paper cites.
GRace: A low-overhead mechanism for detecting data races in GPU programs. In PPoPP
Mai Zheng, Vignesh T. Ravi, Feng Qin, and Gagan Agrawal. 2011 · 2011
Earlier work this paper cites.
GPUVerify: A Verifier for GPU Kernels
Adam Betts, Nathan Chong, Alastair Donaldson, Shaz Qadeer, and Paul Thomson. 2012 · 2012
Earlier work this paper cites.
A quantitative study of irregular programs on GPUs. In 2012 IEEE International Symposium on Workload Characterization (IISWC) . 141–151
M. Burtscher, R. Nasre, and K. Pingali. 2012 · 2012
Earlier work this paper cites.
A systematic study of automated program repair: Fixing 55 out of 105 bugs for $8 each. In 2012 34th International Conference on Software Engineering (ICSE) . 3–13
C. Le Goues, M. Dewey-Vogt, S. Forrest, and W. Weimer. 2012 · 2012
Cited alongside, same era.
Verifying GPU Kernels by Test Amplification. In Proceedings of the 33rd ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI ’12) . ACM, New York, NY, USA, 383–394
Alan Leung, Manish Gupta, Yuvraj Agarwal, Rajesh Gupta, Ranjit Jhala, and Sorin Lerner. 2012 · 2012
Cited alongside, same era.
GKLEE: Concolic Verification and Test Generation for GPUs
Guodong Li, Peng Li, Geof Sawaya, Ganesh Gopalakrishnan, Indradeep Ghosh, and Sreeranga P. Rajan. 2012b · 2012
Cited alongside, same era.
Parametric flows: Automated behavior equivalencing for symbolic analysis of races in CUDA programs. In SC ’12: Proceedings of the International Conference on High Performance Computing, Networking, Storage and Analysis . 1–10
P. Li, G. Li, and G. Gopalakrishnan. 2012a · 2012
Cited alongside, same era.
Verifying CUDA Programs Using SMT-based Context-bounded Model Checking. In Proceedings of the 31st Annual ACM Symposium on Applied Computing (SAC ’16) . ACM, New York, NY, USA, 1648–1653
Phillipe Pereira, Higo Albuquerque, Hendrio Marques, Isabela Silva, Celso Carvalho, Lucas Cordeiro, Vanessa Santos, and Ricardo Ferreira. 2016 · 2016
Later among the works it cites.
BARRACUDA: Binary-level Analysis of Runtime RAces in CUDA Programs. In Proceedings of the 38th ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI 2017) . ACM, New York, NY, USA, 126–140
Ariel Eizenberg, Yuanfeng Peng, Toma Pigli, William Mansky, and Joseph Devietti. 2017 · 2017
Later among the works it cites.
LD: Low-Overhead GPU Race Detection Without Access Monitoring
Pengcheng Li, Xiaoyu Hu, Dong Chen, Jacob Brock, Hao Luo, Eddy Z. Zhang, and Chen Ding. 2017 · 2017
Later among the works it cites.
CURD: A Dynamic CUDA Race Detector. In Proceedings of the 39th ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI 2018) . ACM, New York, NY, USA, 390–403
Yuanfeng Peng, Vinod Grover, and Joseph Devietti. 2018a · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fixing Performance Bugs: An Empirical Study of Open-Source GPGPU Programs. In 2012 41st International Conference on Parallel Processing . 329–339
Y. Yang, P. Xiang, M. Mantor, and H. Zhou. 2012 · 2012
Cited alongside, same era.
Barrier Invariants: A Shared State Abstraction for the Analysis of Data-dependent GPU Kernels
Nathan Chong, Alastair F. Donaldson, Paul H.J. Kelly, Jeroen Ketema, and Shaz Qadeer. 2013 · 2013
Cited alongside, same era.
Interleaving and Lock-Step Semantics for Analysis and Verification of GPU Kernels. In Programming Languages and Systems , Matthias Felleisen and Philippa Gardner (Eds.). Springer Berlin Heidelberg, Berlin, Heidelberg, 270–289
Peter Collingbourne, Alastair F. Donaldson, Jeroen Ketema, and Shaz Qadeer. 2013 · 2013
Cited alongside, same era.
Programming CUDA and OpenCL: A Case Study Using Modern C++ Libraries
D. Demidov, K. Ahnert, K. Rupp, and P. Gottschling. 2013 · 2013
Cited alongside, same era.
GMRace: Detecting Data Races in GPU Programs via a Low-Overhead Scheme
M. Zheng, V. T. Ravi, F. Qin, and G. Agrawal. 2014 · 2013
Cited alongside, same era.
Engineering a Static Verification Tool for GPU Kernels. In Proceedings of the 16th International Conference on Computer Aided Verification - Volume 8559 . Springer-Verlag, Berlin, Heidelberg, 226–242
Ethel Bardsley, Adam Betts, Nathan Chong, Peter Collingbourne, Pantazis Deligiannis, Alastair F. Donaldson, Jeroen Ketema, Daniel Liew, and Shaz Qadeer. 2014 · 2014
Cited alongside, same era.
Practical Symbolic Race Checking of GPU Programs. In Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis (SC ’14) . IEEE Press, Piscataway, NJ, USA, 179–190
Peng Li, Guodong Li, and Ganesh Gopalakrishnan. 2014b · 2014
Cited alongside, same era.
Characterizing and Detecting Performance Bugs for Smartphone Applications. In Proceedings of the 36th International Conference on Software Engineering (ICSE 2014) . ACM, New York, NY, USA, 1013–1024
Yepang Liu, Chang Xu, and Shing-Chi Cheung. 2014 · 2014
Cited alongside, same era.
CURD: A Dynamic CUDA Race Detector
Yuanfeng Peng, Vinod Grover, and Joseph Devietti. 2018b · 2018
Later among the works it cites.
An Empirical Study on TensorFlow Program Bugs. In Proceedings of the 27th ACM SIGSOFT International Symposium on Software Testing and Analysis (ISSTA 2018) . ACM, New York, NY, USA, 129–140
Yuhao Zhang, Yifan Chen, Shing-Chi Cheung, Yingfei Xiong, and Lu Zhang. 2018 · 2018
Later among the works it cites.
Cauchy distribution
2019 · 2019
Closest in time.
CUDA program introduction
2019 · 2019
Closest in time.
GPGPU introduction
2019 · 2019
Closest in time.
Normal distribution
2019 · 2019
Closest in time.
Racecheck Tool
2019 · 2019
Closest in time.
The Simulee project
2019 · 2019
Closest in time.
cuda-convnet2
akrizhevsky. 2019 · 2019
Closest in time.
ArrayFire
arrayfire. 2019 · 2019
Closest in time.