Fetching the paper…
Reading the bibliography…
This paper presents the FormAI dataset, a large collection of 112, 000 AI-generated compilable and independent C programs with vulnerability classification.
Yaqin Zhou, Shangqing Liu, Jingkai Siow, Xiaoning Du, and Yang Liu. 2019b · 1909
Earlier work this paper cites.
Software verification and validation: an overview
D.R. Wallace and R.U. Fujii. 1989 · 1989
Earlier work this paper cites.
Testing software for characteristics other than correctness: Safety, failure tolerance, and security
Je rey Voas. 1996 · 1996
Earlier work this paper cites.
Compilers: Principles, Techniques, And Tools (2nd ed.)
Alfred V. Aho, Monica S. Lam, Ravi Sethi, and Jeffrey D. Ullman. 2006 · 2006
Earlier work this paper cites.
A Survey of Automated Techniques for Formal Software Verification
Vijay D’Silva, Daniel Kroening, and Georg Weissenbacher. 2008 · 2008
Earlier work this paper cites.
SMT-Based Bounded Model Checking for Embedded ANSI-C Software
Lucas Cordeiro, Bernd Fischer, and Joao Marques-Silva. 2012 · 2011
Earlier work this paper cites.
How developers use data race detection tools. In Proceedings of the 5th Workshop on Evaluation and Usability of Programming Languages and Tools . 43–51
Caitlin Sadowski and Jaeheon Yi. 2014 · 2014
Earlier work this paper cites.
ESBMC: 5.0: An Industrial-Strength Model Checker
Mikhail R. Gadelha, Felipe R. Monteiro, Jeremy Morse, Lucas C. Cordeiro, Bernd Fischer, and Denis A. Nicole. 2023 · 2015
Earlier work this paper cites.
Deep learning code fragments for code clone detection. In Proceedings of the 31st IEEE/ACM international conference on automated software engineering . 87–98
Martin White, Michele Tufano, Christopher Vendome, and Denys Poshyvanyk. 2016 · 2016
Earlier work this paper cites.
A Software Assurance Reference Dataset: Thousands of Programs With Known Bugs
Paul E. Black. 2018 · 2018
Earlier work this paper cites.
ESBMC 5.0: an industrial-strength C model checker. In Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering . 888–891
Mikhail R Gadelha, Felipe R Monteiro, Jeremy Morse, Lucas C Cordeiro, Bernd Fischer, and Denis A Nicole. 2018 · 2018
Earlier work this paper cites.
Draper VDISC Dataset - Vulnerability Detection in Source Code
Louis Kim and Rebecca Russell. 2018 · 2018
Earlier work this paper cites.
Automated Vulnerability Detection in Source Code Using Deep Representation Learning. In 2018 17th IEEE International Conference on Machine Learning and Applications (ICMLA) . 757–762
Rebecca Russell, Louis Kim, Lei Hamilton, Tomo Lazovich, Jacob Harer, Onur Ozdemir, Paul Ellingwood, and Marc McConley. 2018 · 2018
Earlier work this paper cites.
Deepsim: deep learning code functional similarity. In Proceedings of the 2018 26th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 141–151
Gang Zhao and Jeff Huang. 2018 · 2018
Earlier work this paper cites.
SMT-based refutation of spurious bug reports in the clang static analyzer. In Proceedings of the 41st International Conference on Software Engineering: Companion Proceedings, ICSE 2019, Montreal, QC, Canada, May 25-31, 2019 , Joanne M. Atlee, Tevfik Bultan, and Jon Whittle (Eds.). IEEE / ACM, 11–14
Mikhail Y. R. Gadelha, Enrico Steffinlongo, Lucas C. Cordeiro, Bernd Fischer, and Denis A. Nicole. 2019 · 2019
Earlier work this paper cites.
A C/C++ Code Vulnerability Dataset with Code Changes and CVE Summaries. In Proceedings of the 17th International Conference on Mining Software Repositories (MSR ’20) . Association for Computing Machinery, New York, NY, USA, 508–512
Jiahao Fan, Yi Li, Shaohua Wang, and Tien N. Nguyen. 2020 · 2020
Cited alongside, same era.
Ensuring Dataset Quality for Machine Learning Certification. In 2020 IEEE International Symposium on Software Reliability Engineering Workshops (ISSREW) . 275–282
S. Picard, C. Chapdelaine, C. Cappi, L. Gardes, E. Jenn, B. Lefevre, and T. Soumarmon. 2020 · 2020
Cited alongside, same era.
Deep Learning Based Vulnerability Detection: Are We There Yet?
Saikat Chakraborty, Rahul Krishna, Yangruibo Ding, and Baishakhi Ray. 2022 · 2021
Cited alongside, same era.
Towards Accountability for Machine Learning Datasets: Practices from Software Engineering and Infrastructure. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’21) . Association for Computing Machinery, New York, NY, USA, 560–575
Ben Hutchinson, Andrew Smart, Alex Hanna, Emily Denton, Christina Greer, Oddur Kjartansson, Parker Barnes, and Margaret Mitchell. 2021 · 2021
A Code Centric Evaluation of C/C++ Vulnerability Datasets for Deep Learning Based Vulnerability Detection Techniques. In Proceedings of the 16th Innovations in Software Engineering Conference . 1–10
Ridhi Jain, Nicole Gervasoni, Mthandazo Ndhlovu, and Sanjay Rawat. 2023 · 2023
Closest in time.
How Secure is Code Generated by ChatGPT?
Raphaël Khoury, Anderson R. Avila, Jacob Brunelle, and Baba Mamadou Camara. 2023 · 2023
Closest in time.
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2023 · 2023
Closest in time.
The Scope of ChatGPT in Software Engineering: A Thorough Investigation
Wei Ma, Shangqing Liu, Wenhan Wang, Qiang Hu, Ye Liu, Cen Zhang, Liming Nie, and Yang Liu. 2023a · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The Juliet 1.1 C/C++ and Java Test Suite
Frederick E. Boland Jr and Paul E. Black. 2012 · 2021
Cited alongside, same era.
Asleep at the Keyboard? Assessing the Security of GitHub Copilot’s Code Contributions
Hammond Pearce, Baleegh Ahmad, Benjamin Tan, Brendan Dolan-Gavitt, and Ramesh Karri. 2021 · 2021
Cited alongside, same era.
Embedded Software Programming Languages: Pros, Cons, and Comparisons of Popular Languages
Risto Avila. 2022 · 2022
Cited alongside, same era.
Examining Zero-Shot Vulnerability Repair with Large Language Models
Hammond Pearce, Benjamin Tan, Baleegh Ahmad, Ramesh Karri, and Brendan Dolan-Gavitt. 2022 · 2022
Cited alongside, same era.
Do Users Write More Insecure Code with AI Assistants?
Neil Perry, Megha Srivastava, Deepak Kumar, and Dan Boneh. 2022 · 2022
Cited alongside, same era.
Competition on Software Verification and Witness Validation: SV-COMP 2023. In Tools and Algorithms for the Construction and Analysis of Systems , Sriram Sankaranarayanan and Natasha Sharygina (Eds.). Springer Nature Switzerland, Cham, 495–522
Dirk Beyer. 2023 · 2023
Cited alongside, same era.
CodeTF: One-stop Transformer Library for State-of-the-art Code LLM
Nghi D. Q. Bui, Hung Le, Yue Wang, Junnan Li, Akhilesh Deepak Gotmare, and Steven C. H. Hoi. 2023 · 2023
Cited alongside, same era.
Yiannis Charalambous, Norbert Tihanyi, Ridhi Jain, Youcheng Sun, Mohamed Amine Ferrag, and Lucas C. Cordeiro. 2023 · 2023
Cited alongside, same era.
Wei Ma, Shangqing Liu, Wenhan Wang, Qiang Hu, Ye Liu, Cen Zhang, Liming Nie, and Yang Liu. 2023b · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
The Programmer’s Assistant: Conversational Interaction with a Large Language Model for Software Development. In Proceedings of the 28th International Conference on Intelligent User Interfaces (IUI ’23) . Association for Computing Machinery, New York, NY, USA, 491–514
Steven I. Ross, Fernando Martinez, Stephanie Houde, Michael Muller, and Justin D. Weisz. 2023 · 2023
Closest in time.
Lost at C: A User Study on the Security Implications of Large Language Model Code Assistants
Gustavo Sandoval, Hammond Pearce, Teo Nys, Ramesh Karri, Siddharth Garg, and Brendan Dolan-Gavitt. 2023 · 2023
Closest in time.
The Curse of Recursion: Training on Generated Data Makes Models Forget
Ilia Shumailov, Zakhar Shumaylov, Yiren Zhao, Yarin Gal, Nicolas Papernot, and Ross Anderson. 2023 · 2023
Closest in time.
Is ChatGPT free and unlimited? In short - yes
Funmi Looi Somoye. 2023 · 2023
Closest in time.
ChatGPT writes insecure code
Jovi Umawing. 2023 · 2023
Closest in time.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Closest in time.
A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT
Jules White, Quchen Fu, Sam Hays, Michael Sandborn, Carlos Olea, Henry Gilbert, Ashraf Elnashar, Jesse Spencer-Smith, and Douglas C. Schmidt. 2023 · 2023
Closest in time.
Prompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native Services
Zhenchang Xing, Qing Huang, Yu Cheng, Liming Zhu, Qinghua Lu, and Xiwei Xu. 2023 · 2023
Closest in time.
Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Thomas L. Griffiths, Yuan Cao, and Karthik Narasimhan. 2023 · 2023
Closest in time.