Fetching the paper…
Reading the bibliography…
The latest paradigm shift in software development brings in the innovation and automation afforded by Large Language Models (LLMs), showcased by Generative Pre-trained Transformer (GPT), which has shown remarkable capacity to generate code autonomously, significantly reducing the manual effort required for various programming tasks.
Search-based software testing: Past, present and future
Phil McMinn · 2011
Earlier work this paper cites.
On the naturalness of software
Abram Hindle, Earl T Barr, Mark Gabel, Zhendong Su, and Premkumar Devanbu · 2016
Earlier work this paper cites.
Directed greybox fuzzing
Marcel Böhme, Van-Thuan Pham, Manh-Dung Nguyen, and Abhik Roychoudhury · 2017
Earlier work this paper cites.
Measuring program comprehension: A large-scale field study with professionals
Xin Xia, Lingfeng Bao, David Lo, Zhenchang Xing, Ahmed E Hassan, and Shanping Li · 2017
Earlier work this paper cites.
Metamorphic testing: A review of challenges and opportunities
Tsong Yueh Chen, Fei-Ching Kuo, Huai Liu, Pak-Lok Poon, Dave Towey, TH Tse, and Zhi Quan Zhou · 2018
Earlier work this paper cites.
Api method recommendation without worrying about the task-api knowledge gap
Qiao Huang, Xin Xia, Zhenchang Xing, David Lo, and Xinyu Wang · 2018
Earlier work this paper cites.
Effective api recommendation without historical software repositories
Xiaoyu Liu, LiGuo Huang, and Vincent Ng · 2018
Earlier work this paper cites.
A deep neural network language model with contexts for source code
Anh Tuan Nguyen, Trong Duc Nguyen, Hung Dang Phan, and Tien N Nguyen · 2018
Earlier work this paper cites.
Paraphrasing with large language models
Sam Witteveen and Martin Andrews · 2019
Earlier work this paper cites.
Metamorphic testing and certified mitigation of fairness violations in nlp models
Pingchuan Ma, Shuai Wang, and Jin Liu · 2020
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Cited alongside, same era.
Automated source code generation and auto-completion using deep learning: Comparing and discussing current language model-related approaches
Juan Cruz-Benito, Sanjay Vishwakarma, Francisco Martin-Fernandez, and Ismael Faro · 2021
Cited alongside, same era.
Paraphrase detection using lstm networks and handcrafted features
Hassan Shahmohammadi, MirHossein Dezfoulian, and Muharram Mansoorizadeh · 2021
Cited alongside, same era.
Autotrainer: An automatic dnn training problem detection and repair system
Xiaoyu Zhang, Juan Zhai, Shiqing Ma, and Chao Shen · 2021
Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models
Priyan Vaithilingam, Tianyi Zhang, and Elena L Glassman · 2022
Later among the works it cites.
Unveiling hidden dnn defects with decision-based metamorphic testing
Yuanyuan Yuan, Qi Pang, and Shuai Wang · 2022
Later among the works it cites.
Programming is hard-or at least it used to be: Educational opportunities and challenges of ai code generation
Brett A Becker, Paul Denny, James Finnie-Ansley, Andrew Luxton-Reilly, James Prather, and Eddie Antonio Santos · 2023
Later among the works it cites.
Dissecting american fuzzy lop: a fuzzbench evaluation
Andrea Fioraldi, Alessandro Mantovani, Dominik Maier, and Davide Balzarotti · 2023
Later among the works it cites.
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gpt4: The ultimate brain
Gerardo Adesso · 2022
Cited alongside, same era.
Cracking the code: Co-coding with ai in creative programming education
Martin Jonsson and Jakob Tholander · 2022
Cited alongside, same era.
An empirical evaluation of github copilot’s code suggestions
Nhan Nguyen and Sarah Nadi · 2022
Cited alongside, same era.
Python-AFL
Python-AFL
Cited in the paper.
SonarQube
SonarQube
Cited in the paper.
SpotBugs
SpotBugs
Cited in the paper.
Octopack: Instruction tuning code large language models
Niklas Muennighoff, Qian Liu, Armel Zebaze, Qinkai Zheng, Binyuan Hui, Terry Yue Zhuo, Swayam Singh, Xiangru Tang, Leandro Von Werra, and Shayne Longpre · 2023
Later among the works it cites.
Lever: Learning to verify language-to-code generation with execution
Ansong Ni, Srini Iyer, Dragomir Radev, Veselin Stoyanov, Wen-tau Yih, Sida Wang, and Xi Victoria Lin · 2023
Later among the works it cites.
Llm is like a box of chocolates: the non-determinism of chatgpt in code generation
Shuyin Ouyang, Jie M Zhang, Mark Harman, and Meng Wang · 2023
Later among the works it cites.