Fetching the paper…
Reading the bibliography…
Recently, the ChatGPT LLM has received great attention: it can be used as a bot for discussing source code, prompting it to suggest changes, provide descriptions or even generate code.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Individual Comparisons by Ranking Methods
F. Wilcoxon. 1945 · 1945
Earlier work this paper cites.
On a Test of Whether One of Two Random Variables Is Stochastically Larger than the Other
Henry B Mann and Donald R. Whitney. 1947 · 1947
Earlier work this paper cites.
Automatic code generation from design patterns
Frank J. Budinsky, Marilyn A. Finnie, John M. Vlissides, and Patsy S. Yu. 1996 · 1996
Earlier work this paper cites.
Comprehending reality-practical barriers to industrial adoption of software maintenance automation. In 11th IEEE International Workshop on Program Comprehension, 2003. IEEE, 196–205
James R Cordy. 2003 · 2003
Earlier work this paper cites.
A systematic study of automated program repair: Fixing 55 out of 105 bugs for $8 each. In 2012 34th International Conference on Software Engineering (ICSE) . IEEE, 3–13
Claire Le Goues, Michael Dewey-Vogt, Stephanie Forrest, and Westley Weimer. 2012 · 2012
Earlier work this paper cites.
Improving automated source code summarization via an eye-tracking study of programmers. In Proceedings of the 36th international conference on Software engineering . 390–401
Paige Rodeghero, Collin McMillan, Paul W McBurney, Nigel Bosch, and Sidney D’Mello. 2014 · 2014
Earlier work this paper cites.
The ManyBugs and IntroClass benchmarks for automated repair of C programs
Claire Le Goues, Neal Holtschulte, Edward K Smith, Yuriy Brun, Premkumar Devanbu, Stephanie Forrest, and Westley Weimer. 2015 · 2015
Earlier work this paper cites.
Deepfix: Fixing common c language errors by deep learning. In Proceedings of the aaai conference on artificial intelligence , Vol. 31
Rahul Gupta, Soham Pal, Aditya Kanade, and Shirish Shevade. 2017 · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Automated clustering and program repair for introductory programming assignments
Sumit Gulwani, Ivan Radiček, and Florian Zuleger. 2018 · 2018
Earlier work this paper cites.
Deep code comment generation. In Proceedings of the 26th conference on program comprehension . 200–210
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
Improving automatic source code summarization via deep reinforcement learning. In Proceedings of the 33rd ACM/IEEE international conference on automated software engineering . 397–407
Yao Wan, Zhou Zhao, Min Yang, Guandong Xu, Haochao Ying, Jian Wu, and Philip S Yu. 2018 · 2018
Earlier work this paper cites.
AutoPandas: neural-backed generators for program synthesis
Rohan Bavishi, Caroline Lemieux, Roy Fox, Koushik Sen, and Ion Stoica. 2019 · 2019
Earlier work this paper cites.
SciBERT: A pretrained language model for scientific text
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Earlier work this paper cites.
Opportunities for software reuse in an uncertain world: From past to emerging trends
Rafael Capilla, Barbara Gallina, Carlos Cetina, and John Favaro. 2019 · 2019
Earlier work this paper cites.
Re-factoring based Program Repair applied to Programming Assignments. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE/ACM, 388–398
Yang Hu, Umair Z. Ahmed, Sergey Mechtaev, Ben Leong, and Abhik Roychoudhury. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Automatic source code summarization with extended tree-lstm. In 2019 International Joint Conference on Neural Networks (IJCNN) . IEEE, 1–8
Yusuke Shido, Yasuaki Kobayashi, Akihiro Yamamoto, Atsushi Miyamoto, and Tadayuki Matsumura. 2019 · 2019
Earlier work this paper cites.
Code generation as a dual task of code summarization
Bolin Wei, Ge Li, Xin Xia, Zhiyi Fu, and Zhi Jin. 2019 · 2019
Earlier work this paper cites.
A survey on research of code comment. In Proceedings of the 2019 3rd International Conference on Management Engineering, Software Engineering and Service Sciences . 45–51
Bai Yang, Zhang Liping, and Zhao Fengrong. 2019 · 2019
Earlier work this paper cites.
A transformer-based approach for source code summarization
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2020 · 2020
Earlier work this paper cites.
Codit: Code editing with tree-based neural models
Saikat Chakraborty, Yangruibo Ding, Miltiadis Allamanis, and Baishakhi Ray. 2020 · 2020
Earlier work this paper cites.
Deep code comment generation with hybrid lexical and syntactical information
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2020 · 2020
Earlier work this paper cites.
Improved code summarization via a graph neural network. In Proceedings of the 28th international conference on program comprehension . 184–195
Alexander LeClair, Sakib Haque, Lingfei Wu, and Collin McMillan. 2020 · 2020
Earlier work this paper cites.
DeepCommenter: a deep code comment generation tool with hybrid lexical and syntactical information. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1571–1575
Boao Li, Meng Yan, Xin Xia, Xing Hu, Ge Li, and David Lo. 2020 · 2020
Earlier work this paper cites.
A self-attentional neural architecture for code completion with multi-task learning. In Proceedings of the 28th International Conference on Program Comprehension . 37–47
Fang Liu, Ge Li, Bolin Wei, Xin Xia, Zhiyi Fu, and Zhi Jin. 2020 · 2020
Earlier work this paper cites.
Quality of automated program repair on real-world defects
Manish Motwani, Mauricio Soto, Yuriy Brun, Rene Just, and Claire Le Goues. 2020 · 2020
Earlier work this paper cites.
Pre-trained models for natural language processing: A survey
Xipeng Qiu, Tianxiang Sun, Yige Xu, Yunfan Shao, Ning Dai, and Xuanjing Huang. 2020 · 2020
Cited alongside, same era.
A human study of comprehension and code summarization. In Proceedings of the 28th International Conference on Program Comprehension . 2–13
Sean Stapleton, Yashmeet Gambhir, Alexander LeClair, Zachary Eberhart, Westley Weimer, Kevin Leach, and Yu Huang. 2020 · 2020
Cited alongside, same era.
Intellicode compose: Code generation using transformer. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1433–1443
Alexey Svyatkovskiy, Shao Kun Deng, Shengyu Fu, and Neel Sundaresan. 2020 · 2020
Cited alongside, same era.
Evaluating representation learning of code changes for predicting patch correctness in program repair. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 981–992
Haoye Tian, Kui Liu, Abdoul Kader Kaboré, Anil Koyuncu, Li Li, Jacques Klein, and Tegawendé F Bissyandé. 2020 · 2020
Cited alongside, same era.
Trust enhancement issues in program repair. In Proceedings of the 44th International Conference on Software Engineering . 2228–2240
Yannic Noller, Ridwan Shariffdeen, Xiang Gao, and Abhik Roychoudhury. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Can OpenAI’s codex fix bugs? an evaluation on QuixBugs. In Proceedings of the Third International Workshop on Automated Program Repair . 69–75
Julian Aron Prenner, Hlib Babii, and Romain Robbes. 2022 · 2022
Later among the works it cites.
Automatic generation of programming exercises and code explanations using large language models. In Proceedings of the 2022 ACM Conference on International Computing Education Research-Volume 1 . 27–43
Sami Sarsa, Paul Denny, Arto Hellas, and Juho Leinonen. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Reinforcement-learning-guided source code summarization using hierarchical attention
Wenhua Wang, Yuqun Zhang, Yulei Sui, Yao Wan, Zhou Zhao, Jian Wu, S Yu Philip, and Guandong Xu. 2020 · 2020
Cited alongside, same era.
Sentiment analysis for software engineering: How far can pre-trained transformer models go?. In 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 70–80
Ting Zhang, Bowen Xu, Ferdian Thung, Stefanus Agus Haryono, David Lo, and Lingxiao Jiang. 2020 · 2020
Cited alongside, same era.
Unified pre-training for program understanding and generation
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021 · 2021
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Cited alongside, same era.
Domain-specific language model pretraining for biomedical natural language processing
Yu Gu, Robert Tinn, Hao Cheng, Michael Lucas, Naoto Usuyama, Xiaodong Liu, Tristan Naumann, Jianfeng Gao, and Hoifung Poon. 2021 · 2021
Cited alongside, same era.
BERT: a review of applications in natural language processing and understanding
MV Koroteev. 2021 · 2021
Cited alongside, same era.
Automatic program repair
Claire Le Goues, Michael Pradel, Abhik Roychoudhury, and Satish Chandra. 2021 · 2021
Cited alongside, same era.
Prefix-Tuning: Optimizing Continuous Prompts for Generation. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Online, 4582–4597
Xiang Lisa Li and Percy Liang. 2021 · 2021
Cited alongside, same era.
On the evaluation of neural code summarization. In Proceedings of the 44th International Conference on Software Engineering . 1597–1608
Ensheng Shi, Yanlin Wang, Lun Du, Junjie Chen, Shi Han, Hongyu Zhang, Dongmei Zhang, and Hongbin Sun. 2022 · 2022
Later among the works it cites.
An Empirical Study of Code Smells in Transformer-based Code Generation Techniques. In 2022 IEEE 22nd International Working Conference on Source Code Analysis and Manipulation (SCAM) . IEEE, 71–82
Mohammed Latif Siddiq, Shafayat H Majumder, Maisha R Mim, Sourov Jajodia, and Joanna CS Santos. 2022 · 2022
Later among the works it cites.
Natural language processing with transformers
Lewis Tunstall, Leandro Von Werra, and Thomas Wolf. 2022 · 2022
Later among the works it cites.
Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models. In Chi conference on human factors in computing systems extended abstracts . 1–7
Priyan Vaithilingam, Tianyi Zhang, and Elena L Glassman. 2022 · 2022
Later among the works it cites.
In-ide code generation from natural language: Promise and challenges
Frank F Xu, Bogdan Vasilescu, and Graham Neubig. 2022 · 2022
Later among the works it cites.
Neural program repair with execution-based backpropagation. In Proceedings of the 44th International Conference on Software Engineering . 1506–1518
He Ye, Matias Martinez, and Martin Monperrus. 2022 · 2022
Later among the works it cites.
Assessing the quality of GitHub copilot’s code generation. In Proceedings of the 18th International Conference on Predictive Models and Data Analytics in Software Engineering . 62–71
Burak Yetistiren, Isik Ozsoy, and Eray Tuzun. 2022 · 2022
Later among the works it cites.
Repairing Bugs in Python Assignments Using Large Language Models
Jialu Zhang, José Cambronero, Sumit Gulwani, Vu Le, Ruzica Piskac, Gustavo Soares, and Gust Verbruggen. 2022a · 2022
Later among the works it cites.
OpenAI’s GPT-4 AI Model Got Lazier and Dumber - ChatGPT
Alistair Barr. 2023 · 2023
Closest in time.
Programming is hard-or at least it used to be: Educational opportunities and challenges of ai code generation. In Proceedings of the 54th ACM Technical Symposium on Computer Science Education V. 1 . 500–506
Brett A Becker, Paul Denny, James Finnie-Ansley, Andrew Luxton-Reilly, James Prather, and Eddie Antonio Santos. 2023 · 2023
Closest in time.
How is ChatGPT’s behavior changing over time?
Lingjiao Chen, Matei Zaharia, and James Zou. 2023 · 2023
Closest in time.
Impact of Code Language Models on Automated Program Repair
Nan Jiang, Kevin Liu, Thibaud Lutellier, and Lin Tan. 2023a · 2023
Closest in time.
Knod: Domain knowledge distilled tree decoder for automated program repair
Nan Jiang, Thibaud Lutellier, Yiling Lou, Lin Tan, Dan Goldwasser, and Xiangyu Zhang. 2023b · 2023
Closest in time.
Studying the effect of AI Code Generators on Supporting Novice Learners in Introductory Programming. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems . 1–23
Majeed Kazemitabaar, Justin Chow, Carl Ka To Ma, Barbara J Ericson, David Weintrop, and Tovi Grossman. 2023 · 2023
Closest in time.
Natural language processing: State of the art, current trends and challenges
Diksha Khurana, Aditya Koli, Kiran Khatter, and Sukhdev Singh. 2023 · 2023
Closest in time.
LeetCode: The World’s Leading Online Programming
LeetCode. Accessed: 14 February 2023 · 2023
Closest in time.
Finding Failure-Inducing Test Cases with ChatGPT
Tsz-On Li, Wenxi Zong, Yibo Wang, Haoye Tian, Ying Wang, and Shing-Chi Cheung. 2023 · 2023
Closest in time.
Experiences from Using Code Explanations Generated by Large Language Models in a Web Software Development E-Book. In Proceedings of the 54th ACM Technical Symposium on Computer Science Education V. 1 . 931–937
Stephen MacNeil, Andrew Tran, Arto Hellas, Joanne Kim, Sami Sarsa, Paul Denny, Seth Bernstein, and Juho Leinonen. 2023 · 2023
Closest in time.
ChatGPT GPT-4 Response Quality And Performance Drop Issues To Be Investigated
Riya Madaan. 2023 · 2023
Closest in time.
An analysis of the automatic bug fixing performance of chatgpt
Dominik Sobania, Martin Briesch, Carol Hanna, and Justyna Petke. 2023 · 2023
Closest in time.
Use Chat GPT to Solve Programming Bugs
Nigar M Shafiq Surameery and Mohammed Y Shakor. 2023 · 2023
Closest in time.
Conversational automated program repair
Chunqiu Steven Xia and Lingming Zhang. 2023 · 2023
Closest in time.
GPT-4: The World’s Most Powerful AI Model - An Analysis of Its Performance Decline
Ikram Zakori. 2023 · 2023
Closest in time.
A convolutional attention network for extreme summarization of source code. In International conference on machine learning . PMLR, 2091–2100
Miltiadis Allamanis, Hao Peng, and Charles Sutton. 2016 · 2091
Closest in time.