Fetching the paper…
Reading the bibliography…
We systematically study the quality of 4,066 ChatGPT-generated code implemented in two popular programming languages, i.e., Java and Python, for 2,033 programming tasks.
On a test of whether one of two random variables is stochastically larger than the other
Henry B Mann and Donald R Whitney. 1947 · 1947
Earlier work this paper cites.
Checkstyle
Oliver Burn. 2003 · 2003
Earlier work this paper cites.
PMD applied . Vol. 10
Tom Copeland. 2005 · 2005
Earlier work this paper cites.
Card sorting: Designing usable categories
Donna Spencer. 2009 · 2009
Earlier work this paper cites.
Flake8: Your tool for style guide enforcement
Ian Cordasco and Tarek Ziade. 2010 · 2010
Earlier work this paper cites.
Cliff’s Delta Calculator: A non-parametric effect size program for two groups of observations
Guillermo Macbeth, Eugenia Razumiejczyk, and Rubén Daniel Ledesma. 2011 · 2011
Earlier work this paper cites.
How practitioners perceive the relevance of software engineering research. In Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering . 415–425
David Lo, Nachiappan Nagappan, and Thomas Zimmermann. 2015 · 2015
Earlier work this paper cites.
How Android app developers manage power consumption? An empirical study by mining power management commits. In Proceedings of the 13th International Conference on Mining Software Repositories . 37–48
Lingfeng Bao, David Lo, Xin Xia, Xinyu Wang, and Cong Tian. 2016 · 2016
Earlier work this paper cites.
A large scale study of multiple programming languages and code quality. In 2016 IEEE 23Rd international conference on software analysis, evolution, and reengineering (SANER) , Vol. 1. IEEE, 563–573
Pavneet Singh Kochhar, Dinusha Wijedasa, and David Lo. 2016 · 2016
Earlier work this paper cites.
Code quality issues in student programs. In Proceedings of the 2017 ACM Conference on Innovation and Technology in Computer Science Education . 110–115
Hieke Keuning, Bastiaan Heeren, and Johan Jeuring. 2017 · 2017
Earlier work this paper cites.
An empirical study of code smells in javascript projects. In 2017 IEEE 24th international conference on software analysis, evolution and reengineering (SANER) . IEEE, 294–305
Amir Saboury, Pooya Musavi, Foutse Khomh, and Giulio Antoniol. 2017 · 2017
Earlier work this paper cites.
Bug characteristics in blockchain systems: a large-scale empirical study. In 2017 IEEE/ACM 14th International Conference on Mining Software Repositories (MSR) . IEEE, 413–424
Zhiyuan Wan, David Lo, Xin Xia, and Liang Cai. 2017 · 2017
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages. In Findings of the Association for Computational Linguistics: EMNLP 2020 . 1536–1547
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Earlier work this paper cites.
Learning to summarize with human feedback
Nisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F Christiano. 2020 · 2020
Earlier work this paper cites.
Configuration smells in continuous delivery pipelines: a linter and a six-month study on GitLab. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 327–337
Carmine Vassallo, Sebastian Proksch, Anna Jancso, Harald C Gall, and Massimiliano Di Penta. 2020 · 2020
Earlier work this paper cites.
Predicting defective lines using a model-agnostic technique
Supatsara Wattanakriengkrai, Patanamon Thongtanunam, Chakkrit Tantithamthavorn, Hideaki Hata, and Kenichi Matsumoto. 2020 · 2020
Earlier work this paper cites.
DomBERT: Domain-oriented Language Model for Aspect-based Sentiment Analysis. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, Online, 1725–1731
Hu Xu, Bing Liu, Lei Shu, and Philip Yu. 2020 · 2020
Earlier work this paper cites.
Sentiment analysis for software engineering: How far can pre-trained transformer models go?. In 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 70–80
Ting Zhang, Bowen Xu, Ferdian Thung, Stefanus Agus Haryono, David Lo, and Lingxiao Jiang. 2020 · 2020
Earlier work this paper cites.
Program synthesis with large language models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, et al · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Cited alongside, same era.
SimTyper: sound type inference for Ruby using type equality prediction
Milod Kazerounian, Jeffrey S Foster, and Bonan Min. 2021 · 2021
Cited alongside, same era.
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . 8696–8708
Yue Wang, Weishi Wang, Shafiq Joty, and Steven CH Hoi. 2021 · 2021
Cited alongside, same era.
Few-shot training LLMs for project-specific code-summarization. In Proceedings of the 37th IEEE/ACM International Conference on Automated Software Engineering . 1–5
Toufique Ahmed and Premkumar Devanbu. 2022 · 2022
Cited alongside, same era.
Self-collaboration Code Generation via ChatGPT
Yihong Dong, Xue Jiang, Zhi Jin, and Ge Li. 2023 · 2023
Closest in time.
Automated Repair of Programs from Large Language Models. In the 45th International Conference on Software Repositories (ICSE)
Zhiyu Fan, Xiang Gao, Abhik Roychoudhury, and Shin Hwei Tan. 2023 · 2023
Closest in time.
Constructing Effective In-Context Demonstration for Code Intelligence Tasks: An Empirical Study
Shuzheng Gao, Xin-Cheng Wen, Cuiyun Gao, Wenxuan Wang, and Michael R Lyu. 2023 · 2023
Closest in time.
GitHub Copilot
GitHub. 2023 · 2023
Closest in time.
Two Cybersecurity Concerns When Using ChatGPT For Software Development
Morey Haber. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bei Chen, Fengji Zhang, Anh Nguyen, Daoguang Zan, Zeqi Lin, Jian-Guang Lou, and Weizhu Chen. 2022 · 2022
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Cited alongside, same era.
Incoder: A generative model for code infilling and synthesis
Daniel Fried, Armen Aghajanyan, Jessy Lin, Sida Wang, Eric Wallace, Freda Shi, Ruiqi Zhong, Wen-tau Yih, Luke Zettlemoyer, and Mike Lewis. 2022 · 2022
Cited alongside, same era.
News summarization and evaluation in the era of gpt-3
Tanya Goyal, Junyi Jessy Li, and Greg Durrett. 2022 · 2022
Cited alongside, same era.
Jigsaw: Large language models meet program synthesis. In Proceedings of the 44th International Conference on Software Engineering . 1219–1231
Naman Jain, Skanda Vaidyanath, Arun Iyer, Nagarajan Natarajan, Suresh Parthasarathy, Sriram Rajamani, and Rahul Sharma. 2022 · 2022
Cited alongside, same era.
AutoPruner: transformer-based call graph pruning. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 520–532
Thanh Le-Cong, Hong Jin Kang, Truong Giang Nguyen, Stefanus Agus Haryono, David Lo, Xuan-Bach D Le, and Quyet Thang Huynh. 2022 · 2022
Cited alongside, same era.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Cited alongside, same era.
An Empirical Evaluation of GitHub Copilot’s Code Suggestions. In Proceedings of the 19th International Conference on Mining Software Repositories (MSR)
Nhan Nguyen and Sarah Nadi. 2022 · 2022
Cited alongside, same era.
Amr Hendy, Mohamed Abdelrehim, Amr Sharaf, Vikas Raunak, Mohamed Gabr, Hitokazu Matsushita, Young Jin Kim, Mohamed Afify, and Hany Hassan Awadalla. 2023 · 2023
Closest in time.
Large language models for software engineering: A systematic literature review
Xinyi Hou, Yanjie Zhao, Yue Liu, Zhou Yang, Kailong Wang, Li Li, Xiapu Luo, David Lo, John Grundy, and Haoyu Wang. 2023 · 2023
Closest in time.
Invalidator: Automated patch correctness assessment via semantic and syntactic reasoning
Thanh Le-Cong, Duc-Minh Luong, Xuan Bach D Le, David Lo, Nhat-Hoa Tran, Bui Quang-Huy, and Quyet-Thang Huynh. 2023 · 2023
Closest in time.
1093. Statistics from a Large Sample
LeetCode. 2023 · 2023
Closest in time.
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2023 · 2023
Closest in time.
Multi-Granularity Detector for Vulnerability Fixes
Truong Giang Nguyen, Thanh Le-Cong, Hong Jin Kang, Ratnadira Widyasari, Chengran Yang, Zhipeng Zhao, Bowen Xu, Jiayuan Zhou, Xin Xia, Ahmed E Hassan, et al · 2023
Closest in time.
Quantifying Memorization Across Neural Language Models. In The Eleventh International Conference on Learning Representations . 70–80
Carlini Nicholas, Ippolito Daphne, Jagielski Matthew, Lee Katherine, Tramer Florian, and Zhang Chiyuan. 2023 · 2023
Closest in time.
Is ChatGPT a cybersecurity threat?
Carly Page. 2023 · 2023
Closest in time.
Pitfalls in Language Models for Code Intelligence: A Taxonomy and Survey
Xinyu She, Yue Liu, Yanjie Zhao, Yiling He, Li Li, Chakkrit Tantithamthavorn, Zhan Qin, and Haoyu Wang. 2023 · 2023
Closest in time.
ChatGPT’s traffic overview
SimilarWeb. 2023 · 2023
Closest in time.
ChatGPT reaches 100 million users two months after launch
TheGuardian. 2022 · 2023
Closest in time.
Is ChatGPT the Ultimate Programming Assistant–How far is it?
Haoye Tian, Weiqi Lu, Tsz On Li, Xunzhu Tang, Shing-Chi Cheung, Jacques Klein, and Tegawendé F Bissyandé. 2023 · 2023
Closest in time.
Automated program repair in the era of large pre-trained language models. In Proceedings of the 45th International Conference on Software Engineering (ICSE 2023). Association for Computing Machinery
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Closest in time.
Conversational automated program repair
Chunqiu Steven Xia and Lingming Zhang. 2023 · 2023
Closest in time.