Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have significantly impacted numerous domains, including Software Engineering (SE).
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning
Haokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin A Raffel. 2022a · 1965
Earlier work this paper cites.
Control flow analysis
Frances E Allen. 1970 · 1970
Earlier work this paper cites.
A deductive approach to program synthesis
Zohar Manna and Richard Waldinger. 1980 · 1980
Earlier work this paper cites.
Transparent logging as a technique for debugging complex distributed systems. In Proceedings of the 5th workshop on ACM SIGOPS European workshop: Models and paradigms for distributed systems structuring . 1–3
Mahadev Satyanarayanan, David C Steere, Masashi Kudo, and Hank Mashburn. 1992 · 1992
Earlier work this paper cites.
Clone detection using abstract syntax trees. In Proceedings. International Conference on Software Maintenance (Cat. No. 98CB36272) . IEEE, 368–377
Ira D Baxter, Andrew Yahin, Leonardo Moura, Marcelo Sant’Anna, and Lorraine Bier. 1998 · 1998
Earlier work this paper cites.
Local type inference
Benjamin C Pierce and David N Turner. 2000 · 2000
Earlier work this paper cites.
An exploratory study of how developers seek, relate, and collect relevant information during software maintenance tasks
Amy J Ko, Brad A Myers, Michael J Coblenz, and Htet Htet Aung. 2006 · 2006
Earlier work this paper cites.
Guidelines for performing systematic literature reviews in software engineering
Barbara Kitchenham, Stuart Charters, et al · 2007
Earlier work this paper cites.
What makes APIs hard to learn? Answers from developers
Martin P Robillard. 2009 · 2009
Earlier work this paper cites.
Intelligent selection of language model training data. In Proceedings of the ACL 2010 conference short papers . 220–224
Robert C Moore and William Lewis. 2010 · 2010
Earlier work this paper cites.
From program verification to program synthesis. In Proceedings of the 37th annual ACM SIGPLAN-SIGACT symposium on Principles of programming languages . 313–326
Saurabh Srivastava, Sumit Gulwani, and Jeffrey S Foster. 2010 · 2010
Earlier work this paper cites.
On integrating orthogonal information retrieval methods to improve traceability recovery. In 2011 27th IEEE International Conference on Software Maintenance (ICSM) . IEEE, 133–142
Malcom Gethers, Rocco Oliveto, Denys Poshyvanyk, and Andrea De Lucia. 2011 · 2011
Earlier work this paper cites.
Supporting requirements engineers in recognising security issues. In Requirements Engineering: Foundation for Software Quality: 17th International Working Conference, REFSQ 2011, Essen, Germany, March 28-30, 2011. Proceedings 17 . Springer, 4–18
Eric Knauss, Siv Houmb, Kurt Schneider, Shareeful Islam, and Jan Jürjens. 2011 · 2011
Earlier work this paper cites.
A field study of API learning obstacles
Martin P Robillard and Robert DeLine. 2011 · 2011
Earlier work this paper cites.
Identifying relevant studies in software engineering
He Zhang, Muhammad Ali Babar, and Paolo Tell. 2011 · 2011
Earlier work this paper cites.
How do professional developers comprehend software?. In 2012 34th International Conference on Software Engineering (ICSE) . IEEE, 255–265
Tobias Roehm, Rebecca Tiarks, Rainer Koschke, and Walid Maalej. 2012 · 2012
Earlier work this paper cites.
Syntax-guided synthesis
Rajeev Alur, Rastislav Bodik, Garvit Juniwal, Milo MK Martin, Mukund Raghothaman, Sanjit A Seshia, Rishabh Singh, Armando Solar-Lezama, Emina Torlak, and Abhishek Udupa. 2013 · 2013
Earlier work this paper cites.
Sentiment analysis of commit comments in GitHub: an empirical study. In Proceedings of the 11th working conference on mining software repositories . 352–355
Emitza Guzman, David Azócar, and Yang Li. 2014 · 2014
Earlier work this paper cites.
Non-functional requirements as qualities, with a spice of ontology. In 2014 IEEE 22nd International Requirements Engineering Conference (RE) . IEEE, 293–302
Feng-Lin Li, Jennifer Horkoff, John Mylopoulos, Renata SS Guizzardi, Giancarlo Guizzardi, Alexander Borgida, and Lin Liu. 2014 · 2014
Earlier work this paper cites.
Code completion with statistical language models. In Proceedings of the 35th ACM SIGPLAN conference on programming language design and implementation . 419–428
Veselin Raychev, Martin Vechev, and Eran Yahav. 2014 · 2014
Earlier work this paper cites.
Towards a big data curated benchmark of inter-project code clones. In 2014 IEEE International Conference on Software Maintenance and Evolution . IEEE, 476–480
Jeffrey Svajlenko, Judith F Islam, Iman Keivanloo, Chanchal K Roy, and Mohammad Mamun Mia. 2014 · 2014
Earlier work this paper cites.
Does automated unit test generation really help software testers? a controlled empirical study
Gordon Fraser, Matt Staats, Phil McMinn, Andrea Arcuri, and Frank Padberg. 2015 · 2015
Earlier work this paper cites.
Choosing your weapons: On sentiment analysis tools for software engineering research. In 2015 IEEE international conference on software maintenance and evolution (ICSME) . IEEE, 531–535
Robbert Jongeling, Subhajit Datta, and Alexander Serebrenik. 2015 · 2015
Earlier work this paper cites.
A study of visual studio usage in practice. In 2016 IEEE 23rd International Conference on Software Analysis, Evolution, and Reengineering (SANER) , Vol. 1. IEEE, 124–134
Sven Amann, Sebastian Proksch, Sarah Nadi, and Mira Mezini. 2016 · 2016
Earlier work this paper cites.
Deep API learning. In Proceedings of the 2016 24th ACM SIGSOFT international symposium on foundations of software engineering . 631–642
Xiaodong Gu, Hongyu Zhang, Dongmei Zhang, and Sunghun Kim. 2016 · 2016
Earlier work this paper cites.
API code recommendation using statistical learning from fine-grained changes. In Proceedings of the 2016 24th ACM SIGSOFT International Symposium on Foundations of Software Engineering . 511–522
Anh Tuan Nguyen, Michael Hilton, Mihai Codoban, Hoan Anh Nguyen, Lily Mast, Eli Rademacher, Tien N Nguyen, and Danny Dig. 2016 · 2016
Earlier work this paper cites.
Query expansion based on crowd knowledge for code search
Liming Nie, He Jiang, Zhilei Ren, Zeyi Sun, and Xiaochen Li. 2016 · 2016
Earlier work this paper cites.
Neuro-symbolic program synthesis
Emilio Parisotto, Abdel-rahman Mohamed, Rishabh Singh, Lihong Li, Dengyong Zhou, and Pushmeet Kohli. 2016 · 2016
Earlier work this paper cites.
From query to usable code: an analysis of stack overflow code snippets. In Proceedings of the 13th International Conference on Mining Software Repositories . 391–402
Di Yang, Aftab Hussain, and Cristina Videira Lopes. 2016 · 2016
Earlier work this paper cites.
Case tool support for variability management in software product lines
Rabih Bashroush, Muhammad Garba, Rick Rabiser, Iris Groher, and Goetz Botterweck. 2017 · 2017
Earlier work this paper cites.
Towards synthesizing complex programs from input-output examples
Xinyun Chen, Chang Liu, and Dawn Song. 2017 · 2017
Earlier work this paper cites.
Leveraging automated sentiment analysis in software engineering. In 2017 IEEE/ACM 14th International Conference on Mining Software Repositories (MSR) . IEEE, 203–214
Md Rakibul Islam and Minhaz F Zibran. 2017 · 2017
Earlier work this paper cites.
Static analysis of android apps: A systematic literature review
Li Li, Tegawendé F Bissyandé, Mike Papadakis, Siegfried Rasthofer, Alexandre Bartel, Damien Octeau, Jacques Klein, and Le Traon. 2017 · 2017
Earlier work this paper cites.
Automatic categorization with deep neural network for open-source java projects. In 2017 IEEE/ACM 39th International Conference on Software Engineering Companion (ICSE-C) . IEEE, 164–166
Anh Tuan Nguyen and Tien N Nguyen. 2017 · 2017
Earlier work this paper cites.
Developing safety-critical software: a practical guide for aviation software and DO-178C compliance
Leanna Rierson. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Supervised deep features for software functional clone detection by exploiting lexical and syntactical information in source code.. In IJCAI . 3034–3040
Huihui Wei and Ming Li. 2017 · 2017
Earlier work this paper cites.
A syntactic neural model for general-purpose code generation
Pengcheng Yin and Graham Neubig. 2017 · 2017
Earlier work this paper cites.
An automated approach to estimating code coverage measures via execution logs. In Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering . 305–316
Boyuan Chen, Jian Song, Peng Xu, Xing Hu, and Zhen Ming Jiang. 2018 · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Deep code search. In Proceedings of the 40th International Conference on Software Engineering . 933–944
Xiaodong Gu, Hongyu Zhang, and Sunghun Kim. 2018 · 2018
Earlier work this paper cites.
Deep learning type inference. In Proceedings of the 2018 26th acm joint meeting on european software engineering conference and symposium on the foundations of software engineering . 152–162
Vincent J Hellendoorn, Christian Bird, Earl T Barr, and Miltiadis Allamanis. 2018 · 2018
Earlier work this paper cites.
Deep code comment generation. In Proceedings of the 26th conference on program comprehension . 200–210
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2018 · 2018
Earlier work this paper cites.
API method recommendation without worrying about the task-API knowledge gap. In Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering . 293–304
Qiao Huang, Xin Xia, Zhenchang Xing, David Lo, and Xinyu Wang. 2018 · 2018
Earlier work this paper cites.
Automatic generation of text descriptive comments for code blocks. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 32
Yuding Liang and Kenny Zhu. 2018 · 2018
Earlier work this paper cites.
Effective API recommendation without historical software repositories. In Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering . 282–292
Xiaoyu Liu, LiGuo Huang, and Vincent Ng. 2018 · 2018
Earlier work this paper cites.
Spt-code: Sequence-to-sequence pre-training for learning source code representations. In Proceedings of the 44th International Conference on Software Engineering . 2006–2018
Changan Niu, Chuanyi Li, Vincent Ng, Jidong Ge, Liguo Huang, and Bin Luo. 2022 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
A systematic review of interaction in search-based software engineering
Aurora Ramirez, Jose Raul Romero, and Christopher L Simons. 2018 · 2018
Earlier work this paper cites.
Improving automatic source code summarization via deep reinforcement learning. In Proceedings of the 33rd ACM/IEEE international conference on automated software engineering . 397–407
Yao Wan, Zhou Zhao, Min Yang, Guandong Xu, Haochao Ying, Jian Wu, and Philip S Yu. 2018 · 2018
Earlier work this paper cites.
An experience report of generating load tests using log-recovered workloads at varying granularities of user behaviour. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 669–681
Jinfu Chen, Weiyi Shang, Ahmed E Hassan, Yong Wang, and Jiangbin Lin. 2019a · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for NLP. In International Conference on Machine Learning . PMLR, 2790–2799
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2019 · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Multi-modal attention network learning for semantic source code retrieval. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 13–25
Yao Wan, Jingdong Shu, Yulei Sui, Guandong Xu, Zhou Zhao, Jian Wu, and Philip Yu. 2019 · 2019
Earlier work this paper cites.
A novel neural source code representation based on abstract syntax tree. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 783–794
Jian Zhang, Xu Wang, Hongyu Zhang, Hailong Sun, Kaixuan Wang, and Xudong Liu. 2019 · 2019
Earlier work this paper cites.
Lancer: Your code tell me what you need. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 1202–1205
Shufan Zhou, Beijun Shen, and Hao Zhong. 2019 · 2019
Earlier work this paper cites.
Software documentation: the practitioners’ perspective. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering . 590–601
Emad Aghajani, Csaba Nagy, Mario Linares-Vásquez, Laura Moreno, Gabriele Bavota, Michele Lanza, and David C Shepherd. 2020 · 2020
Earlier work this paper cites.
Achieving reliable sentiment analysis in the software engineering domain using bert. In 2020 IEEE International conference on software maintenance and evolution (ICSME) . IEEE, 162–173
Eeshita Biswas, Mehmet Efruz Karabulut, Lori Pollock, and K Vijay-Shanker. 2020 · 2020
Earlier work this paper cites.
PyMT5: multi-mode translation of natural language and Python code with transformers
Colin B Clement, Dawn Drain, Jonathan Timcheck, Alexey Svyatkovskiy, and Neel Sundaresan. 2020 · 2020
Earlier work this paper cites.
Codebert: A pre-trained model for programming and natural languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Earlier work this paper cites.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al · 2020
Earlier work this paper cites.
Graphcodebert: Pre-training code representations with data flow
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, Shujie Liu, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu, et al · 2020
Earlier work this paper cites.
Kieker: A monitoring framework for software engineering research
Wilhelm Hasselbring and André van Hoorn. 2020 · 2020
Earlier work this paper cites.
Norbert: Transfer learning for requirements classification. In 2020 IEEE 28th International Requirements Engineering Conference (RE) . IEEE, 169–179
Tobias Hey, Jan Keim, Anne Koziolek, and Walter F Tichy. 2020 · 2020
Earlier work this paper cites.
Automated expansion of abbreviations based on semantic relation and transfer expansion
Yanjie Jiang, Hui Liu, Jiahao Jin, and Lu Zhang. 2020 · 2020
Earlier work this paper cites.
Learning and evaluating contextual embedding of source code. In International conference on machine learning . PMLR, 5110–5121
Aditya Kanade, Petros Maniatis, Gogul Balakrishnan, and Kensen Shi. 2020 · 2020
Earlier work this paper cites.
Scelmo: Source code embeddings from language models
Rafael-Michael Karampatsis and Charles Sutton. 2020 · 2020
Earlier work this paper cites.
Multi-task learning based pre-trained language model for code completion. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 473–485
Fang Liu, Ge Li, Yunfei Zhao, and Zhi Jin. 2020 · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Earlier work this paper cites.
Code and named entity recognition in stackoverflow
Jeniya Tabassum, Mounica Maddela, Wei Xu, and Alan Ritter. 2020 · 2020
Earlier work this paper cites.
Evaluating representation learning of code changes for predicting patch correctness in program repair. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 981–992
Haoye Tian, Kui Liu, Abdoul Kader Kaboré, Anil Koyuncu, Li Li, Jacques Klein, and Tegawendé F Bissyandé. 2020 · 2020
Earlier work this paper cites.
Detecting code clones with graph neural network and flow-augmented abstract syntax tree. In 2020 IEEE 27th International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 261–271
Wenhan Wang, Ge Li, Bo Ma, Xin Xia, and Zhi Jin. 2020a · 2020
Earlier work this paper cites.
Modular tree network for source code representation learning
Wenhan Wang, Ge Li, Sijie Shen, Xin Xia, and Zhi Jin. 2020b · 2020
Earlier work this paper cites.
A deep context-wise method for coreference detection in natural language requirements. In 2020 IEEE 28th International Requirements Engineering Conference (RE) . IEEE, 180–191
Yawen Wang, Lin Shi, Mingyang Li, Qing Wang, and Yun Yang. 2020c · 2020
Earlier work this paper cites.
Sentiment analysis for software engineering: How far can pre-trained transformer models go?. In 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 70–80
Ting Zhang, Bowen Xu, Ferdian Thung, Stefanus Agus Haryono, David Lo, and Lingxiao Jiang. 2020b · 2020
Earlier work this paper cites.
Unified pre-training for program understanding and generation
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021 · 2021
Earlier work this paper cites.
GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow
Sid Black, Gao Leo, Phil Wang, Connor Leahy, and Stella Biderman. 2021 · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Latent execution for neural program synthesis beyond domain-specific languages
Xinyun Chen, Dawn Song, and Yuandong Tian. 2021a · 2021
Earlier work this paper cites.
An empirical study on the usage of transformer models for code completion
Matteo Ciniselli, Nathan Cooper, Luca Pascarella, Antonio Mastropaolo, Emad Aghajani, Denys Poshyvanyk, Massimiliano Di Penta, and Gabriele Bavota. 2021 · 2021
Earlier work this paper cites.
Development of recommendation systems for software engineering: the CROSSMINER experience
Juri Di Rocco, Davide Di Ruscio, Claudio Di Sipio, Phuong T Nguyen, and Riccardo Rubei. 2021 · 2021
Earlier work this paper cites.
Augmenting commit classification by using fine-grained source code changes and a pre-trained deep neural language model
Lobna Ghadhab, Ilyes Jenhani, Mohamed Wiem Mkaouer, and Montassar Ben Messaoud. 2021 · 2021
Earlier work this paper cites.
Logging practices with mobile analytics: An empirical study on firebase. In 2021 IEEE/ACM 8th International Conference on Mobile Software Engineering and Systems (MobileSoft) . IEEE, 56–60
Julian Harty, Haonan Zhang, Lili Wei, Luca Pascarella, Mauricio Aniche, and Weiyi Shang. 2021 · 2021
Earlier work this paper cites.
Measuring coding challenge competence with apps
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, et al · 2021
Earlier work this paper cites.
Shipwright: A human-in-the-loop system for dockerfile repair. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 1148–1160
Jordan Henkel, Denini Silva, Leopoldo Teixeira, Marcelo d’Amorim, and Thomas Reps. 2021 · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Earlier work this paper cites.
Duplicate bug report detection by using sentence embedding and fine-tuning. In 2021 IEEE international conference on software maintenance and evolution (ICSME) . IEEE, 535–544
Haruna Isotani, Hironori Washizaki, Yoshiaki Fukazawa, Tsutomu Nomoto, Saori Ouji, and Shinobu Saito. 2021 · 2021
Earlier work this paper cites.
Automatic detection of five api documentation smells: Practitioners’ perspectives. In 2021 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 318–329
Junaed Younus Khan, Md Tawkat Islam Khondaker, Gias Uddin, and Anindya Iqbal. 2021 · 2021
Earlier work this paper cites.
GitHub Copilot AI Is Leaking Functional API Keys
Amit Kulkarni. 2021 · 2021
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Earlier work this paper cites.
Toward less hidden cost of code completion with acceptance and ranking models. In 2021 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 195–205
Jingxuan Li, Rui Huang, Wei Li, Kai Yao, and Weiguo Tan. 2021 · 2021
Earlier work this paper cites.
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang. 2021 · 2021
Earlier work this paper cites.
Traceability transformed: Generating more accurate links with pre-trained bert models. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 324–335
Jinfeng Lin, Yalin Liu, Qingkai Zeng, Meng Jiang, and Jane Cleland-Huang. 2021 · 2021
Earlier work this paper cites.
Codexglue: A machine learning benchmark dataset for code understanding and generation
Shuai Lu, Daya Guo, Shuo Ren, Junjie Huang, Alexey Svyatkovskiy, Ambrosio Blanco, Colin Clement, Dawn Drain, Daxin Jiang, Duyu Tang, et al · 2021
Earlier work this paper cites.
An empirical study on code comment completion. In 2021 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 159–170
Antonio Mastropaolo, Emad Aghajani, Luca Pascarella, and Gabriele Bavota. 2021a · 2021
Earlier work this paper cites.
Studying the usage of text-to-text transfer transformer to support code-related tasks. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 336–347
Antonio Mastropaolo, Simone Scalabrino, Nathan Cooper, David Nader Palacio, Denys Poshyvanyk, Rocco Oliveto, and Gabriele Bavota. 2021b · 2021
Earlier work this paper cites.
Google Brain unveils trillion-parameter AI language model, the largest yet
Sebastian Moss. 2021 · 2021
Earlier work this paper cites.
Examining zero-shot vulnerability repair with large language models
Hammond Pearce, Benjamin Tan, Baleegh Ahmad, Ramesh Karri, and Brendan Dolan-Gavitt. 2021 · 2021
Earlier work this paper cites.
Cotext: Multi-task learning with code-text transformer
Long Phan, Hieu Tran, Daniel Le, Hieu Nguyen, James Anibal, Alec Peltekian, and Yanfang Ye. 2021 · 2021
Earlier work this paper cites.
Making the most of small Software Engineering datasets with modern machine learning
Julian Aron Prenner and Romain Robbes. 2021 · 2021
Earlier work this paper cites.
Dobf: A deobfuscation pre-training objective for programming languages
Baptiste Roziere, Marie-Anne Lachaux, Marc Szafraniec, and Guillaume Lample. 2021 · 2021
Earlier work this paper cites.
GPT-J-6B: A 6 billion parameter autoregressive language model
Ben Wang and Aran Komatsuzaki. 2021 · 2021
Earlier work this paper cites.
Syncobert: Syntax-guided multi-modal contrastive pre-training for code representation
Xin Wang, Yasheng Wang, Fei Mi, Pingyi Zhou, Yao Wan, Xiao Liu, Li Li, Hao Wu, Jin Liu, and Xin Jiang. 2021b · 2021
Earlier work this paper cites.
Yue Wang, Weishi Wang, Shafiq Joty, and Steven CH Hoi. 2021a · 2021
Earlier work this paper cites.
Quality assessment in systematic literature reviews: A software engineering perspective
Lanxin Yang, He Zhang, Haifeng Shen, Xin Huang, Xin Zhou, Guoping Rong, and Dong Shao. 2021 · 2021
Earlier work this paper cites.
On the impact of sample duplication in machine-learning-based android malware detection
Yanjie Zhao, Li Li, Haoyu Wang, Haipeng Cai, Tegawendé F Bissyandé, Jacques Klein, and John Grundy. 2021 · 2021
Earlier work this paper cites.
Evaluation of Context-Aware Language Models and Experts for Effort Estimation of Software Maintenance Issues. In 2022 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 129–138
Mohammed Alhamed and Tim Storer. 2022 · 2022
Earlier work this paper cites.
National vulnerability database
M Anon. 2022 · 2022
Earlier work this paper cites.
Code generation tools (almost) for free? a study of few-shot, pre-trained language models on code
Patrick Bareiß, Beatriz Souza, Marcelo d’Amorim, and Michael Pradel. 2022 · 2022
Earlier work this paper cites.
The Technology Behind BLOOM Training
Stas Bekman. 2022 · 2022
Earlier work this paper cites.
Gpt-neox-20b: An open-source autoregressive language model
Sid Black, Stella Biderman, Eric Hallahan, Quentin Anthony, Leo Gao, Laurence Golding, Horace He, Connor Leahy, Kyle McDonell, Jason Phang, et al · 2022
Earlier work this paper cites.
On the transferability of pre-trained language models for low-resource programming languages. In Proceedings of the 30th IEEE/ACM International Conference on Program Comprehension . 401–412
Fuxiang Chen, Fatemeh H Fard, David Lo, and Timofey Bryksin. 2022 · 2022
Earlier work this paper cites.
Using a Nearest-Neighbour, BERT-Based Approach for Scalable Clone Detection. In 2022 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 582–591
Muslim Chochlov, Gul Aftab Ahmed, James Vincent Patten, Guoxian Lu, Wei Hou, David Gregg, and Jim Buckley. 2022 · 2022
Earlier work this paper cites.
Fast changeset-based bug localization with BERT. In Proceedings of the 44th International Conference on Software Engineering . 946–957
Agnieszka Ciborowska and Kostadin Damevski. 2022 · 2022
Earlier work this paper cites.
Aligning Offline Metrics and Human Judgments of Value of AI-Pair Programmers
Victor Dibia, Adam Fourney, Gagan Bansal, Forough Poursabzi-Sangdeh, Han Liu, and Saleema Amershi. 2022 · 2022
Earlier work this paper cites.
Piloting Copilot and Codex: Hot Temperature, Cold Prompts, or Black Magic?
Jean-Baptiste Döderlein, Mathieu Acher, Djamel Eddine Khelladi, and Benoit Combemale. 2022 · 2022
Earlier work this paper cites.
Automated handling of anaphoric ambiguity in requirements: a multi-solution study. In Proceedings of the 44th International Conference on Software Engineering . 187–199
Saad Ezzini, Sallam Abualhaija, Chetan Arora, and Mehrdad Sabetzadeh. 2022 · 2022
Earlier work this paper cites.
Automated Repair of Programs from Large Language Models
Zhiyu Fan, Xiang Gao, Abhik Roychoudhury, and Shin Hwei Tan. 2022 · 2022
Earlier work this paper cites.
Flakify: A black-box, language model-based predictor for flaky tests
Sakina Fatima, Taher A Ghaleb, and Lionel Briand. 2022 · 2022
Earlier work this paper cites.
Incoder: A generative model for code infilling and synthesis
Daniel Fried, Armen Aghajanyan, Jessy Lin, Sida Wang, Eric Wallace, Freda Shi, Ruiqi Zhong, Wen-tau Yih, Luke Zettlemoyer, and Mike Lewis. 2022 · 2022
Earlier work this paper cites.
GPT2SP: A transformer-based agile story point estimation approach
Michael Fu and Chakkrit Tantithamthavorn. 2022 · 2022
Earlier work this paper cites.
Assemble foundation models for automatic code summarization. In 2022 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 935–946
Jian Gu, Pasquale Salza, and Harald C Gall. 2022 · 2022
Earlier work this paper cites.
PTM4Tag: sharpening tag recommendation of stack overflow posts with pre-trained models. In Proceedings of the 30th IEEE/ACM International Conference on Program Comprehension . 1–11
Junda He, Bowen Xu, Zhou Yang, DongGyun Han, Chengran Yang, and David Lo. 2022 · 2022
Earlier work this paper cites.
Training compute-optimal large language models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, et al · 2022
Earlier work this paper cites.
Perfect is the enemy of test oracle. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 70–81
Ali Reza Ibrahimzada, Yigit Varli, Dilara Tekinoglu, and Reyhaneh Jabbarvand. 2022 · 2022
Earlier work this paper cites.
Codefill: Multi-token code completion by jointly learning from structure and naming sequences. In Proceedings of the 44th International Conference on Software Engineering . 401–412
Maliheh Izadi, Roberta Gismondi, and Georgios Gousios. 2022 · 2022
Earlier work this paper cites.
Jigsaw: Large language models meet program synthesis. In Proceedings of the 44th International Conference on Software Engineering . 1219–1231
Naman Jain, Skanda Vaidyanath, Arun Iyer, Nagarajan Natarajan, Suresh Parthasarathy, Sriram Rajamani, and Rahul Sharma. 2022 · 2022
Earlier work this paper cites.
Learning to predict user-defined types
Kevin Jesse, Premkumar T Devanbu, and Anand Sawant. 2022 · 2022
Earlier work this paper cites.
Capturing failures of large language models via human cognitive biases
Erik Jones and Jacob Steinhardt. 2022 · 2022
Earlier work this paper cites.
Large language models are few-shot testers: Exploring llm-based general bug reproduction
Sungmin Kang, Juyeon Yoon, and Shin Yoo. 2022 · 2022
Earlier work this paper cites.
Automatic detection and analysis of technical debts in peer-review documentation of r packages. In 2022 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 765–776
Junaed Younus Khan and Gias Uddin. 2022 · 2022
Earlier work this paper cites.
SEGRESS: Software engineering guidelines for reporting secondary studies
Barbara Kitchenham, Lech Madeyski, and David Budgen. 2022 · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Earlier work this paper cites.
Less is more: Summary of long instructions is better for program synthesis
Kirby Kuznia, Swaroop Mishra, Mihir Parmar, and Chitta Baral. 2022 · 2022
Earlier work this paper cites.
Interactive code generation via test-driven user-intent formalization
Shuvendu K Lahiri, Aaditya Naik, Georgios Sakkas, Piali Choudhury, Curtis von Veh, Madanlal Musuvathi, Jeevana Priya Inala, Chenglong Wang, and Jianfeng Gao. 2022 · 2022
Earlier work this paper cites.
Towards JavaScript program repair with generative pre-trained transformer (GPT-2). In Proceedings of the Third International Workshop on Automated Program Repair . 61–68
Márk Lajkó, Viktor Csuvik, and László Vidács. 2022 · 2022
Earlier work this paper cites.
Autopruner: transformer-based call graph pruning. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 520–532
Thanh Le-Cong, Hong Jin Kang, Truong Giang Nguyen, Stefanus Agus Haryono, David Lo, Xuan-Bach D Le, and Quyet Thang Huynh. 2022 · 2022
Earlier work this paper cites.
A Light Bug Triage Framework for Applying Large Pre-trained Language Model. In Proceedings of the 37th IEEE/ACM International Conference on Automated Software Engineering . 1–11
Jaehyung Lee, Kisun Han, and Hwanjo Yu. 2022 · 2022
Earlier work this paper cites.
Generation-Augmented Query Expansion For Code Retrieval
Dong Li, Yelong Shen, Ruoming Jin, Yi Mao, Kuan Wang, and Weizhu Chen. 2022d · 2022
Earlier work this paper cites.
CodeRetriever: A Large Scale Contrastive Pre-Training Method for Code Search. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing . 2898–2910
Xiaonan Li, Yeyun Gong, Yelong Shen, Xipeng Qiu, Hang Zhang, Bolun Yao, Weizhen Qi, Daxin Jiang, Weizhu Chen, and Nan Duan. 2022b · 2022
Earlier work this paper cites.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Earlier work this paper cites.
Do Pre-trained Language Models Indeed Understand Software Engineering Tasks?
Yao Li, Tao Zhang, Xiapu Luo, Haipeng Cai, Sen Fang, and Dawei Yuan. 2022g · 2022
Earlier work this paper cites.
CCTEST: Testing and Repairing Code Completion Systems
Zongjie Li, Chaozheng Wang, Zhibo Liu, Haoxuan Wang, Shuai Wang, and Cuiyun Gao. 2022e · 2022
Earlier work this paper cites.
Deep learning for android malware defenses: a systematic literature review
Yue Liu, Chakkrit Tantithamthavorn, Li Li, and Yepang Liu. 2022b · 2022
Earlier work this paper cites.
PRCBERT: Prompt Learning for Requirement Classification using BERT-based Pretrained Language Models. In Proceedings of the 37th IEEE/ACM International Conference on Automated Software Engineering . 1–13
Xianchang Luo, Yinxing Xue, Zhenchang Xing, and Jiamou Sun. 2022 · 2022
Earlier work this paper cites.
Language models of code are few-shot commonsense learners
Aman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang, and Graham Neubig. 2022 · 2022
Earlier work this paper cites.
Using transfer learning for code-related tasks
Antonio Mastropaolo, Nathan Cooper, David Nader Palacio, Simone Scalabrino, Denys Poshyvanyk, Rocco Oliveto, and Gabriele Bavota. 2022a · 2022
Earlier work this paper cites.
Identification of intra-domain ambiguity using transformer-based machine learning. In Proceedings of the 1st International Workshop on Natural Language-based Software Engineering . 51–58
Ambarish Moharil and Arpit Sharma. 2022 · 2022
Earlier work this paper cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022a · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Earlier work this paper cites.
On the effectiveness of transfer learning for code search
Pasquale Salza, Christoph Schwizer, Jian Gu, and Harald C Gall. 2022 · 2022
Earlier work this paper cites.
Bloom: A 176b-parameter open-access multilingual language model
Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, et al · 2022
Earlier work this paper cites.
Talking about large language models
Murray Shanahan. 2022 · 2022
Earlier work this paper cites.
An exploratory study on code attention in BERT. In Proceedings of the 30th IEEE/ACM International Conference on Program Comprehension . 437–448
Rishab Sharma, Fuxiang Chen, Fatemeh Fard, and David Lo. 2022 · 2022
Earlier work this paper cites.
Benchmarking Language Models for Code Syntax Understanding
Da Shen, Xinyun Chen, Chenguang Wang, Koushik Sen, and Dawn Song. 2022 · 2022
Earlier work this paper cites.
Cross-Modal Contrastive Learning for Code Search. In 2022 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 94–105
Zejian Shi, Yun Xiong, Xiaolong Zhang, Yao Zhang, Shanshan Li, and Yangyong Zhu. 2022 · 2022
Earlier work this paper cites.
Selective annotation makes language models better few-shot learners
Hongjin Su, Jungo Kasai, Chen Henry Wu, Weijia Shi, Tianlu Wang, Jiayi Xin, Rui Zhang, Mari Ostendorf, Luke Zettlemoyer, Noah A Smith, et al · 2022
Earlier work this paper cites.
On the importance of building high-quality training datasets for neural code search. In Proceedings of the 44th International Conference on Software Engineering . 1609–1620
Zhensu Sun, Li Li, Yan Liu, Xiaoning Du, and Li Li. 2022 · 2022
Earlier work this paper cites.
Galactica: A large language model for science
Ross Taylor, Marcin Kardas, Guillem Cucurull, Thomas Scialom, Anthony Hartshorn, Elvis Saravia, Andrew Poulton, Viktor Kerkez, and Robert Stojnic. 2022 · 2022
Earlier work this paper cites.
Transformer-based language models for software vulnerability detection. In Proceedings of the 38th Annual Computer Security Applications Conference . 481–496
Chandra Thapa, Seung Ick Jang, Muhammad Ejaz Ahmed, Seyit Camtepe, Josef Pieprzyk, and Surya Nepal. 2022 · 2022
Earlier work this paper cites.
Using pre-trained models to boost code review automation. In Proceedings of the 44th International Conference on Software Engineering . 2291–2302
Rosalia Tufano, Simone Masiero, Antonio Mastropaolo, Luca Pascarella, Denys Poshyvanyk, and Gabriele Bavota. 2022 · 2022
Earlier work this paper cites.
On the validity of pre-trained transformers for natural language processing in the software engineering domain
Julian Von der Mosel, Alexander Trautsch, and Steffen Herbold. 2022 · 2022
Earlier work this paper cites.
You See What I Want You to See: Poisoning Vulnerabilities in Neural Code Search. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (Singapore, Singapore) (ESEC/FSE 2022) . Association for Computing Machinery, New York, NY, USA, 1233–1245
Yao Wan, Shijie Zhang, Hongyu Zhang, Yulei Sui, Guandong Xu, Dezhong Yao, Hai Jin, and Lichao Sun. 2022a · 2022
Earlier work this paper cites.
Machine/deep learning for software engineering: A systematic literature review
Simin Wang, Liguo Huang, Amiao Gao, Jidong Ge, Tengfei Zhang, Haitao Feng, Ishna Satyarth, Ming Li, He Zhang, and Vincent Ng. 2022a · 2022
Earlier work this paper cites.
ReCode: Robustness Evaluation of Code Generation Models
Shiqi Wang, Zheng Li, Haifeng Qian, Chenghao Yang, Zijian Wang, Mingyue Shang, Varun Kumar, Samson Tan, Baishakhi Ray, Parminder Bhatia, et al · 2022
Earlier work this paper cites.
A systematic literature review on the use of deep learning in software engineering research
Cody Watson, Nathan Cooper, David Nader Palacio, Kevin Moran, and Denys Poshyvanyk. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Earlier work this paper cites.
Practical program repair in the era of large pre-trained language models
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2022 · 2022
Earlier work this paper cites.
A systematic evaluation of large language models of code. In Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming . 1–10
Frank F Xu, Uri Alon, Graham Neubig, and Vincent Josua Hellendoorn. 2022 · 2022
Earlier work this paper cites.
Aspect-based api review classification: How far can pre-trained transformer model go?. In 2022 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 385–395
Chengran Yang, Bowen Xu, Junaed Younus Khan, Gias Uddin, Donggyun Han, Zhou Yang, and David Lo. 2022c · 2022
Earlier work this paper cites.
A survey on deep learning for software engineering
Yanming Yang, Xin Xia, David Lo, and John Grundy. 2022b · 2022
Earlier work this paper cites.
CIRCLE: Continual repair across programming languages. In Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis . 678–690
Wei Yuan, Quanjun Zhang, Tieke He, Chunrong Fang, Nguyen Quoc Viet Hung, Xiaodong Hao, and Hongzhi Yin. 2022 · 2022
Earlier work this paper cites.
When language model meets private library
Daoguang Zan, Bei Chen, Zeqi Lin, Bei Guan, Yongji Wang, and Jian-Guang Lou. 2022a · 2022
Earlier work this paper cites.
CERT: Continual Pre-training on Sketches for Library-oriented Code Generation
Daoguang Zan, Bei Chen, Dejian Yang, Zeqi Lin, Minsu Kim, Bei Guan, Yongji Wang, Weizhu Chen, and Jian-Guang Lou. 2022b · 2022
Earlier work this paper cites.
An extensive study on pre-trained models for program understanding and generation. In Proceedings of the 31st ACM SIGSOFT international symposium on software testing and analysis . 39–51
Zhengran Zeng, Hanzhuo Tan, Haotian Zhang, Jing Li, Yuqun Zhang, and Lingming Zhang. 2022 · 2022
Earlier work this paper cites.
BEQAIN: An Effective and Efficient Identifier Normalization Approach With BERT and the Question Answering System
Jingxuan Zhang, Siyuan Liu, Lina Gong, Haoxiang Zhang, Zhiqiu Huang, and He Jiang. 2022a · 2022
Cited alongside, same era.
Automatic chain of thought prompting in large language models
Zhuosheng Zhang, Aston Zhang, Mu Li, and Alex Smola. 2022d · 2022
Cited alongside, same era.
Enhancing Traceability Link Recovery with Unlabeled Data. In 2022 IEEE 33rd International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 446–457
Jianfei Zhu, Guanping Xiao, Zheng Zheng, and Yulei Sui. 2022 · 2022
Cited alongside, same era.
Monitor-Guided Decoding of Code LMs with Static Analysis of Repository Context. In Thirty-seventh Conference on Neural Information Processing Systems
Lakshya Agrawal, Aditya Kanade, Navin Goyal, Shuvendu K Lahiri, and Sriram Rajamani. 2023 · 2023
Cited alongside, same era.
SUT: Active Defects Probing for Transcompiler Models
Mengnan Qi, Yufan Huang, Maoquan Wang, Yongqiang Yao, Zihan Liu, Bin Gu, Colin Clement, and Neel Sundaresan. 2023 · 2023
Closest in time.
Communicative Agents for Software Development
Chen Qian, Xin Cong, Cheng Yang, Weize Chen, Yusheng Su, Juyuan Xu, Zhiyuan Liu, and Maosong Sun. 2023 · 2023
Closest in time.
Vu Le Anh Quan, Chau Thuan Phat, Kiet Van Nguyen, Phan The Duy, and Van-Hau Pham. 2023 · 2023
Closest in time.
Sajjad Rahmani, AmirHossein Naghshzan, and Latifa Guerrouj. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Baleegh Ahmad, Shailja Thakur, Benjamin Tan, Ramesh Karri, and Hammond Pearce. 2023 · 2023
Cited alongside, same era.
Improving Few-Shot Prompts with Relevant Static Analysis Products
Toufique Ahmed, Kunal Suresh Pai, Premkumar Devanbu, and Earl T Barr. 2023 · 2023
Cited alongside, same era.
Extending Source Code Pre-Trained Language Models to Summarise Decompiled Binarie. In 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 260–271
Ali Al-Kaswan, Toufique Ahmed, Maliheh Izadi, Anand Ashok Sawant, Premkumar Devanbu, and Arie van Deursen. 2023 · 2023
Cited alongside, same era.
GPTCloneBench: A comprehensive benchmark of semantic clones and cross-language clones using GPT-3 model and SemanticCloneBench. In 2023 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 1–13
Ajmain I Alam, Palash R Roy, Farouq Al-Omari, Chanchal K Roy, Banani Roy, and Kevin A Schneider. 2023 · 2023
Cited alongside, same era.
Exploring Distributional Shifts in Large Language Models for Code Analysis
Shushan Arakelyan, Rocktim Jyoti Das, Yi Mao, and Xiang Ren. 2023 · 2023
Cited alongside, same era.
ChatGPT is a Remarkable Tool–For Experts
Amos Azaria, Rina Azoulay, and Shulamit Reches. 2023 · 2023
Cited alongside, same era.
Codeplan: Repository-level coding using llms and planning
Ramakrishna Bairi, Atharv Sonwane, Aditya Kanade, Arun Iyer, Suresh Parthasarathy, Sriram Rajamani, B Ashok, Shashank Shet, et al · 2023
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al · 2023
Cited alongside, same era.
Preventing Abuse of LLMs’ Alignment Deficit by Injection Neutralization (PALADIN)
Sami Ramly. 2023 · 2023
Closest in time.
Tricking LLMs into Disobedience: Understanding, Analyzing, and Preventing Jailbreaks
Abhinav Rao, Sachin Vashistha, Atharva Naik, Somak Aditya, and Monojit Choudhury. 2023b · 2023
Closest in time.
Nikitha Rao, Jason Tsay, Kiran Kate, Vincent J Hellendoorn, and Martin Hirzel. 2023a · 2023
Closest in time.
From Misuse to Mastery: Enhancing Code Generation with Knowledge-Driven AI Chaining. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 976–987
Xiaoxue Ren, Xinyuan Ye, Dehai Zhao, Zhenchang Xing, and Xiaohu Yang. 2023 · 2023
Closest in time.
Risks and benefits of large language models for the environment
Matthias C Rillig, Marlene Ågerstrand, Mohan Bi, Kenneth A Gould, and Uli Sauerland. 2023 · 2023
Closest in time.
ChatGPT as a tool for User Story Quality Evaluation: Trustworthy Out of the Box?
Krishna Ronanki, Beatriz Cabrero-Daniel, and Christian Berger. 2023 · 2023
Closest in time.
Multilingual Adapter-based Knowledge Aggregation on Code Summarization for Low-Resource Languages
Iman Saberi, Fatemeh Fard, and Fuxiang Chen. 2023a · 2023
Closest in time.
Iman Saberi, Fatemeh Fard, and Fuxiang Chen. 2023b · 2023
Closest in time.
Analysis of ChatGPT on Source Code
Ahmed Sadik, Antonello Ceravola, Frank Joublin, and Jibesh Patra. 2023 · 2023
Closest in time.
On Contrastive Learning of Semantic Similarity forCode to Code Search
Anthony Saieva, Saikat Chakraborty, and Gail Kaiser. 2023 · 2023
Closest in time.
Extending the Frontier of ChatGPT: Code Generation and Debugging
Fardin Ahsan Sakib, Saadat Hasan Khan, and AHM Karim. 2023 · 2023
Closest in time.
Adaptive test generation using a large language model
Max Schäfer, Sarah Nadi, Aryaz Eghbali, and Frank Tip. 2023a · 2023
Closest in time.
An empirical evaluation of using large language models for automated unit test generation
Max Schäfer, Sarah Nadi, Aryaz Eghbali, and Frank Tip. 2023b · 2023
Closest in time.
Imanol Schlag, Sainbayar Sukhbaatar, Asli Celikyilmaz, Wen tau Yih, Jason Weston, Jürgen Schmidhuber, and Xian Li. 2023 · 2023
Closest in time.
AutoScrum: Automating Project Planning Using Large Language Models
Martin Schroder. 2023 · 2023
Closest in time.
A Multi-Step Learning Approach to Assist Code Review. In 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 450–460
Oussama Ben Sghaier and Houari Sahraoui. 2023 · 2023
Closest in time.
Entity-augmented code generation
Anton Shapkin, Denis Litvinov, and Timofey Bryksin. 2023 · 2023
Closest in time.
FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU
Ying Sheng, Lianmin Zheng, Binhang Yuan, Zhuohan Li, Max Ryabinin, Beidi Chen, Percy Liang, Christopher Re, Ion Stoica, and Ce Zhang. 2023 · 2023
Closest in time.
Towards Efficient Fine-tuning of Pre-trained Code Models: An Experimental Study and Beyond
Ensheng Shi, Yanlin Wang, Hongyu Zhang, Lun Du, Shi Han, Dongmei Zhang, and Hongbin Sun. 2023a · 2023
Closest in time.
SoTaNa: The Open-Source Software Development Assistant
Ensheng Shi, Fengji Zhang, Yanlin Wang, Bei Chen, Lun Du, Hongyu Zhang, Shi Han, Dongmei Zhang, and Hongbin Sun. 2023c · 2023
Closest in time.
Domain Adaptation for Deep Unit Test Case Generation
Jiho Shin, Sepehr Hashtroudi, Hadi Hemmati, and Song Wang. 2023a · 2023
Closest in time.
Jiho Shin, Clark Tang, Tahmineh Mohati, Maleknaz Nayebi, Song Wang, and Hadi Hemmati. 2023b · 2023
Closest in time.
Exploring the Robustness of Large Language Models for Solving Programming Problems
Atsushi Shirafuji, Yutaka Watanobe, Takumi Ito, Makoto Morishita, Yuki Nakamura, Yusuke Oda, and Jun Suzuki. 2023 · 2023
Closest in time.
Learning performance-improving code edits
Alexander Shypula, Aman Madaan, Yimeng Zeng, Uri Alon, Jacob Gardner, Milad Hashemi, Graham Neubig, Parthasarathy Ranganathan, Osbert Bastani, and Amir Yazdanbakhsh. 2023 · 2023
Closest in time.
A Lightweight Framework for High-Quality Code Generation
Mohammed Latif Siddiq, Beatrice Casey, and Joanna Santos. 2023a · 2023
Closest in time.
Exploring the Effectiveness of Large Language Models in Generating Unit Tests
Mohammed Latif Siddiq, Joanna Santos, Ridwanul Hasan Tanvir, Noshin Ulfat, Fahmid Al Rifat, and Vinicius Carvalho Lopes. 2023b · 2023
Closest in time.
RepairLLaMA: Efficient Representations and Fine-Tuned Adapters for Program Repair
André Silva, Sen Fang, and Martin Monperrus. 2023 · 2023
Closest in time.
Evaluating ChatGPT and GPT-4 for Visual Programming
Adish Singla. 2023 · 2023
Closest in time.
An analysis of the automatic bug fixing performance of chatgpt
Dominik Sobania, Martin Briesch, Carol Hanna, and Justyna Petke. 2023 · 2023
Closest in time.
ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks
Giriprasad Sridhara, Sourav Mazumdar, et al · 2023
Closest in time.
Reinforcement Learning from Automatic Feedback for High-Quality Unit Test Generation
Benjamin Steenhoek, Michele Tufano, Neel Sundaresan, and Alexey Svyatkovskiy. 2023 · 2023
Closest in time.
Clover: Closed-Loop Verifiable Code Generation
Chuyue Sun, Ying Sheng, Oded Padon, and Clark Barrett. 2023d · 2023
Closest in time.
Silent Vulnerable Dependency Alert Prediction with Vulnerability Key Aspect Explanation
Jiamou Sun, Zhenchang Xing, Qinghua Lu, Xiwei Xu, Liming Zhu, Thong Hoang, and Dehai Zhao. 2023f · 2023
Closest in time.
Dexbert: effective, task-agnostic and fine-grained representation learning of Android bytecode
Tiezhu Sun, Kevin Allix, Kisub Kim, Xin Zhou, Dongsun Kim, David Lo, Tegawendé F Bissyandé, and Jacques Klein. 2023a · 2023
Closest in time.
A Prompt Learning Framework for Source Code Summarization
Weisong Sun, Chunrong Fang, Yudu You, Yuchen Chen, Yi Liu, Chong Wang, Jian Zhang, Quanjun Zhang, Hanwei Qian, Wei Zhao, et al · 2023
Closest in time.
Automatic Code Summarization via ChatGPT: How Far Are We?
Weisong Sun, Chunrong Fang, Yudu You, Yun Miao, Yi Liu, Yuekang Li, Gelei Deng, Shenghan Huang, Yuchen Chen, Quanjun Zhang, et al · 2023
Closest in time.
Yuqiang Sun, Daoyuan Wu, Yue Xue, Han Liu, Haijun Wang, Zhengzi Xu, Xiaofei Xie, and Yang Liu. 2023e · 2023
Closest in time.
Copilot for Xcode: Exploring AI-Assisted Programming by Prompting Cloud-based Large Language Models
Chee Wei Tan, Shangxin Guo, Man Fai Wong, and Ching Nam Hang. 2023 · 2023
Closest in time.
CSGVD: A deep learning approach combining sequence and graph embedding for source code vulnerability detection
Wei Tang, Mingwei Tang, Minchao Ban, Ziguo Zhao, and Mingjun Feng. 2023d · 2023
Closest in time.
Just-in-Time Security Patch Detection–LLM At the Rescue for Data Augmentation
Xunzhu Tang, Zhenghan Chen, Kisub Kim, Haoye Tian, Saad Ezzini, and Jacques Klein. 2023a · 2023
Closest in time.
ChatGPT vs SBST: A Comparative Assessment of Unit Test Suite Generation
Yutian Tang, Zhijie Liu, Zhichao Zhou, and Xiapu Luo. 2023c · 2023
Closest in time.
Domain Adaptive Code Completion via Language Models and Decoupled Domain Databases. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 421–433
Ze Tang, Jidong Ge, Shangqing Liu, Tingwei Zhu, Tongtong Xu, Liguo Huang, and Bin Luo. 2023b · 2023
Closest in time.
The potential of LLMs for coding with low-resource and domain-specific programming languages
Artur Tarassow. 2023 · 2023
Closest in time.
VeriGen: A Large Language Model for Verilog Code Generation
Shailja Thakur, Baleegh Ahmad, Hammond Pearce, Benjamin Tan, Brendan Dolan-Gavitt, Ramesh Karri, and Siddharth Garg. 2023 · 2023
Closest in time.
The Best of Both Worlds: Combining Learned Embeddings with Engineered Features for Accurate Prediction of Correct Patches
Haoye Tian, Kui Liu, Yinghua Li, Abdoul Kader Kaboré, Anil Koyuncu, Andrew Habib, Li Li, Junhao Wen, Jacques Klein, and Tegawendé F Bissyandé. 2023a · 2023
Closest in time.
Is ChatGPT the Ultimate Programming Assistant–How far is it?
Haoye Tian, Weiqi Lu, Tsz On Li, Xunzhu Tang, Shing-Chi Cheung, Jacques Klein, and Tegawendé F Bissyandé. 2023b · 2023
Closest in time.
Test-case-driven programming understanding in large language models for better code generation
Zhao Tian and Junjie Chen. 2023 · 2023
Closest in time.
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification
Norbert Tihanyi, Tamas Bisztray, Ridhi Jain, Mohamed Amine Ferrag, Lucas C Cordeiro, and Vasileios Mavroeidis. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Closest in time.
LLM4CBI: Taming LLMs to Generate Effective Test Programs for Compiler Bug Isolation
Haoxin Tu, Zhide Zhou, He Jiang, Imam Nur Bani Yusuf, Yuxian Li, and Lingxiao Jiang. 2023 · 2023
Closest in time.
Predicting Code Coverage without Execution
Michele Tufano, Shubham Chandel, Anisha Agarwal, Neel Sundaresan, and Colin Clement. 2023 · 2023
Closest in time.
Can Large Language Models Write Good Property-Based Tests?
Vasudev Vikram, Caroline Lemieux, and Rohan Padhye. 2023 · 2023
Closest in time.
Frustrated with code quality issues? llms can help!
Nalin Wadhwa, Jui Pradhan, Atharv Sonwane, Surya Prakash Sahu, Nagarajan Natarajan, Aditya Kanade, Suresh Parthasarathy, and Sriram Rajamani. 2023 · 2023
Closest in time.
Boosting Static Resource Leak Detection via LLM-based Resource-Oriented Intention Inference
Chong Wang, Jianan Liu, Xin Peng, Yang Liu, and Yiling Lou. 2023g · 2023
Closest in time.
One Adapter for All Programming Languages? Adapter Tuning for Code Search and Summarization
Deze Wang, Boxing Chen, Shanshan Li, Wei Luo, Shaoliang Peng, Wei Dong, and Xiangke Liao. 2023a · 2023
Closest in time.
Software Testing with Large Language Model: Survey, Landscape, and Vision
Junjie Wang, Yuchao Huang, Chunyang Chen, Zhe Liu, Song Wang, and Qing Wang. 2023c · 2023
Closest in time.
Evaluating AIGC Detectors on Code Content
Jian Wang, Shangqing Liu, Xiaofei Xie, and Yi Li. 2023h · 2023
Closest in time.
Shufan Wang, Sebastien Jean, Sailik Sengupta, James Gung, Nikolaos Pappas, and Yi Zhang. 2023d · 2023
Closest in time.
LeTI: Learning to Generate from Textual Interactions
Xingyao Wang, Hao Peng, Reyhaneh Jabbarvand, and Heng Ji. 2023i · 2023
Closest in time.
Codet5+: Open code large language models for code understanding and generation
Yue Wang, Hung Le, Akhilesh Deepak Gotmare, Nghi DQ Bui, Junnan Li, and Steven CH Hoi. 2023e · 2023
Closest in time.
ChatCoder: Chat-based Refine Requirement Improves LLMs’ Code Generation
Zejun Wang, Jia Li, Ge Li, and Zhi Jin. 2023f · 2023
Closest in time.
Magicoder: Source code is all you need
Yuxiang Wei, Zhe Wang, Jiawei Liu, Yifeng Ding, and Lingming Zhang. 2023a · 2023
Closest in time.
Exploring parameter-efficient fine-tuning techniques for code generation with large language models
Martin Weyssow, Xin Zhou, Kisub Kim, David Lo, and Houari Sahraoui. 2023a · 2023
Closest in time.
Martin Weyssow, Xin Zhou, Kisub Kim, David Lo, and Houari Sahraoui. 2023b · 2023
Closest in time.
A prompt pattern catalog to enhance prompt engineering with chatgpt
Jules White, Quchen Fu, Sam Hays, Michael Sandborn, Carlos Olea, Henry Gilbert, Ashraf Elnashar, Jesse Spencer-Smith, and Douglas C Schmidt. 2023a · 2023
Closest in time.
Jules White, Sam Hays, Quchen Fu, Jesse Spencer-Smith, and Douglas C Schmidt. 2023b · 2023
Closest in time.
Addressing Compiler Errors: Stack Overflow or Large Language Models?
Patricia Widjojo and Christoph Treude. 2023 · 2023
Closest in time.
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
Ratnadira Widyasari, Ting Zhang, Abir Bouraffa, and David Lo. 2023 · 2023
Closest in time.
Natural Language Generation and Understanding of Big Code for AI-Assisted Programming: A Review
Man-Fai Wong, Shangxin Guo, Ching-Nam Hang, Siu-Wai Ho, and Chee-Wei Tan. 2023 · 2023
Closest in time.
Deceptprompt: Exploiting llm-driven code generation via adversarial natural language instructions
Fangzhou Wu, Xiaogeng Liu, and Chaowei Xiao. 2023d · 2023
Closest in time.
Defending ChatGPT against Jailbreak Attack via Self-Reminder
Fangzhao Wu, Yueqi Xie, Jingwei Yi, Jiawei Shao, Justin Curl, Lingjuan Lyu, Qifeng Chen, and Xing Xie. 2023e · 2023
Closest in time.
Is AI the better programming partner? Human-Human Pair Programming vs. Human-AI pAIr Programming
Tongshuang Wu, Kenneth Koedinger, et al · 2023
Closest in time.
How Effective Are Neural Networks for Fixing Security Vulnerabilities
Yi Wu, Nan Jiang, Hung Viet Pham, Thibaud Lutellier, Jordan Davis, Lin Tan, Petr Babkin, and Sameena Shah. 2023a · 2023
Closest in time.
Large language models in fault localisation
Yonghao Wu, Zheng Li, Jie M Zhang, Mike Papadakis, Mark Harman, and Yong Liu. 2023c · 2023
Closest in time.
Automated Program Repair in the Era of Large Pre-Trained Language Models. In Proceedings of the 45th International Conference on Software Engineering (ICSE ’23)
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Closest in time.
Conversational automated program repair
Chunqiu Steven Xia and Lingming Zhang. 2023a · 2023
Closest in time.
Keep the Conversation Going: Fixing 162 out of 337 bugs for 0.42 0.42 each using ChatGPT
Chunqiu Steven Xia and Lingming Zhang. 2023b · 2023
Closest in time.
Impact of Large Language Models on Generating Software Specifications
Danning Xie, Byungwoo Yoo, Nan Jiang, Mijung Kim, Lin Tan, Xiangyu Zhang, and Judy S Lee. 2023b · 2023
Closest in time.
ChatUniTest: a ChatGPT-based automated unit test generation tool
Zhuokui Xie, Yinghao Chen, Chen Zhi, Shuiguang Deng, and Jianwei Yin. 2023a · 2023
Closest in time.
The Program Testing Ability of Large Language Models for Code
Weimin Xiong, Yiwen Guo, and Hao Chen. 2023 · 2023
Closest in time.
LmPa: Improving Decompilation by Synergy of Large Language Model and Program Analysis
Xiangzhe Xu, Zhuo Zhang, Shiwei Feng, Yapeng Ye, Zian Su, Nan Jiang, Siyuan Cheng, Lin Tan, and Xiangyu Zhang. 2023b · 2023
Closest in time.
Guiding ChatGPT to Fix Web UI Tests via Explanation-Consistency Checking
Zhuolin Xu, Yuanzhang Lin, Qiushi Li, and Shin Hwei Tan. 2023a · 2023
Closest in time.
A Closer Look at Different Difficulty Levels Code Generation Abilities of ChatGPT. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 1887–1898
Dapeng Yan, Zhipeng Gao, and Zhiming Liu. 2023a · 2023
Closest in time.
Codetransocean: A comprehensive multilingual benchmark for code translation
Weixiang Yan, Yuchen Tian, Yunzhe Li, Qian Chen, and Wen Wang. 2023b · 2023
Closest in time.
White-box compiler fuzzing empowered by large language models
Chenyuan Yang, Yinlin Deng, Runyu Lu, Jiayi Yao, Jiawei Liu, Reyhaneh Jabbarvand, and Lingming Zhang. 2023a · 2023
Closest in time.
Chengran Yang, Jiakun Liu, Bowen Xu, Christoph Treude, Yunbo Lyu, Ming Li, and David Lo. 2023c · 2023
Closest in time.
A Syntax-Guided Multi-Task Learning Approach for Turducken-Style Code Generation
Guang Yang, Yu Zhou, Xiang Chen, Xiangyu Zhang, Yiran Xu, Tingting Han, and Taolue Chen. 2023f · 2023
Closest in time.
Assessing and Improving Syntactic Adversarial Robustness of Pre-trained Models for Code Translation
Guang Yang, Yu Zhou, Xiangyu Zhang, Xiang Chen, Tingting Han, and Taolue Chen. 2023g · 2023
Closest in time.
Harnessing the power of llms in practice: A survey on chatgpt and beyond
Jingfeng Yang, Hongye Jin, Ruixiang Tang, Xiaotian Han, Qizhang Feng, Haoming Jiang, Bing Yin, and Xia Hu. 2023b · 2023
Closest in time.
Enhancing Code Intelligence Tasks with ChatGPT
Kang Yang, Xinjun Mao, Shangwen Wang, Tanghaoran Zhang, Bo Lin, Yanlin Wang, Yihao Qin, Zhang Zhang, and Xiaoguang Mao. 2023d · 2023
Closest in time.
Generating Data for Symbolic Language with Large Language Models
Jiacheng Ye, Chengzu Li, Lingpeng Kong, and Tao Yu. 2023 · 2023
Closest in time.
CoLadder: Supporting Programmers with Hierarchical Code Generation in Multi-Level Abstraction
Ryan Yen, Jiawen Zhu, Sangho Suh, Haijun Xia, and Jian Zhao. 2023 · 2023
Closest in time.
Burak Yetiştiren, Işık Özsoy, Miray Ayerdem, and Eray Tüzün. 2023 · 2023
Closest in time.
Chinese LLaMA & Alpaca Large Language Models
ymcui. 2023 · 2023
Closest in time.
Autonomous Large Language Model Agents Enabling Intent-Driven Mobile GUI Testing
Juyeon Yoon, Robert Feldt, and Shin Yoo. 2023 · 2023
Closest in time.
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models
Hao Yu, Bo Shen, Dezhi Ran, Jiaxin Zhang, Qi Zhang, Yuchi Ma, Guangtai Liang, Ying Li, Tao Xie, and Qianxiang Wang. 2023a · 2023
Closest in time.
No More Manual Tests? Evaluating and Improving ChatGPT for Unit Test Generation
Zhiqiang Yuan, Yiling Lou, Mingwei Liu, Shiji Ding, Kaixin Wang, Yixuan Chen, and Xin Peng. 2023b · 2023
Closest in time.
Private-library-oriented code generation with large language models
Daoguang Zan, Bei Chen, Yongshun Gong, Junzhi Cao, Fengji Zhang, Bingchao Wu, Bei Guan, Yilong Yin, and Yongji Wang. 2023a · 2023
Closest in time.
Self-taught optimizer (stop): Recursively self-improving code generation
Eric Zelikman, Eliana Lorch, Lester Mackey, and Adam Tauman Kalai. 2023 · 2023
Closest in time.
Understanding Large Language Model Based Fuzz Driver Generation
Cen Zhang, Mingqiang Bai, Yaowen Zheng, Yeting Li, Xiaofei Xie, Yuekang Li, Wei Ma, Limin Sun, and Yang Liu. 2023a · 2023
Closest in time.
Prompt-enhanced software vulnerability detection using chatgpt
Chenyuan Zhang, Hao Liu, Jiutian Zeng, Kejing Yang, Yuhong Li, and Hui Li. 2023m · 2023
Closest in time.
Multilingual Code Co-Evolution Using Large Language Models
Jiyang Zhang, Pengyu Nie, Junyi Jessy Li, and Milos Gligoric. 2023n · 2023
Closest in time.
ToolCoder: Teach Code Generation Models to use APIs with search tools
Kechi Zhang, Ge Li, Jia Li, Zhuo Li, and Zhi Jin. 2023k · 2023
Closest in time.
Self-Edit: Fault-Aware Code Editor for Code Generation
Kechi Zhang, Zhuo Li, Jia Li, Ge Li, and Zhi Jin. 2023l · 2023
Closest in time.
ALGO: Synthesizing Algorithmic Programs with Generated Oracle Verifiers
Kexun Zhang, Danqing Wang, Jingtao Xia, William Yang Wang, and Lei Li. 2023o · 2023
Closest in time.
A Survey of Learning-based Automated Program Repair
Quanjun Zhang, Chunrong Fang, Yuxiang Ma, Weisong Sun, and Zhenyu Chen. 2023b · 2023
Closest in time.
Boosting Automated Patch Correctness Prediction via Pre-trained Language Model
Quanjun Zhang, Chunrong Fang, Weisong Sun, Yan Liu, Tieke He, Xiaodong Hao, and Zhenyu Chen. 2023c · 2023
Closest in time.
Gamma: Revisiting template-based automated program repair via mask prediction. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 535–547
Quanjun Zhang, Chunrong Fang, Tongke Zhang, Bowen Yu, Weisong Sun, and Zhenyu Chen. 2023d · 2023
Closest in time.
Duplicate bug report detection: How far are we?
Ting Zhang, DongGyun Han, Venkatesh Vinayakarao, Ivana Clairine Irsan, Bowen Xu, Ferdian Thung, David Lo, and Lingxiao Jiang. 2023e · 2023
Closest in time.
Cupid: Leveraging chatgpt for more accurate duplicate bug report detection
Ting Zhang, Ivana Clairine Irsan, Ferdian Thung, and David Lo. 2023f · 2023
Closest in time.
Revisiting sentiment analysis for software engineering in the era of large language models
Ting Zhang, Ivana Clairine Irsan, Ferdian Thung, and David Lo. 2023g · 2023
Closest in time.
Evaluating Pre-trained Language Models for Repairing API Misuses
Ting Zhang, Ivana Clairine Irsan, Ferdian Thung, David Lo, Asankhaya Sharma, and Lingxiao Jiang. 2023h · 2023
Closest in time.
STEAM: simulating the interactive behavior of programmers for automatic bug fixing
Yuwei Zhang, Zhi Jin, Ying Xing, and Ge Li. 2023i · 2023
Closest in time.
Neural Program Repair with Program Dependence Analysis and Effective Filter Mechanism
Yuwei Zhang, Ge Li, Zhi Jin, and Ying Xing. 2023j · 2023
Closest in time.
Understanding Programs by Exploiting (Fuzzing) Test Cases
Jianyu Zhao, Yuyang Rong, Yiwen Guo, Yifeng He, and Hao Chen. 2023a · 2023
Closest in time.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Closest in time.
Automatic Model Selection with Large Language Models for Reasoning
Xu Zhao, Yuxi Xie, Kenji Kawaguchi, Junxian He, and Qizhe Xie. 2023b · 2023
Closest in time.
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
Zelin Zhao, Zhaogui Xu, Jialong Zhu, Peng Di, Yuan Yao, and Xiaoxing Ma. 2023c · 2023
Closest in time.
Codegeex: A pre-trained model for code generation with multilingual evaluations on humaneval-x
Qinkai Zheng, Xiao Xia, Xu Zou, Yuxiao Dong, Shan Wang, Yufei Xue, Zihan Wang, Lei Shen, Andi Wang, Yang Li, et al · 2023
Closest in time.
Outline, then details: Syntactically guided coarse-to-fine code generation
Wenqing Zheng, SP Sharan, Ajay Kumar Jaiswal, Kevin Wang, Yihan Xi, Dejia Xu, and Zhangyang Wang. 2023b · 2023
Closest in time.
A survey of large language models for code: Evolution, benchmarking, and future trends
Zibin Zheng, Kaiwen Ning, Yanlin Wang, Jingwen Zhang, Dewu Zheng, Mingxi Ye, and Jiachi Chen. 2023a · 2023
Closest in time.
A study on robustness and reliability of large language model code generation
Li Zhong and Zilong Wang. 2023 · 2023
Closest in time.
Codebertscore: Evaluating code generation with pretrained models of code
Shuyan Zhou, Uri Alon, Sumit Agarwal, and Graham Neubig. 2023a · 2023
Closest in time.
UniversalNER: Targeted Distillation from Large Language Models for Open Named Entity Recognition
Wenxuan Zhou, Sheng Zhang, Yu Gu, Muhao Chen, and Hoifung Poon. 2023c · 2023
Closest in time.
Large Language Models Are Human-Level Prompt Engineers
Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba. 2023b · 2023
Closest in time.
Automating Method Naming with Context-Aware Prompt-Tuning
Jie Zhu, Lingwei Li, Li Yang, Xiaoxiao Ma, and Chun Zuo. 2023 · 2023
Closest in time.
Large Language Models Are State-of-the-Art Evaluators of Code Generation
Terry Yue Zhuo. 2023 · 2023
Closest in time.
Pop Quiz! Do Pre-trained Code Models Possess Knowledge of Correct API Names?
Terry Yue Zhuo, Xiaoning Du, Zhenchang Xing, Jiamou Sun, Haowei Quan, Li Li, and Liming Zhu. 2023 · 2023
Closest in time.
Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models
Mayank Agarwal, Yikang Shen, Bailin Wang, Yoon Kim, and Jie Chen. 2024 · 2024
Closest in time.
Automatic Semantic Augmentation of Language Model Prompts (for Code Summarization)
Toufique Ahmed, Kunal Suresh Pai, Premkumar Devanbu, and Earl T. Barr. 2024 · 2024
Closest in time.
APIGen: Generative API Method Recommendation
Yujia Chen, Cuiyun Gao, Muyijie Zhu, Qing Liao, Yong Wang, and Guoai Xu. 2024 · 2024
Closest in time.
PyTy: Repairing Static Type Errors in Python
Yiu Wai Chow, Luca Di Grazia, and Michael Pradel. 2024 · 2024
Closest in time.
Large language models of code fail at completing code with potential bugs
Tuan Dinh, Jinman Zhao, Samson Tan, Renato Negrinho, Leonard Lausen, Sheng Zha, and George Karypis. 2024 · 2024
Closest in time.
De-Hallucinator: Iterative Grounding for LLM-Based Code Completion
Aryaz Eghbali and Michael Pradel. 2024 · 2024
Closest in time.
Rapid: Zero-shot Domain Adaptation for Code Search with Pre-trained Models
Guodong Fan, Shizhan Chen, Cuiyun Gao, Jianmao Xiao, Tao Zhang, and Zhiyong Feng. 2024 · 2024
Closest in time.
Incivility detection in open source code review and issue discussions
Isabella Ferreira, Ahlaam Rafiq, and Jinghui Cheng. 2024 · 2024
Closest in time.
Shuzheng Gao, Wenxin Mao, Cuiyun Gao, Li Li, Xing Hu, Xin Xia, and Michael R Lyu. 2024 · 2024
Closest in time.
Large Language Models are Few-Shot Summarizers: Multi-Intent Comment Generation via In-Context Learning
Mingyang Geng, Shangwen Wang, Dezun Dong, Haotian Wang, Ge Li, Zhi Jin, Xiaoguang Mao, and Xiangke Liao. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Y Wu, YK Li, et al · 2024
Closest in time.
Leveraging Print Debugging to Improve Code Generation in Large Language Models
Xueyu Hu, Kun Kuang, Jiankai Sun, Hongxia Yang, and Fei Wu. 2024 · 2024
Closest in time.
Crashtranslator: Automatically reproducing mobile application crashes directly from stack trace. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Yuchao Huang, Junjie Wang, Zhe Liu, Yawen Wang, Song Wang, Chunyang Chen, Yuanzhe Hu, and Qing Wang. 2024 · 2024
Closest in time.
LLM-Powered Code Vulnerability Repair with Reinforcement Learning and Semantic Reward
Nafis Tanveer Islam, Joseph Khoury, Andrew Seong, Gonzalo De La Torre Parra, Elias Bou-Harb, and Peyman Najafirad. 2024 · 2024
Closest in time.
Code Security Vulnerability Repair Using Reinforcement Learning with Large Language Models
Nafis Tanveer Islam and Peyman Najafirad. 2024 · 2024
Closest in time.
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using ChatGPT
Azmain Kabir, Shaowei Wang, Yuan Tian, Muhammad Asaduzzaman, Wenbin Zhang, et al · 2024
Closest in time.
ChatGPT and Human Synergy in Black-Box Testing: A Comparative Analysis
Hiroyuki Kirinuki and Haruto Tanno. 2024 · 2024
Closest in time.
Rewriting the Code: A Simple Method for Large Language Model Augmented Code Search
Haochen Li, Xin Zhou, and Zhiqi Shen. 2024c · 2024
Closest in time.
DevEval: Evaluating Code Generation in Practical Software Projects
Jia Li, Ge Li, Yunfei Zhao, Yongmin Li, Zhi Jin, Hao Zhu, Huanyu Liu, Kaibo Liu, Lecheng Wang, Zheng Fang, et al · 2024
Closest in time.
On the Reliability and Explainability of Language Models for Program Generation
Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, and Li Li. 2024a · 2024
Closest in time.
SpecGen: Automated Generation of Formal Program Specifications via Large Language Models
Lezhi Ma, Shangqing Liu, Yi Li, Xiaofei Xie, and Lei Bu. 2024a · 2024
Closest in time.
T-FREX: A Transformer-based Feature Extraction Method from Mobile App Reviews
Quim Motger, Alessio Miaschi, Felice Dell’Orletta, Xavier Franch, and Jordi Marco. 2024 · 2024
Closest in time.
Model driven engineering for machine learning components: A systematic literature review
Hira Naveed, Chetan Arora, Hourieh Khalajzadeh, John Grundy, and Omar Haggag. 2024 · 2024
Closest in time.
Can Large Language Models Write Parallel Code?
Daniel Nichols, Joshua H Davis, Zhaojun Xie, Arjun Rajaram, and Abhinav Bhatele. 2024 · 2024
Closest in time.
Domain knowledge matters: Improving prompts with fix templates for repairing python type errors. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Yun Peng, Shuzheng Gao, Cuiyun Gao, Yintong Huo, and Michael Lyu. 2024 · 2024
Closest in time.
Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering
Tal Ridnik, Dedy Kredo, and Itamar Friedman. 2024 · 2024
Closest in time.
Fernando Vallecillos Ruiz, Anastasiia Grishina, Max Hort, and Leon Moonen. 2024 · 2024
Closest in time.
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
Pranab Sahoo, Ayush Kumar Singh, Sriparna Saha, Vinija Jain, Samrat Mondal, and Aman Chadha. 2024 · 2024
Closest in time.
Finetuning Large Language Models for Vulnerability Detection
Alexey Shestov, Anton Cheshkov, Rodion Levichev, Ravil Mussabayev, Pavel Zadorozhny, Evgeny Maslov, Chibirev Vadim, and Egor Bulychev. 2024 · 2024
Closest in time.
Dataflow Analysis-Inspired Deep Learning for Efficient Vulnerability Detection. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Benjamin Steenhoek, Hongyang Gao, and Wei Le. 2024 · 2024
Closest in time.
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs’ Vulnerability Reasoning
Yuqiang Sun, Daoyuan Wu, Yue Xue, Han Liu, Wei Ma, Lyuye Zhang, Miaolei Shi, and Yang Liu. 2024b · 2024
Closest in time.
Zhensu Sun, Xiaoning Du, Fu Song, Shangwen Wang, and Li Li. 2024a · 2024
Closest in time.
Debugbench: Evaluating debugging capability of large language models
Runchu Tian, Yining Ye, Yujia Qin, Xin Cong, Yankai Lin, Zhiyuan Liu, and Maosong Sun. 2024 · 2024
Closest in time.
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
Chong Wang, Jian Zhang, Yebo Feng, Tianlin Li, Weisong Sun, Yang Liu, and Xin Peng. 2024c · 2024
Closest in time.
OOP: Object-Oriented Programming Evaluation Benchmark for Large Language Models
Shuai Wang, Liang Ding, Li Shen, Yong Luo, Bo Du, and Dacheng Tao. 2024a · 2024
Closest in time.
SparseCoder: Identifier-Aware Sparse Transformer for File-Level Code Summarization
Yanlin Wang, Yanxian Huang, Daya Guo, Hongyu Zhang, and Zibin Zheng. 2024b · 2024
Closest in time.
Automatic recognizing relevant fragments of APIs using API references
Di Wu, Yang Feng, Hongyu Zhang, and Baowen Xu. 2024 · 2024
Closest in time.
Fuzz4all: Universal fuzzing with large language models
Chunqiu Steven Xia, Matteo Paltenghi, Jia Le Tian, Michael Pradel, and Lingming Zhang. 2024 · 2024
Closest in time.
UniLog: Automatic Logging via LLM and In-Context Learning. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–12
Junjielong Xu, Ziang Cui, Yuan Zhao, Xu Zhang, Shilin He, Pinjia He, Liqun Li, Yu Kang, Qingwei Lin, Yingnong Dang, et al · 2024
Closest in time.
Large language models for test-free fault localization. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–12
Aidan ZH Yang, Claire Le Goues, Ruben Martins, and Vincent Hellendoorn. 2024 · 2024
Closest in time.
Kechi Zhang, Jia Li, Ge Li, Xianjie Shi, and Zhi Jin. 2024b · 2024
Closest in time.
Selene: Pioneering Automated Proof in Software Verification
Lichen Zhang, Shuai Lu, and Nan Duan. 2024c · 2024
Closest in time.
APPT: Boosting Automated Patch Correctness Prediction via Fine-tuning Pre-trained Models
Quanjun Zhang, Chunrong Fang, Weisong Sun, Yan Liu, Tieke He, Xiaodong Hao, and Zhenyu Chen. 2024a · 2024
Closest in time.
Experimenting a New Programming Practice with LLMs
Simiao Zhang, Jiaping Wang, Guoliang Dong, Jun Sun, Yueling Zhang, and Geguang Pu. 2024d · 2024
Closest in time.
Summarizing source code using a neural attention model. In 54th Annual Meeting of the Association for Computational Linguistics 2016 . Association for Computational Linguistics, 2073–2083
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung, and Luke Zettlemoyer. 2016 · 2083
Closest in time.