Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have gained considerable traction within the Software Engineering (SE) community, impacting various SE tasks from code completion to test generation, from program repair to code summarization.
EvoSuite: Automatic Test Suite Generation for Object-Oriented Software. In Proceedings of the 19th ACM SIGSOFT ESEC/FSE (Szeged, Hungary) (ESEC/FSE ’11) . ACM, New York, NY, USA, 416–419
Gordon Fraser and Andrea Arcuri. 2011 · 2011
Earlier work this paper cites.
A hitchhiker’s guide to statistical tests for assessing randomized algorithms in software engineering
Andrea Arcuri and Lionel Briand. 2014 · 2014
Earlier work this paper cites.
Defects4J: A Database of Existing Faults to Enable Controlled Testing Studies for Java Programs. In Proceedings of the 2014 International Symposium on Software Testing and Analysis (San Jose, CA, USA) (ISSTA 2014) . ACM, NY, USA, 437–440
René Just, Darioush Jalali, and Michael D. Ernst. 2014 · 2014
Earlier work this paper cites.
Reformulating branch coverage as a many-objective optimization problem. In 2015 IEEE 8th international conference on software testing, verification and validation (ICST) . IEEE, 1–10
Annibale Panichella, Fitsum Meshesha Kifetew, and Paolo Tonella. 2015 · 2015
Earlier work this paper cites.
A systematic review on code clone detection
Qurat Ul Ain, Wasi Haider Butt, Muhammad Waseem Anwar, Farooque Azam, and Bilal Maqbool. 2019 · 2019
Earlier work this paper cites.
Embedding Java Classes with code2vec: Improvements from Variable Obfuscation. In 2020 IEEE/ACM 17th International Conference on Mining Software Repositories (MSR) . IEEE, New York, NY, USA, 243–253
Rhys Compton, Eibe Frank, Panos Patros, and Abigail Koay. 2020 · 2020
Earlier work this paper cites.
Unit test case generation with transformers and focal context
Michele Tufano, Dawn Drain, Alexey Svyatkovskiy, Shao Kun Deng, and Neel Sundaresan. 2020 · 2020
Earlier work this paper cites.
Adversarial examples for models of code
Noam Yefet, Uri Alon, and Eran Yahav. 2020 · 2020
Earlier work this paper cites.
Assessing Robustness of ML-Based Program Analysis Tools using Metamorphic Program Transformations. In 2021 36th IEEE/ACM International Conference on Automated Software Engineering (ASE) . 1377–1381
Leonhard Applis, Annibale Panichella, and Arie van Deursen. 2021 · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Training Data Leakage Analysis in Language Models. (February 2021)
Huseyin Atahan Inan, Osman Ramadan, Lukas Wutschitz, Daniel Jones, Victor Rühle, James Withers, and Robert Sim. 2021 · 2021
Earlier work this paper cites.
Sbst tool competition 2021. In 2021 IEEE/ACM 14th International Workshop on Search-Based Software Testing (SBST) . IEEE, 20–27
Sebastiano Panichella, Alessio Gambi, Fiorella Zampetti, and Vincenzo Riccio. 2021 · 2021
Earlier work this paper cites.
Is github’s copilot as bad as humans at introducing vulnerabilities in code?
Owura Asare, Meiyappan Nagappan, and N Asokan. 2022 · 2022
Earlier work this paper cites.
Multi-lingual evaluation of code generation models
Ben Athiwaratkun, Sanjay Krishna Gouda, Zijian Wang, Xiaopeng Li, Yuchen Tian, Ming Tan, Wasi Uddin Ahmad, Shiqi Wang, Qing Sun, Mingyue Shang, et al · 2022
Earlier work this paper cites.
Counterfactual explanations for models of code. In Proceedings of the 44th International Conference on Software Engineering: Software Engineering in Practice . 125–134
Jürgen Cito, Isil Dillig, Vijayaraghavan Murali, and Satish Chandra. 2022 · 2022
Earlier work this paper cites.
Codex Hacks HackerRank: Memorization Issues and a Framework for Code Synthesis Evaluation
Anjan Karmakar, Julian Aron Prenner, Marco D’Ambros, and Romain Robbes. 2022 · 2022
Cited alongside, same era.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Cited alongside, same era.
Natural attack for pre-trained models of code. In Proceedings of the 44th International Conference on Software Engineering . 1482–1493
Zhou Yang, Jieke Shi, Junda He, and David Lo. 2022 · 2022
Cited alongside, same era.
Hugging Face – The AI community building the future
2023 · 2023
Cited alongside, same era.
LeetCode - The World’s Leading Online Programming Learning Platform
2023 · 2023
Cited alongside, same era.
CodaMosa: Escaping Coverage Plateaus in Test Generation with Pre-Trained Large Language Models. In Proc. of the 45th International Conference on Software Engineering (ICSE ’23) . IEEE Press, Melbourne, Victoria, Australia, 919–931
Caroline Lemieux, Jeevana Priya Inala, Shuvendu K. Lahiri, and Siddhartha Sen. 2023 · 2023
Closest in time.
Application of Large Language Models to Software Engineering Tasks: Opportunities, Risks, and Implications
Ipek Ozkaya. 2023 · 2023
Closest in time.
Examining Zero-Shot Vulnerability Repair with Large Language Models. In 2023 IEEE Symposium on Security and Privacy (SP) . IEEE, Los Alamitos, CA, USA, 2339–2356
H. Pearce, B. Tan, B. Ahmad, R. Karri, and B. Dolan-Gavitt. 2023 · 2023
Closest in time.
On the Challenges of Using Black-Box APIs for Toxicity Evaluation in Research
Luiza Pozzobon, Beyza Ermis, Patrick Lewis, and Sara Hooker. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The (ab)use of Open Source Code to Train Large Language Models. In 2023 IEEE/ACM 2nd International Workshop on Natural Language-Based Software Engineering (NLBSE)
Ali Al-Kaswan and Maliheh Izadi. 2023 · 2023
Cited alongside, same era.
A3Test: Assertion-Augmented Automated Test Case Generation
Saranya Alagarsamy, Chakkrit Tantithamthavorn, and Aldeida Aleti. 2023 · 2023
Cited alongside, same era.
The Falcon Series of Language Models: Towards Open Frontier Models
Ebtesam Almazrouei, Hamza Alobeidli, Abdulaziz Alshamsi, Alessandro Cappelli, Ruxandra Cojocaru, Maitha Alhammadi, Mazzotta Daniele, Daniel Heslow, Julien Launay, Quentin Malartic, Badreddine Noune, Baptiste Pannier, and Guilherme Penedo. 2023 · 2023
Cited alongside, same era.
Searching for Quality: Genetic Algorithms and Metamorphic Testing for Software Engineering ML. In Proc. of the Genetic and Evolutionary Computation Conference . 1490–1498
Leonhard Applis, Annibale Panichella, and Ruben Marang. 2023 · 2023
Cited alongside, same era.
Authors. 2023
2023
Cited alongside, same era.
How is ChatGPT’s behavior changing over time?
Lingjiao Chen, Matei Zaharia, and James Zou. 2023 · 2023
Cited alongside, same era.
JUGE: An infrastructure for benchmarking Java unit test generators
Xavier Devroey, Alessio Gambi, Juan Pablo Galeotti, René Just, Fitsum Kifetew, Annibale Panichella, and Sebastiano Panichella. 2023 · 2023
Cited alongside, same era.
Mohammed Latif Siddiq, Joanna Santos, Ridwanul Hasan Tanvir, Noshin Ulfat, Fahmid Al Rifat, and Vinicius Carvalho Lopes. 2023 · 2023
Closest in time.
ChatGPT vs SBST: A Comparative Assessment of Unit Test Suite Generation
Yutian Tang, Zhijie Liu, Zhichao Zhou, and Xiapu Luo. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Closest in time.
A systematic review of Green AI
Roberto Verdecchia, June Sallou, and Luís Cruz. 2023 · 2023
Closest in time.
Large Language Models in Fault Localisation
Yonghao Wu, Zheng Li, Jie M Zhang, Mike Papadakis, Mark Harman, and Yong Liu. 2023 · 2023
Closest in time.
Automated program repair in the era of large pre-trained language models. In Proceedings of the 45th International Conference on Software Engineering (ICSE 2023). Association for Computing Machinery
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Closest in time.
Keep the Conversation Going: Fixing 162 out of 337 bugs for $0.42 each using ChatGPT
Chunqiu Steven Xia and Lingming Zhang. 2023 · 2023
Closest in time.
Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility
Wentao Ye, Mingfeng Ou, Tianyi Li, Xuetao Ma, Yifan Yanggong, Sai Wu, Jie Fu, Gang Chen, Junbo Zhao, et al · 2023
Closest in time.
Siren’s Song in the AI Ocean: A Survey on Hallucination in Large Language Models
Yue Zhang, Yafu Li, Leyang Cui, Deng Cai, Lemao Liu, Tingchen Fu, Xinting Huang, Enbo Zhao, Yu Zhang, Yulong Chen, et al · 2023
Closest in time.