Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have significantly improved their ability to perform tasks in the field of code generation.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Using benchmarking to advance research: A challenge to software engineering. In 25th International Conference on Software Engineering, 2003. Proceedings. IEEE, 74–83
Susan Elliott Sim, Steve Easterbrook, and Richard C Holt. 2003 · 2003
Earlier work this paper cites.
The education of a software engineer. In Proceedings. 19th International Conference on Automated Software Engineering, 2004. IEEE, xviii–xxvii
Mehdi Jazayeri. 2004 · 2004
Earlier work this paper cites.
Communication and co-ordination practices in software engineering projects
Ian R McChesney and Seamus Gallagher. 2004 · 2004
Earlier work this paper cites.
Software engineering: a practitioner’s approach
Roger S Pressman. 2005 · 2005
Earlier work this paper cites.
Collaboration in software engineering: A roadmap. In Future of Software Engineering (FOSE’07) . IEEE, 214–225
Jim Whitehead. 2007 · 2007
Earlier work this paper cites.
Collaborative software engineering: challenges and prospects
Ivan Mistrík, John Grundy, Andre Van der Hoek, and Jim Whitehead. 2010 · 2010
Earlier work this paper cites.
Applications of ontologies in requirements engineering: a systematic review of the literature
Diego Dermeval, Jéssyka Vilela, Ig Ibert Bittencourt, Jaelson Castro, Seiji Isotani, Patrick Brito, and Alan Silva. 2016 · 2016
Earlier work this paper cites.
Abstract Syntax Networks for Code Generation and Semantic Parsing
M. Rabinovich, M. Stern, and D. Klein. 2017 · 2017
Earlier work this paper cites.
Attention is All You Need. In Advances in Neural Information Processing Systems , Vol. 30
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin. 2017 · 2017
Earlier work this paper cites.
Can Neural Machine Translation Be Improved with User Feedback?. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT)
Julia Kreutzer, Shahram Khadivi, Evgeny Matusov, and Stefan Riezler. 2018 · 2018
Earlier work this paper cites.
Learning to mine aligned code and natural language pairs from stack overflow. In Proceedings of the 15th international conference on mining software repositories . 476–486
Pengcheng Yin, Bowen Deng, Edgar Chen, Bogdan Vasilescu, and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
Code2Vec: Learning Distributed Representations of Code
U. Alon, M. Zilberstein, O. Levy, and E. Yahav. 2019 · 2019
Earlier work this paper cites.
CodeBERT: A Pre-trained Model for Programming and Natural Languages
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, and D. Jiang et al. 2020 · 2020
Earlier work this paper cites.
Intellicode Compose: Code Generation Using Transformer. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1433–1443
A. Svyatkovskiy, S. K. Deng, S. Fu, and N. Sundaresan. 2020 · 2020
Earlier work this paper cites.
Unit Test Case Generation with Transformers and Focal Context
M. Tufano, D. Drain, A. Svyatkovskiy, S. Deng, and N. Sundaresan. 2020 · 2020
Earlier work this paper cites.
Leveraging Code Generation to Improve Code Retrieval and Summarization via Dual Learning. In Proceedings of The Web Conference 2020 . 2309–2319
W. Ye, R. Xie, J. Zhang, T. Hu, X. Wang, and S. Zhang. 2020 · 2020
Earlier work this paper cites.
Program synthesis with large language models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, et al · 2021
Earlier work this paper cites.
InferCode: Self-supervised Learning of Code Representations by Predicting Subtrees. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . 1186–1197
N. D. Bui, Y. Yu, and L. Jiang. 2021 · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Measuring coding challenge competence with apps
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, et al · 2021
Earlier work this paper cites.
Requirement engineering challenges: A systematic mapping study on the academic and the industrial perspective
Muhammad Tukur, Sani Umar, and Jameleddine Hassine. 2021 · 2021
Earlier work this paper cites.
GPT-J-6B: A 6 billion parameter autoregressive language model
Ben Wang and Aran Komatsuzaki. 2021 · 2021
Earlier work this paper cites.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi. 2021 · 2021
Earlier work this paper cites.
Constitutional AI: Harmlessness from AI Feedback
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, Carol Chen, Catherine Olsson, Christopher Olah, Danny Hernandez, Dawn Drain, Deep Ganguli, Dustin Li, Eli Tran-Johnson, Ethan Perez, Jamie Kerr, Jared Mueller, Jeffrey Ladish, Joshua Landau, Kamal Ndousse, Kamile Lukosiute, Liane Lovitt, Michael Sellitto, Nelson Elhage, Nicholas Schiefer, Noemí Mercado, Nova DasSarma, Robert Lasenby, Robin Larson, Sam Ringer, Scott Johnston, Shauna Kravec, Sheer El Showk, Stanislav Fort, Tamera Lanham, Timothy Telleen-Lawton, Tom Conerly, Tom Henighan, Tristan Hume, Samuel R. Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan. 2022 · 2022
Earlier work this paper cites.
Improving Alignment of Dialogue Agents via Targeted Human Judgements
Amelia Glaese, Nat McAleese, Maja Trebacz, John Aslanides, Vlad Firoiu, Timo Ewalds, Maribeth Rauh, Laura Weidinger, Martin J. Chadwick, Phoebe Thacker, Lucy Campbell-Gillingham, Jonathan Uesato, Po-Sen Huang, Ramona Comanescu, Fan Yang, Abigail See, Sumanth Dathathri, Rory Greig, Charlie Chen, Doug Fritz, Jaume Sanchez Elias, Richard Green, Sona Mokrá, Nicholas Fernando, Boxi Wu, Rachel Foley, Susannah Young, Iason Gabriel, William Isaac, John Mellor, Demis Hassabis, Koray Kavukcuoglu, Lisa Anne Hendricks, and Geoffrey Irving. 2022 · 2022
Earlier work this paper cites.
Large Language Models Can Self-Improve
Jiaxin Huang, Shixiang Shane Gu, Le Hou, Yuexin Wu, Xuezhi Wang, Hongkun Yu, and Jiawei Han. 2022 · 2022
Earlier work this paper cites.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Earlier work this paper cites.
An Empirical Evaluation of GitHub Copilot’s Code Suggestions. In Proceedings of the 19th International Conference on Mining Software Repositories . 1–5
N. Nguyen and S. Nadi. 2022 · 2022
Earlier work this paper cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Earlier work this paper cites.
Training Language Models to Follow Instructions with Human Feedback. In Proceedings of the Annual Conference on Neural Information Processing Systems (NeurIPS)
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F. Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Earlier work this paper cites.
What is it like to program with artificial intelligence?
Advait Sarkar, Andrew D Gordon, Carina Negreanu, Christian Poelitz, Sruti Srinivasa Ragavan, and Ben Zorn. 2022 · 2022
Cited alongside, same era.
An Empirical Study of Code Smells in Transformer-Based Code Generation Techniques. In 2022 IEEE 22nd International Working Conference on Source Code Analysis and Manipulation (SCAM) . 71–82
M. L. Siddiq, S. H. Majumder, M. R. Mim, S. Jajodia, and J. C. Santos. 2022 · 2022
Cited alongside, same era.
Choose Your Programming Copilot: A Comparison of the Program Synthesis Performance of GitHub Copilot and Genetic Programming. In Proceedings of the Genetic and Evolutionary Computation Conference . 1019–1027
D. Sobania, M. Briesch, and F. Rothlauf. 2022 · 2022
Cited alongside, same era.
Expectation vs. Experience: Evaluating the Usability of Code Generation Tools Powered by Large Language Models. In CHI Conference on Human Factors in Computing Systems Extended Abstracts . 1–7
P. Vaithilingam, T. Zhang, and E. L. Glassman. 2022 · 2022
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Patrick Schramowski, and Jacob Andreas. 2023 · 2023
Later among the works it cites.
Reflexion: Language agents with verbal reinforcement learning
Noah Shinn, Nelson Labash, Mark Skreta, Shunyu Yao, and Karthik Narasimhan. 2023 · 2023
Later among the works it cites.
Is chatgpt a good nlg evaluator? a preliminary study
Jiaan Wang, Yunlong Liang, Fandong Meng, Zengkui Sun, Haoxiang Shi, Zhixu Li, Jinan Xu, Jianfeng Qu, and Jie Zhou. 2023a · 2023
Later among the works it cites.
A survey on large language model based autonomous agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A systematic evaluation of large language models of code. In Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming . 1–10
Frank F Xu, Uri Alon, Graham Neubig, and Vincent Josua Hellendoorn. 2022 · 2022
Cited alongside, same era.
Generating Natural Language Proofs with Verifier-Guided Search. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing (EMNLP) . 89–105
Kaiyu Yang, Jia Deng, and Danqi Chen. 2022 · 2022
Cited alongside, same era.
Productivity assessment of neural code completion. In Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming . 21–29
Albert Ziegler, Eirini Kalliamvakou, X Alice Li, Andrew Rice, Devon Rifkin, Shawn Simister, Ganesh Sittampalam, and Edward Aftandilian. 2022 · 2022
Cited alongside, same era.
Prompt engineering guidelines for LLMs in Requirements Engineering
Simon Arvidsson and Johan Axell. 2023 · 2023
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Cited alongside, same era.
A categorical archive of chatgpt failures
Ali Borji. 2023 · 2023
Cited alongside, same era.
Improving factuality and reasoning in language models through multiagent debate
Yilun Du, Shuang Li, Antonio Torralba, Joshua B Tenenbaum, and Igor Mordatch. 2023 · 2023
Cited alongside, same era.
Codeparrot
Hugging Face. 2023 · 2023
Cited alongside, same era.
Yixuan Weng, Minjun Zhu, Fei Xia, Bin Li, Shizhu He, Kang Liu, and Jun Zhao. 2023 · 2023
Later among the works it cites.
The rise and potential of large language model based agents: A survey
Zhiheng Xi, Wenxiang Chen, Xin Guo, Wei He, Yiwen Ding, Boyang Hong, Ming Zhang, Junzhe Wang, Senjie Jin, Enyu Zhou, et al · 2023
Later among the works it cites.
Decomposition Enhances Reasoning via Self-Evaluation Guided Decoding
Yuxi Xie, Kenji Kawaguchi, Yiran Zhao, Xu Zhao, MinYen Kan, Junxian He, and Qizhe Xie. 2023 · 2023
Later among the works it cites.
ReAct: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Shixiang Shane Cao, Karthik Narasimhan, and Wen-tau Huang. 2023 · 2023
Later among the works it cites.
Large language models meet nl2code: A survey. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 7443–7464
Daoguang Zan, Bei Chen, Fengji Zhang, Dianjie Lu, Bingchao Wu, Bei Guan, Wang Yongji, and Jian-Guang Lou. 2023 · 2023
Later among the works it cites.
Exploring collaboration mechanisms for llm agents: A social psychology view
Jintian Zhang, Xin Xu, and Shumin Deng. 2023c · 2023
Later among the works it cites.
Self-Edit: Fault-Aware Code Editor for Code Generation
Kechi Zhang, Zhuo Li, Jia Li, Ge Li, and Zhi Jin. 2023b · 2023
Later among the works it cites.
Unifying the perspectives of nlp and software engineering: A survey on language models for code
Ziyin Zhang, Chaoyu Chen, Bingchang Liu, Cong Liao, Zi Gong, Hang Yu, Jianguo Li, and Rui Wang. 2023a · 2023
Later among the works it cites.
Big Code Models Leaderboard
Hugging Face Accessed 2024 · 2024
Closest in time.
Deepseek llm: Scaling open-source language models with longtermism
Xiao Bi, Deli Chen, Guanting Chen, Shanhuang Chen, Damai Dai, Chengqi Deng, Honghui Ding, Kai Dong, Qiushi Du, Zhe Fu, et al · 2024
Closest in time.
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair
Islem Bouzenia, Premkumar Devanbu, and Michael Pradel. 2024 · 2024
Closest in time.
De-Hallucinator: Iterative Grounding for LLM-Based Code Completion
Aryaz Eghbali and Michael Pradel. 2024 · 2024
Closest in time.
LLM-based Test-driven Interactive Code Generation: User Study and Empirical Evaluation
Sarah Fakhoury, Aaditya Naik, Georgios Sakkas, Saikat Chakraborty, and Shuvendu K Lahiri. 2024 · 2024
Closest in time.
Llm-based nlg evaluation: Current status and challenges
Mingqi Gao, Xinyu Hu, Jie Ruan, Xiao Pu, and Xiaojun Wan. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Y Wu, YK Li, et al · 2024
Closest in time.
Large language model based multi-agents: A survey of progress and challenges
Taicheng Guo, Xiuying Chen, Yaqi Wang, Ruidi Chang, Shichao Pei, Nitesh V Chawla, Olaf Wiest, and Xiangliang Zhang. 2024a · 2024
Closest in time.
Ahmed E Hassan, Gustavo A Oliva, Dayi Lin, Boyuan Chen, Zhen Ming, et al · 2024
Closest in time.
SWE-bench: Can Language Models Resolve Real-world Github Issues?. In The Twelfth International Conference on Learning Representations
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik R Narasimhan. 2024 · 2024
Closest in time.
Hao Li, Cor-Paul Bezemer, and Ahmed E Hassan. 2024 · 2024
Closest in time.
Toward a theory of causation for interpreting neural code models
David N Palacio, Alejandro Velasco, Nathan Cooper, Alvaro Rodriguez, Kevin Moran, and Denys Poshyvanyk. 2024 · 2024
Closest in time.
Codepori: Large scale model for autonomous software development by using multi-agents
Zeeshan Rasheed, Muhammad Waseem, Mika Saari, Kari Systä, and Pekka Abrahamsson. 2024 · 2024
Closest in time.
Talking about large language models
Murray Shanahan. 2024 · 2024
Closest in time.
Introducing Devin, the first AI software engineer
Scott Wu. Accessed 2024 · 2024
Closest in time.
Shuyuan Xu, Zelong Li, Kai Mei, and Yongfeng Zhang. 2024 · 2024
Closest in time.
Swe-agent: Agent-computer interfaces enable automated software engineering
John Yang, Carlos E Jimenez, Alexander Wettig, Kilian Lieret, Shunyu Yao, Karthik Narasimhan, and Ofir Press. 2024 · 2024
Closest in time.
Kechi Zhang, Jia Li, Ge Li, Xianjie Shi, and Zhi Jin. 2024 · 2024
Closest in time.