Fetching the paper…
Reading the bibliography…
Although large language models (LLMs) have demonstrated impressive ability in code generation, they are still struggling to address the complicated intent provided by humans.
Agile software development methods: Review and analysis
Pekka Abrahamsson, Outi Salo, Jussi Ronkainen, and Juhani Warsta. 2002 · 2002
Earlier work this paper cites.
The Waterfall Model in Large-Scale Development. In PROFES (Lecture Notes in Business Information Processing, Vol. 32) . Springer, 386–400
Kai Petersen, Claes Wohlin, and Dejan Baca. 2009 · 2009
Earlier work this paper cites.
Software development lifecycle models
Nayan B. Ruparelia. 2010 · 2010
Earlier work this paper cites.
Requirements engineering
Linda A Macaulay. 2012 · 2012
Earlier work this paper cites.
Code completion with statistical language models. In PLDI . ACM, 419–428
Veselin Raychev, Martin T. Vechev, and Eran Yahav. 2014 · 2014
Earlier work this paper cites.
Deep API learning. In SIGSOFT FSE . ACM, 631–642
Xiaodong Gu, Hongyu Zhang, Dongmei Zhang, and Sunghun Kim. 2016 · 2016
Earlier work this paper cites.
Data Recombination for Neural Semantic Parsing. In ACL (1) . The Association for Computer Linguistics
Robin Jia and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Latent Predictor Networks for Code Generation. In ACL (1) . The Association for Computer Linguistics
Wang Ling, Phil Blunsom, Edward Grefenstette, Karl Moritz Hermann, Tomás Kociský, Fumin Wang, and Andrew W. Senior. 2016 · 2016
Earlier work this paper cites.
DeepAM: Migrate APIs with Multi-modal Sequence to Sequence Learning. In IJCAI . ijcai.org, 3675–3681
Xiaodong Gu, Hongyu Zhang, Dongmei Zhang, and Sunghun Kim. 2017 · 2017
Earlier work this paper cites.
Abstract Syntax Networks for Code Generation and Semantic Parsing. In ACL (1) . Association for Computational Linguistics, 1139–1149
Maxim Rabinovich, Mitchell Stern, and Dan Klein. 2017 · 2017
Earlier work this paper cites.
A Syntactic Neural Model for General-Purpose Code Generation. In ACL (1) . Association for Computational Linguistics, 440–450
Pengcheng Yin and Graham Neubig. 2017 · 2017
Earlier work this paper cites.
TRANX: A Transition-based Neural Abstract Syntax Parser for Semantic Parsing and Code Generation. In EMNLP (Demonstration) . Association for Computational Linguistics, 7–12
Pengcheng Yin and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
A Grammar-Based Structural CNN Decoder for Code Generation. In AAAI . AAAI Press, 7055–7062
Zeyu Sun, Qihao Zhu, Lili Mou, Yingfei Xiong, Ge Li, and Lu Zhang. 2019 · 2019
Earlier work this paper cites.
Code Generation as a Dual Task of Code Summarization. In NeurIPS . 6559–6569
Bolin Wei, Ge Li, Xin Xia, Zhiyi Fu, and Zhi Jin. 2019 · 2019
Earlier work this paper cites.
Language Models are Few-Shot Learners. In NeurIPS 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Earlier work this paper cites.
CodeBLEU: a Method for Automatic Evaluation of Code Synthesis
Shuo Ren, Daya Guo, Shuai Lu, Long Zhou, Shujie Liu, Duyu Tang, Neel Sundaresan, Ming Zhou, Ambrosio Blanco, and Shuai Ma. 2020 · 2020
Earlier work this paper cites.
TreeGen: A Tree-Based Transformer Architecture for Code Generation. In AAAI . AAAI Press, 8984–8991
Zeyu Sun, Qihao Zhu, Yingfei Xiong, Yican Sun, Lili Mou, and Lu Zhang. 2020 · 2020
Earlier work this paper cites.
Leap-Of-Thought: Teaching Pre-Trained Models to Systematically Reason Over Implicit Knowledge. In NeurIPS
Alon Talmor, Oyvind Tafjord, Peter Clark, Yoav Goldberg, and Jonathan Berant. 2020 · 2020
Earlier work this paper cites.
Program Synthesis with Large Language Models
Jacob Austin, Augustus Odena, Maxwell I. Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie J. Cai, Michael Terry, Quoc V. Le, and Charles Sutton. 2021 · 2021
Earlier work this paper cites.
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Pondé de Oliveira Pinto, Jared Kaplan, Harrison Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Joshua Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba. 2021 · 2021
Earlier work this paper cites.
Training Verifiers to Solve Math Word Problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Earlier work this paper cites.
Measuring Coding Challenge Competence With APPS. In NeurIPS Datasets and Benchmarks
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
Neural Program Generation Modulo Static Analysis. In NeurIPS . 18984–18996
Rohan Mukherjee, Yeming Wen, Dipak Chaudhari, Thomas W. Reps, Swarat Chaudhuri, and Christopher M. Jermaine. 2021 · 2021
Cited alongside, same era.
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation. In EMNLP (1) . 8696–8708
Yue Wang, Weishi Wang, Shafiq R. Joty, and teven C. H. Hoi. 2021 · 2021
Cited alongside, same era.
Wenhu Chen, Xueguang Ma, Xinyi Wang, and William W. Cohen. 2022 · 2022
Cited alongside, same era.
Scaling Instruction-Finetuned Language Models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Y. Zhao, Yanping Huang, Andrew M. Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2022 · 2022
Cited alongside, same era.
Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution
Chrisantha Fernando, Dylan Banarse, Henryk Michalewski, Simon Osindero, and Tim Rocktäschel. 2023 · 2023
Closest in time.
Self-Verification Improves Few-Shot Clinical Information Extraction
Zelalem Gero, Chandan Singh, Hao Cheng, Tristan Naumann, Michel Galley, Jianfeng Gao, and Hoifung Poon. 2023 · 2023
Closest in time.
Large Language Models Can Self-Improve. In EMNLP . Association for Computational Linguistics, 1051–1068
Jiaxin Huang, Shixiang Gu, Le Hou, Yuexin Wu, Xuezhi Wang, Hongkun Yu, and Jiawei Han. 2023 · 2023
Closest in time.
SkCoder: A Sketch-based Approach for Automatic Code Generation. In ICSE
Jia Li, Yongmin Li, Ge Li, Zhi Jin, Yiyang Hao, and Xing Hu. 2023 · 2023
Closest in time.
Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
Pengfei Liu, Weizhe Yuan, Jinlan Fu, Zhengbao Jiang, Hiroaki Hayashi, and Graham Neubig. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
InCoder: A Generative Model for Code Infilling and Synthesis
Daniel Fried, Armen Aghajanyan, Jessy Lin, Sida Wang, Eric Wallace, Freda Shi, Ruiqi Zhong, Wen-tau Yih, Luke Zettlemoyer, and Mike Lewis. 2022 · 2022
Cited alongside, same era.
PAL: Program-aided Language Models
Luyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon, Pengfei Liu, Yiming Yang, Jamie Callan, and Graham Neubig. 2022 · 2022
Cited alongside, same era.
UniXcoder: Unified Cross-Modal Pre-training for Code Representation. In ACL (1) . Association for Computational Linguistics, 7212–7225
Daya Guo, Shuai Lu, Nan Duan, Yanlin Wang, Ming Zhou, and Jian Yin. 2022 · 2022
Cited alongside, same era.
Large Language Models are Zero-Shot Reasoners. In NeurIPS
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Cited alongside, same era.
Solving Quantitative Reasoning Problems with Language Models. In NeurIPS
Aitor Lewkowycz, Anders Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay V. Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, Yuhuai Wu, Behnam Neyshabur, Guy Gur-Ari, and Vedant Misra. 2022 · 2022
Cited alongside, same era.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Cited alongside, same era.
Language Models of Code are Few-Shot Commonsense Learners. In EMNLP . Association for Computational Linguistics, 1384–1403
Aman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang, and Graham Neubig. 2022 · 2022
Cited alongside, same era.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Cited alongside, same era.
Closest in time.
Self-Refine: Iterative Refinement with Self-Feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, Sean Welleck, Bodhisattwa Prasad Majumder, Shashank Gupta, Amir Yazdanbakhsh, and Peter Clark. 2023 · 2023
Closest in time.
Adversarial Robustness of Prompt-based Few-Shot Learning for Natural Language Understanding. In ACL (Findings) . Association for Computational Linguistics, 2196–2208
Venkata Prabhakara Sarath Nookala, Gaurav Verma, Subhabrata Mukherjee, and Srijan Kumar. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Large Language Models for Education: Grading Open-Ended Questions Using ChatGPT. In SBES . ACM, 293–302
Gustavo Pinto, Isadora Cardoso-Pereira, Danilo Monteiro, Danilo Lucena, Alberto Luiz Oliveira Tavares de Souza, and Kiev Gama. 2023 · 2023
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Yongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li, Weiming Lu, and Yueting Zhuang. 2023 · 2023
Closest in time.
ProgPrompt: Generating Situated Robot Task Plans using Large Language Models. In ICRA . IEEE, 11523–11530
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg. 2023 · 2023
Closest in time.
LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models. In ICCV . IEEE, 2986–2997
Chan Hee Song, Brian M. Sadler, Jiaman Wu, Wei-Lun Chao, Clayton Washington, and Yu Su. 2023 · 2023
Closest in time.
Self-Criticism: Aligning Large Language Models with their Understanding of Helpfulness, Honesty, and Harmlessness. In EMNLP (Industry Track) . Association for Computational Linguistics, 650–662
Xiaoyu Tan, Shaojie Shi, Xihe Qiu, Chao Qu, Zhenting Qi, Yinghui Xu, and Yuan Qi. 2023 · 2023
Closest in time.
Zihao Wang, Shaofei Cai, Anji Liu, Xiaojian Ma, and Yitao Liang. 2023a · 2023
Closest in time.
Self-Evaluation Guided Beam Search for Reasoning. In Thirty-seventh Conference on Neural Information Processing Systems
Yuxi Xie, Kenji Kawaguchi, Yiran Zhao, Xu Zhao, Min-Yen Kan, Junxian He, and Qizhe Xie. 2023 · 2023
Closest in time.
ReAct: Synergizing Reasoning and Acting in Language Models. In ICLR . OpenReview.net
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik R. Narasimhan, and Yuan Cao. 2023 · 2023
Closest in time.
Planning with Large Language Models for Code Generation
Shun Zhang, Zhenfang Chen, Yikang Shen, Mingyu Ding, Joshua B. Tenenbaum, and Chuang Gan. 2023a · 2023
Closest in time.
CodeGeeX: A Pre-Trained Model for Code Generation with Multilingual Evaluations on HumanEval-X
Qinkai Zheng, Xiao Xia, Xu Zou, Yuxiao Dong, Shan Wang, Yufei Xue, Zihan Wang, Lei Shen, Andi Wang, Yang Li, Teng Su, Zhilin Yang, and Jie Tang. 2023 · 2023
Closest in time.
Large Language Models are Human-Level Prompt Engineers. In ICLR . OpenReview.net
Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba. 2023 · 2023
Closest in time.
PromptBench: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
Kaijie Zhu, Jindong Wang, Jiaheng Zhou, Zichen Wang, Hao Chen, Yidong Wang, Linyi Yang, Wei Ye, Neil Zhenqiang Gong, Yue Zhang, and Xing Xie. 2023 · 2023
Closest in time.
Yihong Dong, Xue Jiang, Huanyu Liu, Zhi Jin, and Ge Li. 2024 · 2024
Closest in time.