Fetching the paper…
Reading the bibliography…
Iteratively improving and repairing source code with large language models (LLMs), known as refinement, has emerged as a popular way of generating programs that would be too complex to construct in one shot.
Arm-acquiring bandits
Peter Whittle · 1981
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Christopher M. Bishop · 2006
Earlier work this paper cites.
Efficient selectivity and backup operators in monte-carlo tree search
Rémi Coulom · 2006
Earlier work this paper cites.
Generating all polynomial invariants in simple loops
Enric Rodríguez-Carbonell and Deepak Kapur · 2007
Earlier work this paper cites.
Generating all polynomial invariants in simple loops
E. Rodríguez-Carbonell and D. Kapur · 2007
Earlier work this paper cites.
Z3: An efficient smt solver
Leonardo De Moura and Nikolaj Bjørner · 2008
Earlier work this paper cites.
Progressive strategies for monte-carlo tree search
Guillaume M Jb Chaslot, Mark HM Winands, H Jaap van den Herik, Jos WHM Uiterwijk, and Bruno Bouzy · 2008
Earlier work this paper cites.
Automatically finding patches using genetic programming
Westley Weimer, ThanhVu Nguyen, Claire Le Goues, and Stephanie Forrest · 2009
Earlier work this paper cites.
Monte-carlo planning in large pomdps
David Silver and Joel Veness · 2010
Earlier work this paper cites.
An empirical evaluation of thompson sampling
Olivier Chapelle and Lihong Li · 2011
Earlier work this paper cites.
Using dynamic analysis to discover polynomial and array invariants
ThanhVu Nguyen, Deepak Kapur, Westley Weimer, and Stephanie Forrest · 2012
Earlier work this paper cites.
Semfix: Program repair via semantic analysis
Hoang Duong Thien Nguyen, Dawei Qi, Abhik Roychoudhury, and Satish Chandra · 2013
Earlier work this paper cites.
Automated feedback generation for introductory programming assignments
Rishabh Singh, Sumit Gulwani, and Armando Solar-Lezama · 2013
Earlier work this paper cites.
Monte carlo tree search in continuous action spaces with execution uncertainty
Timothy Yee, Viliam Lisỳ, Michael H Bowling, and S Kambhampati · 2016
Earlier work this paper cites.
A tutorial on thompson sampling
Daniel J Russo, Benjamin Van Roy, Abbas Kazerouni, Ian Osband, Zheng Wen, et al · 2018
Cited alongside, same era.
Learning loop invariants for program verification
Xujie Si, Hanjun Dai, Mukund Raghothaman, Mayur Naik, and Le Song · 2018
Cited alongside, same era.
Automated program repair
Claire Le Goues, Michael Pradel, and Abhik Roychoudhury · 2019
Cited alongside, same era.
Learning nonlinear loop invariants with gated continuous logic networks
Jianan Yao, Gabriel Ryan, Justin Wong, Suman Jana, and Ronghui Gu · 2020
Cited alongside, same era.
Global guidance for local generalization in model checking
Hari Govind Vediramana Krishnan, YuTing Chen, Sharon Shoham, and Arie Gurfinkel · 2020
Cited alongside, same era.
Measuring coding challenge competence with apps
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, et al · 2021
Is self-repair a silver bullet for code generation?, 2023
Theo X. Olausson, Jeevana Priya Inala, Chenglong Wang, Jianfeng Gao, and Armando Solar-Lezama · 2023
Later among the works it cites.
Reflexion: an autonomous agent with dynamic memory and self-reflection
Noah Shinn, Beck Labash, and Ashwin Gopinath · 2023
Later among the works it cites.
Parsel: Algorithmic reasoning with language models by composing decompositions
Eric Zelikman, Qian Huang, Gabriel Poesia, Noah D Goodman, and Nick Haber · 2023
Later among the works it cites.
Automated repair of programs from large language models
Zhiyu Fan, Xiang Gao, Martin Mirchev, Abhik Roychoudhury, and Shin Hwei Tan · 2023
Later among the works it cites.
Automated program repair in the era of large pre-trained language models
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Cited alongside, same era.
Program synthesis with large language models, 2021
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, and Charles Sutton · 2021
Cited alongside, same era.
Fast and flexible: Human program induction in abstract reasoning tasks
Aysja Johnson, Wai Keen Vong, Brenden M Lake, and Todd M Gureckis · 2021
Cited alongside, same era.
Generating sequences by learning to self-correct
Sean Welleck, Ximing Lu, Peter West, Faeze Brahman, Tianxiao Shen, Daniel Khashabi, and Yejin Choi · 2022
Cited alongside, same era.
CodeRL: Mastering code generation through pretrained models and deep reinforcement learning
Hung Le, Yue Wang, Akhilesh Deepak Gotmare, Silvio Savarese, and Steven Hoi · 2022
Cited alongside, same era.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Cited alongside, same era.
Badge: prioritizing ui events with hierarchical multi-armed bandits for automated ui testing
Dezhi Ran, Hao Wang, Wenyu Wang, and Tao Xie · 2023
Later among the works it cites.
Whitefox: White-box compiler fuzzing empowered by large language models
Chenyuan Yang, Yinlin Deng, Runyu Lu, Jiayi Yao, Jiawei Liu, Reyhaneh Jabbarvand, and Lingming Zhang · 2023
Later among the works it cites.
Finding inductive loop invariants using large language models, 2023
Adharsh Kamath, Aditya Senthilnathan, Saikat Chakraborty, Pantazis Deligiannis, Shuvendu K. Lahiri, Akash Lal, Aseem Rastogi, Subhajit Roy, and Rahul Sharma · 2023
Later among the works it cites.
Self-refine: Iterative refinement with self-feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, et al · 2024
Closest in time.
Worldcoder, a model-based llm agent: Building world models by writing code and interacting with the environment
Hao Tang, Darren Key, and Kevin Ellis · 2024
Closest in time.
Hypothesis search: Inductive reasoning with language models
Ruocheng Wang, Eric Zelikman, Gabriel Poesia, Yewen Pu, Nick Haber, and Noah D. Goodman · 2024
Closest in time.
Phenomenal yet puzzling: Testing inductive reasoning capabilities of language models with hypothesis refinement
Linlu Qiu, Liwei Jiang, Ximing Lu, Melanie Sclar, Valentina Pyatkin, Chandra Bhagavatula, Bailin Wang, Yoon Kim, Yejin Choi, Nouha Dziri, and Xiang Ren · 2024
Closest in time.
Cycle: Learning to self-refine the code generation
Yangruibo Ding, Marcus J Min, Gail Kaiser, and Baishakhi Ray · 2024
Closest in time.