Fetching the paper…
Reading the bibliography…
While large language models (LLMs) now excel at code generation, a key aspect of software development is the art of refactoring: consolidating code into libraries of reusable and readable programs.
A formal theory of inductive inference. Part I
Ray J Solomonoff · 1964
Earlier work this paper cites.
Three approaches for defining the concept of information quantity
Andrei Nikolaevich Kolmogorov · 1965
Earlier work this paper cites.
On the length of programs for computing finite binary sequences
Gregory J Chaitin · 1966
Earlier work this paper cites.
Turtle Geometry: The Computer as a Medium for Exploring Mathematics
Harold Abelson and Andrea diSessa · 1986
Earlier work this paper cites.
Perception as Bayesian inference
David C Knill and Whitman Richards · 1996
Earlier work this paper cites.
The empirical case for two systems of reasoning
Steven A Sloman · 1996
Earlier work this paper cites.
AntiPatterns: refactoring software, architectures, and projects in crisis
William H Brown, Raphael C Malveau, Hays W “Skip” McCormick, and Thomas J Mowbray · 1998
Earlier work this paper cites.
PCFG models of linguistic tree representations
Mark Johnson · 1998
Earlier work this paper cites.
Programming by demonstration: An inductive learning formulation
Tessa A. Lau and Daniel S. Weld · 1998
Earlier work this paper cites.
The mathematics of statistical machine translation: Parameter estimation
Peter F. Brown, Stephen A. Della Pietra, Vincent J. Della Pietra, and Robert L. Mercer · 2003
Earlier work this paper cites.
In two minds: dual-process accounts of reasoning
Jonathan St.B.T. Evans · 2003
Earlier work this paper cites.
Hierarchical Bayesian inference in the visual cortex
Tai Sing Lee and David Mumford · 2003
Earlier work this paper cites.
Object perception as Bayesian inference
Daniel Kersten, Pascal Mamassian, and Alan Yuille · 2004
Earlier work this paper cites.
Vision as Bayesian inference: analysis by synthesis?
Alan Yuille and Daniel Kersten · 2006
Earlier work this paper cites.
Learning programs: A hierarchical Bayesian approach
Percy Liang, Michael I. Jordan, and Dan Klein · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
David L. Chen and Raymond J. Mooney · 2011
Earlier work this paper cites.
Thinking, fast and slow
Daniel Kahneman · 2011
Earlier work this paper cites.
Learning dependency-based compositional semantics
Percy Liang, Michael Jordan, and Dan Klein · 2011
Earlier work this paper cites.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Yoav Artzi and Luke Zettlemoyer · 2013
Earlier work this paper cites.
Bootstrap learning via modular concept discovery
Eyal Dechter, Jonathan Malmaud, Ryan P. Adams, and Joshua B. Tenenbaum · 2013
Earlier work this paper cites.
Integrating programming by example and natural language programming
Mehdi Manshadi, Daniel Gildea, and James Allen · 2013
Earlier work this paper cites.
Mining idioms from source code
Miltiadis Allamanis and Charles Sutton · 2014
Earlier work this paper cites.
Inductive programming meets the real world
Sumit Gulwani, José Hernández-Orallo, Emanuel Kitzelmann, Stephen H Muggleton, Ute Schmid, and Benjamin Zorn · 2015
Earlier work this paper cites.
Human-level concept learning through probabilistic program induction
Brenden M. Lake, Ruslan Salakhutdinov, and Joshua B. Tenenbaum · 2015
Earlier work this paper cites.
Compositional program synthesis from natural language and examples
Mohammad Raza, Sumit Gulwani, and Natasa Milic-Frayling · 2015
Earlier work this paper cites.
Galileo: Perceiving physical object properties by integrating a physics engine with deep learning
Jiajun Wu, Ilker Yildirim, Joseph J. Lim, Bill Freeman, and Joshua B. Tenenbaum · 2015
Earlier work this paper cites.
Neural module networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Earlier work this paper cites.
Program synthesis from polymorphic refinement types
Nadia Polikarpova, Ivan Kuraj, and Armando Solar-Lezama · 2016
Earlier work this paper cites.
Agile software development methods: Review and analysis
Pekka Abrahamsson, Outi Salo, Jussi Ronkainen, and Juhani Warsta · 2017
Cited alongside, same era.
Deepcoder: Learning to write programs
Matej Balog, Alexander L. Gaunt, Marc Brockschmidt, Sebastian Nowozin, and Daniel Tarlow · 2017
Cited alongside, same era.
Robustfill: Neural program learning under noisy I/O
Jacob Devlin, Jonathan Uesato, Surya Bhupatiraju, Rishabh Singh, Abdel-rahman Mohamed, and Pushmeet Kohli · 2017
Cited alongside, same era.
Program synthesis
Sumit Gulwani, Oleksandr Polozov, and Rishabh Singh · 2017
Cited alongside, same era.
Learning to reason: End-to-end module networks for visual question answering
Ronghang Hu, Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Kate Saenko · 2017
Cited alongside, same era.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
Language models can teach themselves to program better
Patrick Haluptzok, Matthew Bowers, and Adam Tauman Kalai · 2022
Later among the works it cites.
Competition-level code generation with AlphaCode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Later among the works it cites.
What makes good in-context examples for GPT-3?
Jiachang Liu, Dinghan Shen, Yizhe Zhang, Bill Dolan, Lawrence Carin, and Weizhu Chen · 2022
Later among the works it cites.
Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity
Yao Lu, Max Bartolo, Alastair Moore, Sebastian Riedel, and Pontus Stenetorp · 2022
Later among the works it cites.
Synchromesh: Reliable code generation from pre-trained language models
Gabriel Poesia, Alex Polozov, Vu Le, Ashish Tiwari, Gustavo Soares, Christopher Meek, and Sumit Gulwani · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C. Lawrence Zitnick, and Ross B. Girshick · 2017
Cited alongside, same era.
Neuro-symbolic program synthesis
Emilio Parisotto, Abdel-rahman Mohamed, Rishabh Singh, Lihong Li, Dengyong Zhou, and Pushmeet Kohli · 2017
Cited alongside, same era.
The end of Moore’s law: A new beginning for information technology
Thomas N. Theis and H.-S. Philip Wong · 2017
Cited alongside, same era.
Neural scene de-rendering
Jiajun Wu, Joshua B. Tenenbaum, and Pushmeet Kohli · 2017
Cited alongside, same era.
Mining Semantic Loop Idioms
Miltiadis Allamanis, Earl T. Barr, Christian Bird, Premkumar Devanbu, Mark Marron, and Charles Sutton · 2018
Cited alongside, same era.
Learning with latent language
Jacob Andreas, Dan Klein, and Sergey Levine · 2018
Cited alongside, same era.
Learning to infer graphics programs from hand-drawn images
Kevin Ellis, Daniel Ritchie, Armando Solar-Lezama, and Josh Tenenbaum · 2018
Cited alongside, same era.
Skill induction and planning with latent language
Pratyusha Sharma, Antonio Torralba, and Jacob Andreas · 2022
Later among the works it cites.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Huai hsin Chi, and Denny Zhou · 2022
Later among the works it cites.
Memorizing transformers
Yuhuai Wu, Markus Norman Rabe, DeLesley Hutchins, and Christian Szegedy · 2022
Later among the works it cites.
ReAct: Synergizing Reasoning and Acting in Language Models, 2022
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2022
Later among the works it cites.
Natural Language to Code Generation in Interactive Data Science Notebooks, 2022
Pengcheng Yin, Wen-Ding Li, Kefan Xiao, Abhishek Rao, Yeming Wen, Kensen Shi, Joshua Howland, Paige Bailey, Michele Catasta, Henryk Michalewski, Alex Polozov, and Charles Sutton · 2022
Later among the works it cites.
Parsel: A (De-)compositional Framework for Algorithmic Reasoning with Language Models, 2022
Eric Zelikman, Qian Huang, Gabriel Poesia, Noah D. Goodman, and Nick Haber · 2022
Later among the works it cites.
An extensive study on pre-trained models for program understanding and generation
Zhengran Zeng, Hanzhuo Tan, Haotian Zhang, Jing Li, Yuqun Zhang, and Lingming Zhang · 2022
Later among the works it cites.
Top-down synthesis for library learning
Matthew Bowers, Theo X. Olausson, Lionel Wong, Gabriel Grand, Joshua B. Tenenbaum, Kevin Ellis, and Armando Solar-Lezama · 2023
Closest in time.
FlashFill++: Scaling Programming by Example by Cutting to the Chase
José Cambronero, Sumit Gulwani, Vu Le, Daniel Perelman, Arjun Radhakrishna, Clint Simon, and Ashish Tiwari · 2023
Closest in time.
babble: Learning better abstractions with e-graphs and anti-unification
David Cao, Rose Kunkel, Chandrakana Nandi, Max Willsey, Zachary Tatlock, and Nadia Polikarpova · 2023
Closest in time.
Chain-of-verification reduces hallucination in large language models
Shehzaad Dhuliawala, Mojtaba Komeili, Jing Xu, Roberta Raileanu, Xian Li, Asli Celikyilmaz, and Jason Weston · 2023
Closest in time.
StarCoder: May the source be with you!, 2023
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, Jia Li, Jenny Chim, Qian Liu, Evgenii Zheltonozhskii, Terry Yue Zhuo, Thomas Wang, Olivier Dehaene, Mishig Davaadorj, Joel Lamy-Poirier, João Monteiro, Oleh Shliazhko, Nicolas Gontier, Nicholas Meade, Armel Zebaze, Ming-Ho Yee, Logesh Kumar Umapathi, Jian Zhu, Benjamin Lipkin, Muhtasham Oblokulov, Zhiruo Wang, Rudra Murthy, Jason Stillerman, Siva Sankalp Patel, Dmitry Abulkhanov, Marco Zocca, Manan Dey, Zhihan Zhang, Nour Fahmy, Urvashi Bhattacharyya, Wenhao Yu, Swayam Singh, Sasha Luccioni, Paulo Villegas, Maxim Kunakov, Fedor Zhdanov, Manuel Romero, Tony Lee, Nadav Timor, Jennifer Ding, Claire Schlesinger, Hailey Schoelkopf, Jan Ebert, Tri Dao, Mayank Mishra, Alex Gu, Jennifer Robinson, Carolyn Jane Anderson, Brendan Dolan-Gavitt, Danish Contractor, Siva Reddy, Daniel Fried, Dzmitry Bahdanau, Yacine Jernite, Carlos Muñoz Ferrandis, Sean Hughes, Thomas Wolf, Arjun Guha, Leandro von Werra, and Harm de Vries · 2023
Closest in time.
Lost in the Middle: How Language Models Use Long Contexts, 2023
Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang · 2023
Closest in time.
Antonio Valerio Miceli-Barone, Fazl Barez, Ioannis Konstas, and Shay B. Cohen · 2023
Closest in time.
CodeGen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong · 2023
Closest in time.
OpenAI · 2023
Closest in time.
Generative Agents: Interactive Simulacra of Human Behavior, 2023
Joon Sung Park, Joseph C. O’Brien, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom · 2023
Closest in time.
Reflexion: Language Agents with Verbal Reinforcement Learning, 2023
Noah Shinn, Federico Cassano, Beck Labash, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao · 2023
Closest in time.
Voyager: An Open-Ended Embodied Agent with Large Language Models, 2023
Guanzhi Wang, Yuqi Xie, Yunfan Jiang, Ajay Mandlekar, Chaowei Xiao, Yuke Zhu, Linxi Fan, and Anima Anandkumar · 2023
Closest in time.
ChatGPT gets its “Wolfram Superpowers”, 2023
Stephen Wolfram · 2023
Closest in time.
Large language models are human-level prompt engineers
Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba · 2023
Closest in time.