Fetching the paper…
Reading the bibliography…
The Codex model has demonstrated extraordinary competence in synthesizing code from natural language problem descriptions.
Natural language processing in an automatic programming domain
Jerrold M Ginsparg. 1978 · 1978
Earlier work this paper cites.
Automatic Programming through Natural Language Dialogue: A Survey
G. Heidorn. 1986 · 1986
Earlier work this paper cites.
NaturalJava: A Natural Language Interface for Programming in Java
David Price, Ellen Riloff, Joseph Zachary, and On Harvey. 2000 · 2000
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, and Ming Zhou. 2020 · 2002
Earlier work this paper cites.
Language Models are Few-Shot Learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2005
Earlier work this paper cites.
SourceFinder: Finding Malware Source-Code from Publicly Available Repositories
Md Omar Faruk Rokon, Risul Islam, Ahmad Darki, Vagelis E. Papalexakis, and Michalis Faloutsos. 2020 · 2005
Earlier work this paper cites.
Programming With Unrestricted Natural Language. In Proceedings of the Australasian Language Technology Workshop 2005 . Sydney, Australia, 191–199
David Vadas and James R. Curran. 2005 · 2005
Earlier work this paper cites.
NLP (Natural Language Processing) for NLP (Natural Language Programming). 319–330
Rada Mihalcea, Hugo Liu, and Henry Lieberman. 2006 · 2006
Earlier work this paper cites.
GraphCodeBERT: Pre-training Code Representations with Data Flow
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, Shujie Liu, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu, Michele Tufano, Shao Kun Deng, Colin Clement, Dawn Drain, Neel Sundaresan, Jian Yin, Daxin Jiang, and Ming Zhou. 2021 · 2009
Earlier work this paper cites.
Automating string processing in spreadsheets using input-output examples
Sumit Gulwani. 2011 · 2011
Earlier work this paper cites.
On the Naturalness of Software. In Proceedings of the 34th International Conference on Software Engineering (Zurich, Switzerland) (ICSE ’12) . 11 pages
Abram Hindle, Earl T. Barr, Zhendong Su, Mark Gabel, and Premkumar Devanbu. 2012 · 2012
Earlier work this paper cites.
A Machine Learning Framework for Programming by Example. In Proceedings of the 30th International Conference on International Conference on Machine Learning - Volume 28 (Atlanta, GA, USA) (ICML’13) . JMLR.org, I–187–I–195
Aditya Krishna Menon, Omer Tamuz, Sumit Gulwani, Butler Lampson, and Adam Tauman Kalai. 2013 · 2013
Earlier work this paper cites.
Detecting and Mitigating Secret-Key Leaks in Source Code Repositories. In 2015 IEEE/ACM 12th Working Conference on Mining Software Repositories . 396–400
Vibha Singhal Sinha, Diptikalyan Saha, Pankaj Dhoolia, Rohan Padhye, and Senthil Mani. 2015 · 2015
Cited alongside, same era.
Neuro-Symbolic Program Synthesis
Emilio Parisotto, Abdel rahman Mohamed, Rishabh Singh, Lihong Li, Dengyong Zhou, and Pushmeet Kohli. 2016 · 2016
Cited alongside, same era.
Program synthesis
Sumit Gulwani, Oleksandr Polozov, Rishabh Singh, et al · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
A syntactic neural model for general-purpose code generation
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba. 2021 · 2021
Later among the works it cites.
Solving Linear Algebra by Program Synthesis
Iddo Drori and Nakul Verma. 2021 · 2021
Later among the works it cites.
What do pre-trained code models know about code?. In 2021 36th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 1332–1336
Anjan Karmakar and Romain Robbes. 2021 · 2021
Later among the works it cites.
Can OpenAI Codex and Other Large Language Models Help Us Fix Security Bugs?
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pengcheng Yin and Graham Neubig. 2017 · 2017
Cited alongside, same era.
A Survey of Machine Learning for Big Code and Naturalness
Miltiadis Allamanis, Earl T. Barr, Premkumar Devanbu, and Charles Sutton. 2018 · 2018
Cited alongside, same era.
The adverse effects of code duplication in machine learning models of code. In Proceedings of the 2019 ACM SIGPLAN International Symposium on New Ideas, New Paradigms, and Reflections on Programming and Software . 143–153
Miltiadis Allamanis. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Establishing Benchmarks for Learning Program Representations.. In SATToSE
Anjan Karmakar. 2019 · 2019
Cited alongside, same era.
Unified Pre-training for Program Understanding and Generation
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2021 · 2021
Cited alongside, same era.
Program Synthesis with Large Language Models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, and Charles Sutton. 2021 · 2021
Cited alongside, same era.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?. In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency (Virtual Event, Canada) (FAccT ’21) . Association for Computing Machinery, New York, NY, USA, 610–623
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Cited alongside, same era.
Hammond Pearce, Benjamin Tan, Baleegh Ahmad, Ramesh Karri, and Brendan Dolan-Gavitt. 2021b · 2021
Later among the works it cites.
Automatic Program Repair with OpenAI’s Codex: Evaluating QuixBugs
Julian Aron Prenner and Romain Robbes. 2021 · 2021
Later among the works it cites.
Solving Probability and Statistics Problems by Program Synthesis
Leonard Tang, Elizabeth Ke, Nikhil Singh, Nakul Verma, and Iddo Drori. 2021 · 2021
Later among the works it cites.
Yue Wang, Weishi Wang, Shafiq Joty, and Steven C. H. Hoi. 2021 · 2021
Later among the works it cites.
A survey of machine learning on source code
Miltiadis Allamanis. 2022 · 2022
Closest in time.
Iddo Drori, Sunny Tran, Roman Wang, Newman Cheng, Kevin Liu, Leonard Tang, Elizabeth Ke, Nikhil Singh, Taylor L. Patti, Jayson Lynch, Avi Shporer, Nakul Verma, Eugene Wu, and Gilbert Strang. 2022 · 2022
Closest in time.
Competition-Level Code Generation with AlphaCode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, Thomas Hubert, Peter Choy, Cyprien de Masson d’Autume, Igor Babuschkin, Xinyun Chen, Po-Sen Huang, Johannes Welbl, Sven Gowal, Alexey Cherepanov, James Molloy, Daniel J. Mankowitz, Esme Sutherland Robson, Pushmeet Kohli, Nando de Freitas, Koray Kavukcuoglu, and Oriol Vinyals. 2022 · 2022
Closest in time.
Pop Quiz! Can a Large Language Model Help With Reverse Engineering?
Hammond Pearce, Benjamin Tan, Prashanth Krishnamurthy, Farshad Khorrami, Ramesh Karri, and Brendan Dolan-Gavitt. 2022 · 2022
Closest in time.
Synchromesh: Reliable code generation from pre-trained language models
Gabriel Poesia, Oleksandr Polozov, Vu Le, Ashish Tiwari, Gustavo Soares, Christopher Meek, and Sumit Gulwani. 2022 · 2022
Closest in time.