Fetching the paper…
Reading the bibliography…
Large language models (LLMs) for automatic code generation have achieved breakthroughs in several programming tasks.
B. Chess and G. McGraw, “Static analysis for security,” IEEE security & privacy , vol. 2, no. 6, pp. 76–79, 2004
2004
Earlier work this paper cites.
L. Yujian and L. Bo, “A normalized levenshtein distance metric,” TPAMI , 2007
2007
Earlier work this paper cites.
N. Ayewah, W. Pugh, D. Hovemeyer, J. D. Morgenthaler, and J. Penix, “Using static analysis to find bugs,” IEEE Software , vol. 25, no. 5, pp. 22–29, 2008
2008
Earlier work this paper cites.
G. Chatzieleftheriou and P. Katsaros, “Test-driving static analysis tools in search of c code vulnerabilities,” in 2011 IEEE 35th Annual Computer Software and Applications Conference Workshops , 2011, pp. 96–103
2011
Earlier work this paper cites.
L. Szekeres, M. Payer, T. Wei, and D. Song, “SoK: Eternal War in Memory,” in IEEE Symposium on Security and Privacy , 2013
2013
Earlier work this paper cites.
M. Fredrikson, S. Jha, and T. Ristenpart, “Model inversion attacks that exploit confidence information and basic countermeasures,” in ACM CCS , 2015
2015
Earlier work this paper cites.
A. Gosain and G. Sharma, “Static analysis: A survey of techniques and tools,” in Intelligent Computing and Applications . Springer, 2015, pp. 581–591
2015
Earlier work this paper cites.
K. Goseva-Popstojanova and A. Perhinschi, “On the capability of static code analysis to detect security vulnerabilities,” Information and Software Technology , vol. 68, pp. 18–33, 2015
2015
Earlier work this paper cites.
M. Beller, R. Bholanath, S. McIntosh, and A. Zaidman, “Analyzing the state of static analysis: A large-scale evaluation in open source software,” in 2016 IEEE 23rd International Conference on Software Analysis, Evolution, and Reengineering (SANER) , vol. 1, 2016, pp. 470–481
2016
Earlier work this paper cites.
M. Christakis and C. Bird, “What developers want and need from program analysis: An empirical study,” in Proceedings of the 31st IEEE/ACM International Conference on Automated Software Engineering , ser. ASE ’16. New York, NY, USA: Association for Computing Machinery, 2016, p. 332–343. [Online]. Available: https://doi.org/10.1145/2970276.2970347
2016
Earlier work this paper cites.
Y. Shoshitaishvili, R. Wang, C. Salls, N. Stephens, M. Polino, A. Dutcher, J. Grosen, S. Feng, C. Hauser, C. Kruegel, and G. Vigna, “SoK: (State of) The Art of War: Offensive Techniques in Binary Analysis,” in IEEE Symposium on Security and Privacy (SP) , 2016
2016
Earlier work this paper cites.
L. Wang, A. Schwing, and S. Lazebnik, “Diverse and accurate image description using a variational auto-encoder with an additive gaussian encoding space,” in NeurIPS , 2017
2017
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in NAACL , 2019
2019
Earlier work this paper cites.
A. Deshpande, J. Aneja, L. Wang, A. G. Schwing, and D. Forsyth, “Fast, diverse and accurate image captioning guided by part-of-speech,” in CVPR , 2019
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” in NeurIPS , 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, P. J. Liu et al. , “Exploring the limits of transfer learning with a unified text-to-text transformer.” JMLR , 2020
2020
Earlier work this paper cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “CodeBERT: A pre-trained model for programming and natural languages,” in EMNLP , 2020
2020
Earlier work this paper cites.
H. Yin, P. Molchanov, J. Alvarez, Z. Li, A. Mallya, D. Hoiem, N. Jha, and J. Kautz, “Dreaming to distill: Data-free knowledge transfer via deepinversion,” in CVPR , 2020
2020
Earlier work this paper cites.
Y. Nakamura, S. Hanaoka, Y. Nomura, N. Hayashi, O. Abe, S. Yada, S. Wakamiya, and E. Aramaki, “Kart: Privacy leakage framework of language models pre-trained with clinical records,” arXiv , 2020
2020
Earlier work this paper cites.
A. Fioraldi, D. Maier, H. Eißfeldt, and M. Heuse, “AFL++ : Combining Incremental Steps of Fuzzing Research ,” in USENIX Workshop on Offensive Technologies (WOOT) , 2020
2020
Earlier work this paper cites.
A. Holtzman, J. Buys, L. Du, M. Forbes, and Y. Choi, “The curious case of neural text degeneration,” in ICLR , 2020
2020
Earlier work this paper cites.
M. Chen, J. Tworek, H. Jun, Q. Yuan, H. Ponde, J. Kaplan, H. Edwards, Y. Burda, N. Joseph, G. Brockman, A. Ray, R. Puri, G. Krueger, M. Petrov, H. Khlaaf, G. Sastry, P. Mishkin, B. Chan, S. Gray, N. Ryder, M. Pavlov, A. Power, L. Kaiser, M. Bavarian, C. Winter, P. Tillet, F. P. Such, D. W. Cummings, M. Plappert, F. Chantzis, E. Barnes, A. Herbert-Voss, W. H. Guss, A. Nichol, I. Babuschkin, S. A. Balaji, S. Jain, A. Carr, J. Leike, J. Achiam, V. Misra, E. Morikawa, A. Radford, M. M. Knight, M. Brundage, M. Murati, K. Mayer, P. Welinder, B. McGrew, D. Amodei, S. McCandlish, I. Sutskever, and W. Zaremba, “Evaluating large language models trained on code,” arXiv , 2021
2021
Earlier work this paper cites.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi, “Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation,” in EMNLP , 2021
2021
Cited alongside, same era.
D. Guo, S. Ren, S. Lu, Z. Feng, D. Tang, S. Liu, L. Zhou, N. Duan, A. Svyatkovskiy, S. Fu, M. Tufano, S. K. Deng, C. B. Clement, D. Drain, N. Sundaresan, J. Yin, D. Jiang, and M. Zhou, “Graphcodebert: Pre-training code representations with data flow,” in ICLR , 2021
2021
Cited alongside, same era.
W. Ahmad, S. Chakraborty, B. Ray, and K.-W. Chang, “Unified pre-training for program understanding and generation,” in NAACL , 2021
2021
Cited alongside, same era.
A. Mahendran and A. Vedaldi, “Understanding deep image representations by inverting them,” in CVPR , 2021
2021
Cited alongside, same era.
K.-C. Wang, Y. FU, K. Li, A. Khisti, R. Zemel, and A. Makhzani, “Variational model inversion attacks,” in NeurIPS , 2021
Y. Lu, M. Bartolo, A. Moore, S. Riedel, and P. Stenetorp, “Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity,” in ACL , May 2022
2022
Later among the works it cites.
P. Bareiß, B. Souza, M. d’Amorim, and M. Pradel, “Code generation tools (almost) for free? a study of few-shot, pre-trained language models on code,” arXiv , 2022
2022
Later among the works it cites.
S. Lipp, S. Banescu, and A. Pretschner, “An empirical study on the effectiveness of static c code analyzers for vulnerability detection,” in Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis , 2022, pp. 544–555
2022
Later among the works it cites.
B. Rozière, J. Gehring, F. Gloeckle, S. Sootla, I. Gat, X. E. Tan, Y. Adi, J. Liu, T. Remez, J. Rapin et al. , “Code llama: Open foundation models for code,” arXiv , 2023
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson, A. Oprea, and C. Raffel, “Extracting Training Data from Large Language Models,” in USENIX Security Symposium , 2021
2021
Cited alongside, same era.
H. W. Chung, L. Hou, S. Longpre, B. Zoph, Y. Tay, W. Fedus, E. Li, X. Wang, M. Dehghani, S. Brahma et al. , “Scaling instruction-finetuned language models,” arXiv , 2022
2022
Cited alongside, same era.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” arXiv , 2022
2022
Cited alongside, same era.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” arXiv , 2022
2022
Cited alongside, same era.
D. Fried, A. Aghajanyan, J. Lin, S. Wang, E. Wallace, F. Shi, R. Zhong, W.-t. Yih, L. Zettlemoyer, and M. Lewis, “Incoder: A generative model for code infilling and synthesis,” arXiv , 2022
2022
Cited alongside, same era.
Y. Li, D. Choi, J. Chung, N. Kushman, J. Schrittwieser, R. Leblond, T. Eccles, J. Keeling, F. Gimeno, A. D. Lago, T. Hubert, P. Choy, C. de Masson d’Autume, I. Babuschkin, X. Chen, P.-S. Huang, J. Welbl, S. Gowal, A. Cherepanov, J. Molloy, D. J. Mankowitz, E. S. Robson, P. Kohli, N. de Freitas, K. Kavukcuoglu, and O. Vinyals, “Competition-level code generation with alphacode,” Science , vol. 378, no. 6624, pp. 1092–1097, 2022. [Online]. Available: https://www.science.org/doi/abs/10.1126/science.abq1158
2022
Cited alongside, same era.
S. Imai, “Is github copilot a substitute for human pair-programming? an empirical study,” in Proceedings of the ACM/IEEE 44th International Conference on Software Engineering: Companion Proceedings , 2022, pp. 319–321
2022
Cited alongside, same era.
R. Li, L. B. Allal, Y. Zi, N. Muennighoff, D. Kocetkov, C. Mou, M. Marone, C. Akiki, J. Li, J. Chim et al. , “Starcoder: may the source be with you!” arXiv , 2023
2023
Closest in time.
OpenAI, “Gpt-4 technical report,” 2023
2023
Closest in time.
Z. Luo, C. Xu, P. Zhao, Q. Sun, X. Geng, W. Hu, C. Tao, J. Ma, Q. Lin, and D. Jiang, “Wizardcoder: Empowering code large language models with evol-instruct,” arXiv , 2023
2023
Closest in time.
J. Liu, C. S. Xia, Y. Wang, and L. Zhang, “Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation,” arXiv , 2023
2023
Closest in time.
J. He and M. Vechev, “Large language models for code: Security hardening and adversarial testing,” Workshop on Challenges in Deployable Generative AI at International Conference on Machine Learning (ICML) , 2023
2023
Closest in time.
H. Touvron, L. Martin, K. Stone, P. Albert, A. Almahairi, Y. Babaei, N. Bashlykov, S. Batra, P. Bhargava, S. Bhosale et al. , “Llama 2: Open foundation and fine-tuned chat models,” arXiv , 2023
2023
Closest in time.
C. Xu, Q. Sun, K. Zheng, X. Geng, P. Zhao, J. Feng, C. Tao, and D. Jiang, “Wizardlm: Empowering large language models to follow complex instructions,” arXiv , 2023
2023
Closest in time.
OpenAI, “Chatgpt: Optimizing language models for dialogue,” Nov. 2022, https://openai.com/blog/chatgpt/ , as of August 24, 2026
2026
Closest in time.
T. Dohmke, “Github copilot is generally available to all developers,” Jun. 2022, https://github.blog/2022-06-21-github-copilot-is-generally-available-to-all-developers/ , as of August 24, 2026
2026
Closest in time.
S. Zhao, “Github copilot is generally available for businesses,” Dec. 2022, https://github.blog/2022-12-07-github-copilot-is-generally-available-for-businesses/ , as of August 24, 2026
2026
Closest in time.
MITRE, “CWE - Common Weakness Enumeration,” 2022, https://cwe.mitre.org , as of August 24, 2026
2026
Closest in time.
G. Inc, “Github codeql,” 2022, https://codeql.github.com/ , as of August 24, 2026
2026
Closest in time.
N. C. for Assured Software, “Juliet C/C++ 1.3,” Oct. 2017, https://samate.nist.gov/SARD/test-suites/112 , as of August 24, 2026
2026
Closest in time.
OpenAI, “OpenAI API Documentation,” 2022, https://beta.openai.com/docs/introduction , as of August 24, 2026
2026
Closest in time.
HuggingFace, “Big code models leaderboard,” Oct. 2023, https://huggingface.co/spaces/bigcode/bigcode-models-leaderboard , as of August 24, 2026
2026
Closest in time.
P. Thakkar, “Copilot internals,” 2022, https://thakkarparth007.github.io/copilot-explorer/posts/copilot-internals , as of August 24, 2026
2026
Closest in time.
SeatGeek, “Thefuzz,” 2022, https://github.com/seatgeek/thefuzz , as of August 24, 2026
2026
Closest in time.