Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have seen widespread success in code generation tasks for different scenarios, both everyday and professional.
SEI CERT, “Sei CERT c coding standard,” https://wiki.sei.cmu.edu/confluence/display/c/SEI+CERT+C+Coding+Standard, Pittsburgh, PA, 2016, version 3.1 (Latest stable release)
2016
Earlier work this paper cites.
J. Li, B. Zhao, and C. Zhang, “Fuzzing: a survey,” Cybersecurity , vol. 1, pp. 1–13, 2018
2018
Earlier work this paper cites.
2020
Earlier work this paper cites.
M. Chen, J. Tworek, H. Jun, Q. Yuan, H. P. de Oliveira Pinto, J. Kaplan, H. Edwards, Y. Burda, N. Joseph, G. Brockman, A. Ray, R. Puri, G. Krueger, M. Petrov, H. Khlaaf, G. Sastry, P. Mishkin, B. Chan, S. Gray, N. Ryder, M. Pavlov, A. Power, L. Kaiser, M. Bavarian, C. Winter, P. Tillet, F. P. Such, D. Cummings, M. Plappert, F. Chantzis, E. Barnes, A. Herbert-Voss, W. H. Guss, A. Nichol, A. Paino, N. Tezak, J. Tang, I. Babuschkin, S. Balaji, S. Jain, W. Saunders, C. Hesse, A. N. Carr, J. Leike, J. Achiam, V. Misra, E. Morikawa, A. Radford, M. Knight, M. Brundage, M. Murati, K. Mayer, P. Welinder, B. McGrew, D. Amodei, S. McCandlish, I. Sutskever, and W. Zaremba, “Evaluating large language models trained on code,” 2021
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
H. Pearce, B. Ahmad, B. Tan, B. Dolan-Gavitt, and R. Karri, “Asleep at the keyboard? assessing the security of github copilot’s code contributions,” in 2022 IEEE Symposium on Security and Privacy (SP) , 2022, pp. 754–768
2022
Earlier work this paper cites.
M. L. Siddiq and J. C. S. Santos, “Securityeval dataset: mining vulnerability examples to evaluate machine learning-based code generation techniques,” ser. MSR4P&S 2022. New York, NY, USA: Association for Computing Machinery, 2022, p. 29–33. [Online]. Available: https://doi.org/10.1145/3549035.3561184
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
N. Tihanyi, T. Bisztray, R. Jain, M. A. Ferrag, L. C. Cordeiro, and V. Mavroeidis, “The formai dataset: Generative ai in software security through the lens of formal verification,” in Proceedings of the 19th International Conference on Predictive Models and Data Analytics in Software Engineering , ser. PROMISE 2023. New York, NY, USA: Association for Computing Machinery, 2023, p. 33–43. [Online]. Available: https://doi.org/10.1145/3617555.3617874
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
S. Chaudhary, “Code alpaca: An instruction-following llama model for code generation,” https://github.com/sahil280114/codealpaca, 2023
2023
Earlier work this paper cites.
J. He and M. Vechev, “Large language models for code: Security hardening and adversarial testing,” in Proceedings of the 2023 ACM SIGSAC Conference on Computer and Communications Security , ser. CCS ’23. New York, NY, USA: Association for Computing Machinery, 2023, p. 1865–1879. [Online]. Available: https://doi.org/10.1145/3576915.3623175
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
T. Schick, J. Dwivedi-Yu, R. Dessí, R. Raileanu, M. Lomeli, E. Hambro, L. Zettlemoyer, N. Cancedda, and T. Scialom, “Toolformer: language models can teach themselves to use tools,” in Proceedings of the 37th International Conference on Neural Information Processing Systems , ser. NIPS ’23. Red Hook, NY, USA: Curran Associates Inc., 2023
2023
Earlier work this paper cites.
OpenAI, “Openai o1 system card,” 2024. [Online]. Available: https://arxiv.org/abs/2412.16720
2024
Earlier work this paper cites.
S. Pichai, “Q3 earnings call: CEO’s remarks,” https://blog.google/inside-google/message-ceo/alphabet-earnings-q3-2024/
2024
Cited alongside, same era.
V. Majdinasab, M. J. Bishop, S. Rasheed, A. Moradidakhel, A. Tahir, and F. Khomh, “ Assessing the Security of GitHub Copilot’s Generated Code - A Targeted Replication Study ,” in 2024 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . Los Alamitos, CA, USA: IEEE Computer Society, Mar. 2024, pp. 435–444. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/SANER60148.2024.00051
2024
Cited alongside, same era.
2024
Cited alongside, same era.
J. Yang, C. E. Jimenez, A. Wettig, K. Lieret, S. Yao, K. Narasimhan, and O. Press, “Swe-agent: Agent-computer interfaces enable automated software engineering,” in Advances in Neural Information Processing Systems , A. Globerson, L. Mackey, D. Belgrave, A. Fan, U. Paquet, J. Tomczak, and C. Zhang, Eds., vol. 37. Curran Associates, Inc., 2024, pp. 50 528–50 652. [Online]. Available: https://proceedings.neurips.cc/paper_ files/paper/2024/file/5a7c947568c1b1328ccc5230172e1e7c-Paper-Conference.pdf
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2024
Cited alongside, same era.
J. He, M. Vero, G. Krasnopolska, and M. Vechev, “Instruction tuning for secure code generation,” in Proceedings of the 41st International Conference on Machine Learning , ser. ICML’24. JMLR.org, 2024
2024
Cited alongside, same era.
Aider. (2024) Aider llm leaderboards. [Online]. Available: https://aider.chat/docs/leaderboards/
2024
Cited alongside, same era.
C. E. Jimenez, J. Yang, A. Wettig, S. Yao, K. Pei, O. Press, and K. R. Narasimhan, “SWE-bench: Can language models resolve real-world github issues?” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=VTF8yNQM66
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Y. Wei, Z. Wang, J. Liu, Y. Ding, and L. Zhang, “Magicoder: Empowering code generation with OSS-instruct,” in Proceedings of the 41st International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, vol. 235. PMLR, 21–27 Jul 2024, pp. 52 632–52 657. [Online]. Available: https://proceedings.mlr.press/v235/wei24h.html
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Y. Chen, Z. Hu, C. Zhi, J. Han, S. Deng, and J. Yin, “Chatunitest: A framework for llm-based test generation,” in Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering , ser. FSE 2024. New York, NY, USA: Association for Computing Machinery, 2024, p. 572–576. [Online]. Available: https://doi.org/10.1145/3663529.3663801
2024
Later among the works it cites.
Anthropic, “Claude 3.7 sonnet and claude code,” https://www.anthropic.com/news/claude-3-7-sonnet, 2025, accessed: 2025-04-23
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
GitHub, “Github copilot features,” https://github.com/features/copilot, 2025, accessed: 2025-04-23
2025
Closest in time.
E. Kalliamvakou, “Quantifying github copilot’s impact on developer productivity and happiness,” https://github.blog/news-insights/research/research-quantifying-github-copilots-impact-on-developer-productivity-and-happiness/, 2022, accessed: 2025-04-23
2025
Closest in time.
Y. Wang, Y. Wang, D. Guo, J. Chen, R. Zhang, Y. Ma, and Z. Zheng, “ RLCoder: Reinforcement Learning for Repository-Level Code Completion ,” in 2025 IEEE/ACM 47th International Conference on Software Engineering (ICSE) . Los Alamitos, CA, USA: IEEE Computer Society, May 2025, pp. 165–177. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/ICSE55347.2025.00014
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
GitHub, “Codeql,” https://codeql.github.com, 2025, accessed: 2025-04-23
2025
Closest in time.
PyCQA, “Bandit: A security linter from pycqa,” https://github.com/PyCQA/bandit?tab=readme-ov-file, 2025, accessed: 2025-04-23
2025
Closest in time.
MITRE Corporation, “Common weakness enumeration (cwe),” Online, 2025, accessed: 26 April, 2025. [Online]. Available: https://cwe.mitre.org/
2025
Closest in time.