Fetching the paper…
Reading the bibliography…
Recent advancements have led to the widespread adoption of code-oriented large language models (Code LLMs) for programming tasks.
L. Van der Maaten and G. Hinton, “Visualizing data using t-sne.” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
Z. Li, D. Zou, S. Xu, X. Ou, H. Jin, S. Wang, Z. Deng, and Y. Zhong, “Vuldeepecker: A deep learning-based system for vulnerability detection,” in NDSS , 2018
2018
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” in NeurIPS , 2020
2020
Earlier work this paper cites.
M. Ohm, H. Plate, A. Sykosch, and M. Meier, “Backstabber’s knife collection: A review of open source software supply chain attacks,” in DIMVA , 2020
2020
Earlier work this paper cites.
M. Javaheripi, M. Samragh, B. D. Rouhani, T. Javidi, and F. Koushanfar, “Curtail: Characterizing and thwarting adversarial deep learning,” IEEE Transactions on Dependable and Secure Computing , vol. 18, no. 2, pp. 736–752, 2020
2020
Earlier work this paper cites.
P. Bielik and M. Vechev, “Adversarial robustness for code,” in ICML , 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang et al. , “Codebert: A pre-trained model for programming and natural languages,” in EMNLP , 2020
2020
Earlier work this paper cites.
N. Friedman, “Introducing github copilot: your ai pair programmer,” https://docs.github.com/zh/copilot/quickstart , 2021
2021
Earlier work this paper cites.
O. Wojciech Zaremba, Greg Brockman, “Openai codex,” https://openai.com/blog/openai-codex , 2021
2021
Earlier work this paper cites.
R. Schuster, C. Song, E. Tromer, and V. Shmatikov, “You autocomplete me: Poisoning vulnerabilities in neural code completion,” in USENIX Security , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi, “Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation,” in EMNLP , 2021
2021
Earlier work this paper cites.
Y. Zhou, X. Zhang, J. Shen, T. Han, T. Chen, and H. Gall, “Adversarial robustness of deep code comment generation,” ACM Transactions on Software Engineering and Methodology , vol. 31, no. 4, pp. 1–30, 2022
2022
Earlier work this paper cites.
Y. He, Y. Liu, L. Wu, Z. Yang, K. Ren, and Z. Qin, “Msdroid: Identifying malicious snippets for android malware detection,” IEEE Transactions on Dependable and Secure Computing , vol. 20, no. 3, pp. 2025–2039, 2022
2022
Earlier work this paper cites.
W. Sun, C. Fang, Y. Chen, G. Tao, T. Han, and Q. Zhang, “Code search based on context-aware code translation,” in ICSE , 2022
2022
Earlier work this paper cites.
W. Jiang, T. Zhang, H. Qiu, H. Li, and G. Xu, “Incremental learning, incremental backdoor threats,” IEEE Transactions on Dependable and Secure Computing , vol. 21, no. 2, pp. 559–572, 2022
2022
Earlier work this paper cites.
Y. Li, Y. Bai, Y. Jiang, Y. Yang, S.-T. Xia, and B. Li, “Untargeted backdoor watermark: Towards harmless and stealthy dataset copyright protection,” in NeurIPS , 2022
2022
Earlier work this paper cites.
S. Lu, D. Guo, S. Ren, J. Huang, A. Svyatkovskiy, A. Blanco, C. Clement, D. Drain, D. Jiang, D. Tang et al. , “Codexglue: A machine learning benchmark dataset for code understanding and generation,” in NeurIPS , 2022
2022
Earlier work this paper cites.
Q. Zheng, X. Xia, X. Zou, Y. Dong, S. Wang, Y. Xue et al. , “Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x,” in SIGKDD , 2023
2023
Cited alongside, same era.
F. Zhang, B. Chen, Y. Zhang, J. Keung, J. Liu, D. Zan, Y. Mao, J.-G. Lou, and W. Chen, “Repocoder: Repository-level code completion through iterative retrieval and generation,” in EMNLP , 2023
2023
Cited alongside, same era.
Y. Li, S. Liu, K. Chen, X. Xie, T. Zhang, and Y. Liu, “Multi-target backdoor attacks for code pre-trained models,” in ACL , 2023
2023
Cited alongside, same era.
A. Jha and C. K. Reddy, “Codeattack: Code-based adversarial attacks for pre-trained programming language models,” in AAAI , 2023
2023
Cited alongside, same era.
K. Greshake, S. Abdelnabi, S. Mishra, C. Endres, T. Holz, and M. Fritz, “Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection,” in CCS , 2023
W. Ma, Y. Song, M. Xue, S. Wen, and Y. Xiang, “The “code” of ethics: A holistic audit of ai code generators,” IEEE Transactions on Dependable and Secure Computing , vol. 21, no. 5, pp. 4997–5031, 2024
2024
Closest in time.
X. Shen, Z. Chen, M. Backes, Y. Shen, and Y. Zhang, “” do anything now”: Characterizing and evaluating in-the-wild jailbreak prompts on large language models,” in CCS , 2024
2024
Closest in time.
Q. Zhan, Z. Liang, Z. Ying, and D. Kang, “Injecagent: Benchmarking indirect prompt injections in tool-integrated large language model agents,” in ACL , 2024
2024
Closest in time.
J. Chen, X. Hu, Z. Li, C. Gao, X. Xia, and D. Lo, “Code search is all you need? improving code suggestions with code search,” in ICSE , 2024
2024
Closest in time.
“2023 cwe top 10 known exploited vulnerabilities weaknesses,” https://cwe.mitre.org/top25/archive/2023/2023_kev_list.html , 2024, accessed: October 30, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
A. Alhamdan and C.-A. Staicu, “Sanddriller: A fully-automated approach for testing language-based javascript sandboxes,” in USENIX Security , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
C. S. Xia, Y. Wei, and L. Zhang, “Automated program repair in the era of large pre-trained language models,” in ICSE , 2023
2023
Cited alongside, same era.
Y. Li, M. Zhu, X. Yang, Y. Jiang, T. Wei, and S.-T. Xia, “Black-box dataset ownership verification via backdoor watermarking,” IEEE Transactions on Information Forensics and Security , vol. 18, pp. 2318–2332, 2023
2023
Cited alongside, same era.
B. He, J. Liu, Y. Li, S. Liang, J. Li, X. Jia, and X. Cao, “Generating transferable 3d adversarial point cloud via random perturbation factorization,” in AAAI , 2023
2023
Cited alongside, same era.
H. Peng, S. Guo, D. Zhao, X. Zhang, J. Han, S. Ji, X. Yang, and M. Zhong, “Textcheater: A query-efficient textual adversarial attack in the hard-label setting,” IEEE Transactions on Dependable and Secure Computing , vol. 21, no. 4, pp. 3901–3916, 2023
2023
Cited alongside, same era.
2024
Closest in time.
“Visual studio code,” https://code.visualstudio.com/ , Microsoft Corporation, 2024, accessed: May 9, 2024
2024
Closest in time.
W. Ma, Y. Song, M. Xue, S. Wen, and Y. Xiang, “The “code” of ethics: A holistic audit of ai code generators,” IEEE Transactions on Dependable and Secure Computing , vol. 21, no. 5, pp. 4997–5013, 2024
2024
Closest in time.
J. Zhang, J. Chi, Z. Li, K. Cai, Y. Zhang, and Y. Tian, “Badmerging: Backdoor attacks against model merging,” in CCS , 2024
2024
Closest in time.
M. Ya, Y. Li, T. Dai, B. Wang, Y. Jiang, and S.-T. Xia, “Towards faithful xai evaluation via generalization-limited backdoor watermark,” in ICLR , 2024
2024
Closest in time.
W. Jiang, H. Li, G. Xu, H. Ren, H. Yang, T. Zhang, and S. Yu, “Rethinking the design of backdoor triggers and adversarial perturbations: A color space perspective,” IEEE Transactions on Dependable and Secure Computing , pp. 1–18, 2024
2024
Closest in time.
2024
Closest in time.
R. Zhang, H. Li, R. Wen, W. Jiang, Y. Zhang, M. Backes, Y. Shen, and Y. Zhang, “Instruction backdoor attacks against customized { \{ LLMs } \} ,” in USENIX Security , 2024
2024
Closest in time.
M. Shen, C. Li, Q. Li, H. Lu, L. Zhu, and K. Xu, “Transferability of white-box perturbations: query-efficient adversarial attacks against commercial dnn services,” in USENIX Security , 2024
2024
Closest in time.
J. Liu, C. S. Xia, Y. Wang, and L. Zhang, “Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation,” in NeurIPS , 2024
2024
Closest in time.
Y. He, H. She, X. Qian, X. Zheng, Z. Chen, Z. Qin, and L. Cavallaro, “On benchmarking code llms for android malware analysis,” in ISSTA , 2025
2025
Closest in time.
“Criminal cases involving large models acting as intermediaries to generate malicious code,” https://x.com/r_cky0/status/1859656430888026524 , 2025, accessed: January 6, 2025
2025
Closest in time.
B. Yi, T. Huang, S. Chen, T. Li, Z. Liu, C. Zhixuan, and Y. Li, “Probe before you talk: Towards black-box defense against backdoor unalignment for large language models,” in ICLR , 2025
2025
Closest in time.
2025
Closest in time.