M. L. Siddiq, S. H. Majumder, M. R. Mim, S. Jajodia, and J. C. S. Santos, “An empirical study of code smells in transformer-based code generation techniques,” in 2022 IEEE 22nd International Working Conference on Source Code Analysis and Manipulation (SCAM) , 2022, pp. 71–82
2022
Later among the works it cites.
Z. Sun, X. Du, F. Song, M. Ni, and L. Li, “Coprotector: Protect open-source code against unauthorized training usage with data poisoning,” in Proceedings of the ACM Web Conference 2022 , ser. WWW ’22. New York, NY, USA: Association for Computing Machinery, 2022, p. 652–660
2022
Later among the works it cites.
F. F. Xu, U. Alon, G. Neubig, and V. J. Hellendoorn, “A systematic evaluation of large language models of code,” in Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming , ser. MAPS 2022. New York, NY, USA: Association for Computing Machinery, 2022, p. 1–10. [Online]. Available: https://doi.org/10.1145/3520312.3534862
2022
Later among the works it cites.
H. Ye, M. Martinez, and M. Monperrus, “Neural program repair with execution-based backpropagation,” in Proceedings of the 44th International Conference on Software Engineering , ser. ICSE ’22. New York, NY, USA: Association for Computing Machinery, 2022, p. 1506–1518. [Online]. Available: https://doi.org/10.1145/3510003.3510222
2022
Later among the works it cites.
X. Zheng and J. Jiang, “An empirical study of memorization in NLP,” in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Dublin, Ireland: Association for Computational Linguistics, May 2022, pp. 6265–6278. [Online]. Available: https://aclanthology.org/2022.acl-long.434
2022
Later among the works it cites.
H. Pearce, B. Ahmad, B. Tan, B. Dolan-Gavitt, and R. Karri, “Asleep at the keyboard? assessing the security of github copilot’s code contributions,” in 43rd IEEE Symposium on Security and Privacy, SP 2022, San Francisco, CA, USA, May 22-26, 2022 . IEEE, 2022, pp. 754–768. [Online]. Available: https://doi.org/10.1109/SP46214.2022.9833571
2022
Later among the works it cites.
Z. Zeng, H. Tan, H. Zhang, J. Li, Y. Zhang, and L. Zhang, “An extensive study on pre-trained models for program understanding and generation,” in Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis , ser. ISSTA 2022. New York, NY, USA: Association for Computing Machinery, 2022, p. 39–51. [Online]. Available: https://doi.org/10.1145/3533767.3534390
2022
Later among the works it cites.
J. He, B. Xu, Z. Yang, D. Han, C. Yang, and D. Lo, “Ptm4tag: Sharpening tag recommendation of stack overflow posts with pre-trained models,” ser. ICPC ’22. New York, NY, USA: Association for Computing Machinery, 2022, p. 1–11. [Online]. Available: https://doi.org/10.1145/3524610.3527897
2022
Later among the works it cites.
X. Tang, S. Mahloujifar, L. Song, V. Shejwalkar, M. Nasr, A. Houmansadr, and P. Mittal, “Mitigating membership inference attacks by { \{ Self-Distillation } \} through a novel ensemble architecture,” in 31st USENIX Security Symposium (USENIX Security 22) , 2022, pp. 1433–1450
2022
Later among the works it cites.
M. Wei, Y. Huang, J. Yang, J. Wang, and S. Wang, “Cocofuzzing: Testing neural code models with coverage-guided fuzzing,” IEEE Transactions on Reliability , pp. 1–14, 2022
2022
Later among the works it cites.
G. Ramakrishnan and A. Albarghouthi, “Backdoors in neural models of source code,” in 2022 26th International Conference on Pattern Recognition (ICPR) . Los Alamitos, CA, USA: IEEE Computer Society, aug 2022, pp. 2892–2899. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/ICPR56361.2022.9956690
2022
Later among the works it cites.
J. Li, Z. Li, H. Zhang, G. Li, Z. Jin, X. Hu, and X. Xia, “Poison attack and defense on deep source code processing models,” 2022. [Online]. Available: https://arxiv.org/abs/2210.17029
Original
2022
Later among the works it cites.
F. Mireshghallah, K. Goyal, A. Uniyal, T. Berg-Kirkpatrick, and R. Shokri, “Quantifying privacy risks of masked language models using membership inference attacks,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing . Abu Dhabi, United Arab Emirates: Association for Computational Linguistics, Dec. 2022, pp. 8332–8347. [Online]. Available: https://aclanthology.org/2022.emnlp-main.570
2022
Later among the works it cites.
X. Hou, Y. Zhao, Y. Liu, Z. Yang, K. Wang, L. Li, X. Luo, D. Lo, J. Grundy, and H. Wang, “Large language models for software engineering: A systematic literature review,” 2023
2023
Closest in time.
C. Yang, B. Xu, F. Thung, Y. Shi, T. Zhang, Z. Yang, X. Zhou, J. Shi, J. He, D. Han, and D. Lo, Answer Summarization for Technical Queries: Benchmark and New Approach . New York, NY, USA: Association for Computing Machinery, 2023. [Online]. Available: https://doi.org/10.1145/3551349.3560421
2023
Closest in time.
T.-D. Nguyen, Z. Yang, X. B. D. Le, Patanamon, Thongtanunam, and D. Lo, “Adversarial attacks on code models with discriminative graph patterns,” 2023
2023
Closest in time.
L. Niu, S. Mirza, Z. Maradni, and C. Pöpper, “CodexLeaks: Privacy leaks from code generation language models in GitHub copilot,” in 32nd USENIX Security Symposium (USENIX Security 23) . Anaheim, CA: USENIX Association, Aug. 2023, pp. 2133–2150
2023
Closest in time.
“Aws codewhisperer: Features,” https://aws.amazon.com/codewhisperer/features/ , accessed: March 29, 2023
2023
Closest in time.
S. K. Basak, L. Neil, B. Reaves, and L. Williams, “Secretbench: A dataset of software secrets,” in Proceedings of the 20th International Conference on Mining Software Repositories , ser. MSR ’23, 2023
2023
Closest in time.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” in The Eleventh International Conference on Learning Representations , 2023
2023
Closest in time.
J. He, Z. Xin, B. Xu, T. Zhang, K. Kim, Z. Yang, F. Thung, I. Irsan, and D. Lo, “Representation learning for stack overflow posts: How far are we?” 2023
2023
Closest in time.
K. Jesse, T. Ahmed, P. T. Devanbu, and E. Morgan, “Large language models and simple, stupid bugs,” in 2023 IEEE/ACM 20th International Conference on Mining Software Repositories (MSR) . Los Alamitos, CA, USA: IEEE Computer Society, may 2023, pp. 563–575. [Online]. Available: https://doi.ieeecomputersociety.org/10.1109/MSR59073.2023.00082
2023
Closest in time.
“A massively spiffy yet delicately unobtrusive compression library,” https://zlib.net/ , accessed on March 27, 2023
2023
Closest in time.
D. Cotroneo, C. Improta, P. Liguori, and R. Natella, “Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks,” 2023
2023
Closest in time.
T. T. sitter Contributors, “Tree-sitter: An A incremental parsing library,” https://tree-sitter.github.io/tree-sitter/ , Accessed on 25th March, 2023
2023
Closest in time.
Z. Yang, J. Shi, M. H. Asyrofi, B. Xu, X. Zhou, D. Han, and D. Lo, “Prioritizing speech test cases,” 2023. [Online]. Available: https://arxiv.org/abs/2302.00330
Original
2023
Closest in time.
Z. Yang, C. Wang, J. Shi, T. Hoang, P. Kochhar, Q. Lu, Z. Xing, and D. Lo, “What do users ask in open-source ai repositories? an empirical study of github issues,” in Proceedings of the 20th International Conference on Mining Software Repositories , ser. MSR ’23, 2023
2023
Closest in time.
A. Jha and C. K. Reddy, “Codeattack: code-based adversarial attacks for pre-trained programming language models,” in Proceedings of the Thirty-Seventh AAAI Conference on Artificial Intelligence and Thirty-Fifth Conference on Innovative Applications of Artificial Intelligence and Thirteenth Symposium on Educational Advances in Artificial Intelligence , ser. AAAI’23/IAAI’23/EAAI’23. AAAI Press, 2023. [Online]. Available: https://doi.org/10.1609/aaai.v37i12.26739
2023
Closest in time.
D. Dua and C. Graff, “Uci machine learning repository,” http://archive.ics.uci.edu/ml , 2017, accessed: 2023-03-25
2023
Closest in time.
Z. Yang, Z. Zhao, C. Wang, J. Shi, D. Kim, D. Han, and D. Lo, “Unveiling memorization in code models,” in Proceedings of the IEEE/ACM 46th International Conference on Software Engineering , ser. ICSE ’24. New York, NY, USA: Association for Computing Machinery, 2024. [Online]. Available: https://doi.org/10.1145/3597503.3639074
2024
Closest in time.