Fetching the paper…
Reading the bibliography…
Language models for code (CodeLMs) have emerged as powerful tools for code-related tasks, outperforming traditional methods and standard machine learning approaches.
RoPGen: Towards Robust Code Authorship Attribution via Automatic Coding Style Transformation. In Proceedings of the 44th IEEE/ACM International Conference on Software Engineering . ACM, Pittsburgh, PA, USA, 1906–1918
Zhen Li, Qian (Guenevere) Chen, Chen Chen, Yayi Zou, and Shouhuai Xu. 2022a · 1918
Earlier work this paper cites.
Equation of state calculations by fast computing machines
Nicholas Metropolis, Arianna W Rosenbluth, Marshall N Rosenbluth, Augusta H Teller, and Edward Teller. 1953 · 1953
Earlier work this paper cites.
Combining Graph-Based Learning With Automated Data Collection for Code Vulnerability Detection
Huanting Wang, Guixin Ye, Zhanyong Tang, Shin Hwei Tan, Songfang Huang, Dingyi Fang, Yansong Feng, Lizhong Bian, and Zheng Wang. 2021b · 1958
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Qualitative Methods in Empirical Studies of Software Engineering
Carolyn B. Seaman. 1999 · 1999
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation. In Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics . ACL, Philadelphia, PA, USA, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Guidelines for performing systematic literature reviews in software engineering
Keele Staffs et al · 2007
Earlier work this paper cites.
How Reliable Are Systematic Reviews in Empirical Software Engineering?
Stephen G. MacDonell, Martin J. Shepperd, Barbara A. Kitchenham, and Emilia Mendes. 2010 · 2010
Earlier work this paper cites.
Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition
George E Dahl, Dong Yu, Li Deng, and Alex Acero. 2011 · 2011
Earlier work this paper cites.
Repeatability of systematic literature reviews. In Proceedings of the 15th International Conference on Evaluation and Assessment in Software Engineering . IET - The Institute of Engineering and Technology / IEEE Xplore, Durham, UK, 46–55
Barbara A. Kitchenham, Pearl Brereton, Zhi Li, David Budgen, and Andrew James Burn. 2011 · 2011
Earlier work this paper cites.
Identifying relevant studies in software engineering
He Zhang, Muhammad Ali Babar, and Paolo Tell. 2011 · 2011
Earlier work this paper cites.
Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al · 2012
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks. In Proceedings of the 26th Annual Conference on Neural Information Processing Systems . OpenReview.net, Lake Tahoe, Nevada, United States, 1106–1114
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012 · 2012
Earlier work this paper cites.
Malicious PDF detection using metadata and structural features. In Proceedings of the 28th Annual Computer Security Applications Conference . ACM, Orlando, FL, USA, 239–248
Charles Smutz and Angelos Stavrou. 2012 · 2012
Earlier work this paper cites.
DREBIN: Effective and Explainable Detection of Android Malware in Your Pocket. In Proceedings of the 21st Annual Network and Distributed System Security Symposium . The Internet Society, San Diego, California, USA, 1–16
Daniel Arp, Michael Spreitzenbarth, Malte Hubner, Hugo Gascon, and Konrad Rieck. 2014 · 2014
Earlier work this paper cites.
On the Properties of Neural Machine Translation: Encoder-Decoder Approaches. In Proceedings of the 8th Workshop on Syntax, Semantics and Structure in Statistical Translation . Association for Computational Linguistics, Doha, Qatar, 103–111
Kyunghyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Convolutional Neural Networks for Sentence Classification. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing . ACL, Doha, Qatar, 1746–1751
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Experiments in program synthesis with grammatical evolution: A focus on Integer Sorting. In Proceedings of the IEEE Congress on Evolutionary Computation . IEEE, Beijing, China, 1504–1511
Michael O’Neill, Miguel Nicolau, and Alexandros Agapitos. 2014 · 2014
Earlier work this paper cites.
Towards a Big Data Curated Benchmark of Inter-project Code Clones. In Proceedings of the 30th IEEE International Conference on Software Maintenance and Evolution . IEEE Computer Society, Victoria, BC, Canada, 476–480
Jeffrey Svajlenko, Judith F. Islam, Iman Keivanloo, Chanchal Kumar Roy, and Mohammad Mamun Mia. 2014 · 2014
Earlier work this paper cites.
Intriguing properties of neural networks. In Proceedings of the 2nd International Conference on Learning Representations, Conference Track Proceedings . OpenReview.net, Banff, AB, Canada, 1–10
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian J. Goodfellow, and Rob Fergus. 2014 · 2014
Earlier work this paper cites.
Guidelines for snowballing in systematic literature studies and a replication in software engineering. In Proceedings of the 18th International Conference on Evaluation and Assessment in Software Engineering . ACM, London, England, United Kingdom, 38:1–38:10
Claes Wohlin. 2014 · 2014
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate. In Proceedings of the 3rd International Conference on Learning Representations . OpenReview.net, San Diego, CA, USA, 1–15
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
De-anonymizing Programmers via Code Stylometry. In Proceedings of the 24th USENIX Security Symposium . USENIX Association, Washington, D.C., USA, 255–270
Aylin Caliskan Islam, Richard E. Harang, Andrew Liu, Arvind Narayanan, Clare R. Voss, Fabian Yamaguchi, and Rachel Greenstadt. 2015 · 2015
Earlier work this paper cites.
Obfuscator-LLVM - Software Protection for the Masses. In Proceedings of the 1st IEEE/ACM International Workshop on Software Protection . IEEE Computer Society, Florence, Italy, 3–9
Pascal Junod, Julien Rinaldini, Johan Wehrli, and Julie Michielin. 2015 · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition . IEEE Computer Society, Las Vegas, NV, USA, 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
On the naturalness of software
Abram Hindle, Earl T Barr, Mark Gabel, Zhendong Su, and Premkumar Devanbu. 2016 · 2016
Earlier work this paper cites.
Convolutional Neural Networks over Tree Structures for Programming Language Processing. In Proceedings of the 30th AAAI Conference on Artificial Intelligence . AAAI Press, Phoenix, Arizona, USA, 1287–1293
Lili Mou, Ge Li, Lu Zhang, Tao Wang, and Zhi Jin. 2016 · 2016
Earlier work this paper cites.
Stealing Machine Learning Models via Prediction APIs. In Proceedings of the 25th USENIX Security Symposium . USENIX Association, Austin, TX, USA, 601–618
Florian Tramèr, Fan Zhang, Ari Juels, Michael K. Reiter, and Thomas Ristenpart. 2016 · 2016
Earlier work this paper cites.
Towards Evaluating the Robustness of Neural Networks. In Proceedings of the 2017 IEEE Symposium on Security and Privacy . IEEE Computer Society, San Jose, CA, USA, 39–57
Nicholas Carlini and David A. Wagner. 2017 · 2017
Earlier work this paper cites.
BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain
Tianyu Gu, Brendan Dolan-Gavitt, and Siddharth Garg. 2017 · 2017
Earlier work this paper cites.
Automatically generating commit messages from diffs using neural machine translation. In Proceedings of the 32nd IEEE/ACM International Conference on Automated Software Engineering . IEEE Computer Society, Urbana, IL, USA, 135–146
Siyuan Jiang, Ameer Armaly, and Collin McMillan. 2017 · 2017
Earlier work this paper cites.
A Survey of App Store Analysis for Software Engineering
William J. Martin, Federica Sarro, Yue Jia, Yuanyuan Zhang, and Mark Harman. 2017 · 2017
Earlier work this paper cites.
Practical Black-Box Attacks against Machine Learning. In Proceedings of the 2017 ACM on Asia Conference on Computer and Communications Security . ACM, Abu Dhabi, United Arab Emirates, 506–519
Nicolas Papernot, Patrick D. McDaniel, Ian J. Goodfellow, Somesh Jha, Z. Berkay Celik, and Ananthram Swami. 2017 · 2017
Earlier work this paper cites.
Machine Learning Models that Remember Too Much. In Proceedings of the 2017 ACM Conference on Computer and Communications Security . ACM, Dallas, TX, USA, 587–601
Congzheng Song, Thomas Ristenpart, and Vitaly Shmatikov. 2017 · 2017
Earlier work this paper cites.
Axiomatic Attribution for Deep Networks. In Proceedings of the 34th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 70) . PMLR, Sydney, NSW, Australia, 3319–3328
Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017 · 2017
Earlier work this paper cites.
Neural Network-based Graph Embedding for Cross-Platform Binary Code Similarity Detection. In Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security . ACM, Dallas, TX, USA, 363–376
Xiaojun Xu, Chang Liu, Qian Feng, Heng Yin, Le Song, and Dawn Song. 2017 · 2017
Earlier work this paper cites.
Large-Scale and Language-Oblivious Code Authorship Identification. In Proceedings of the 2018 ACM Conference on Computer and Communications Security . ACM, Toronto, ON, Canada, 101–114
Mohammed Abuhamad, Tamer AbuHmed, Aziz Mohaisen, and DaeHun Nyang. 2018 · 2018
Earlier work this paper cites.
Generating Natural Language Adversarial Examples. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Brussels, Belgium, 2890–2896
Moustafa Alzantot, Yash Sharma, Ahmed Elgohary, Bo-Jhang Ho, Mani B. Srivastava, and Kai-Wei Chang. 2018 · 2018
Earlier work this paper cites.
EMBER: An Open Dataset for Training Static PE Malware Machine Learning Models
Hyrum S. Anderson and Phil Roth. 2018 · 2018
Earlier work this paper cites.
Deep code search. In Proceedings of the 40th International Conference on Software Engineering . ACM, Gothenburg, Sweden, 933–944
Xiaodong Gu, Hongyu Zhang, and Sunghun Kim. 2018 · 2018
Earlier work this paper cites.
Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression Learning. In Proceedings of the 2018 IEEE Symposium on Security and Privacy . IEEE Computer Society, San Francisco, California, USA, 19–35
Matthew Jagielski, Alina Oprea, Battista Biggio, Chang Liu, Cristina Nita-Rotaru, and Bo Li. 2018 · 2018
Earlier work this paper cites.
Trojaning Attack on Neural Networks. In Proceedings of the 25th Annual Network and Distributed System Security Symposium . The Internet Society, San Diego, California, USA, 1–15
Yingqi Liu, Shiqing Ma, Yousra Aafer, Wen-Chuan Lee, Juan Zhai, Weihang Wang, and Xiangyu Zhang. 2018 · 2018
Earlier work this paper cites.
Towards Deep Learning Models Resistant to Adversarial Attacks. In Proceedings of the 6th International Conference on Learning Representations . OpenReview.net, Vancouver, BC, Canada, 1–28
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt, Dimitris Tsipras, and Adrian Vladu. 2018 · 2018
Earlier work this paper cites.
Adversarial Binaries for Authorship Identification
Xiaozhu Meng, Barton P. Miller, and Somesh Jha. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
Retrieval on source code: a neural code search. In Proceedings of the 2nd ACM SIGPLAN International Workshop on Machine Learning and Programming Languages . ACM, Philadelphia, PA, USA, 31–41
Saksham Sachdev, Hongyu Li, Sifei Luan, Seohyun Kim, Koushik Sen, and Satish Chandra. 2018 · 2018
Earlier work this paper cites.
Spectral Signatures in Backdoor Attacks. In Proceedings of the 2018 32nd Conference on Neural Information Processing Systems . OpenReview.net, Montréal, Canada, 8011–8021
Brandon Tran, Jerry Li, and Aleksander Madry. 2018 · 2018
Earlier work this paper cites.
code2vec: learning distributed representations of code
Uri Alon, Meital Zilberstein, Omer Levy, and Eran Yahav. 2019b · 2019
Earlier work this paper cites.
Generative Code Modeling with Graphs. In Proceedings of the 7th International Conference on Learning Representations . OpenReview.net, New Orleans, LA, USA, 1–24
Marc Brockschmidt, Miltiadis Allamanis, Alexander L. Gaunt, and Oleksandr Polozov. 2019 · 2019
Earlier work this paper cites.
Deep Integration: A Multi-Label Architecture for Road Scene Recognition
Long Chen, Wujing Zhan, Wei Tian, Yuhang He, and Qin Zou. 2019d · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Minneapolis, MN, USA, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
CodeSearchNet Challenge: Evaluating the State of Semantic Code Search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Earlier work this paper cites.
Maybe deep neural networks are the best choice for modeling source code
Rafael-Michael Karampatsis and Charles Sutton. 2019 · 2019
Earlier work this paper cites.
A neural model for generating natural language summaries of program subroutines. In Proceedings of the 41st International Conference on Software Engineering . IEEE / ACM, Montreal, QC, Canada, 795–806
Alexander LeClair, Siyuan Jiang, and Collin McMillan. 2019 · 2019
Earlier work this paper cites.
Graph Matching Networks for Learning the Similarity of Graph Structured Objects. In Proceedings of the 36th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 97) . PMLR, Long Beach, California, USA, 3835–3845
Yujia Li, Chenjie Gu, Thomas Dullien, Oriol Vinyals, and Pushmeet Kohli. 2019 · 2019
Earlier work this paper cites.
Adversarial Authorship Attribution in Open-Source Projects. In Proceedings of the 9th ACM Conference on Data and Application Security and Privacy . ACM, Richardson, TX, USA, 291–302
Alina Matyukhina, Natalia Stakhanova, Mila Dalla Preda, and Celine Perley. 2019 · 2019
Earlier work this paper cites.
Misleading Authorship Attribution of Source Code using Adversarial Learning. In Proceedings of the 28th USENIX Security Symposium . USENIX Association, Santa Clara, CA, USA, 479–496
Erwin Quiring, Alwin Maier, and Konrad Rieck. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
A Systematic Review of Interaction in Search-Based Software Engineering
Aurora Ramírez, José Raúl Romero, and Christopher L. Simons. 2019 · 2019
Earlier work this paper cites.
Pythia: AI-assisted Code Completion System. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery . ACM, Anchorage, AK, USA, 2727–2735
Alexey Svyatkovskiy, Ying Zhao, Shengyu Fu, and Neel Sundaresan. 2019 · 2019
Earlier work this paper cites.
Multi-modal Attention Network Learning for Semantic Source Code Retrieval. In Proceedings of the 34th IEEE/ACM International Conference on Automated Software Engineering . IEEE, San Diego, CA, USA, 13–25
Yao Wan, Jingdong Shu, Yulei Sui, Guandong Xu, Zhou Zhao, Jian Wu, and Philip S. Yu. 2019 · 2019
Earlier work this paper cites.
A novel neural source code representation based on abstract syntax tree. In Proceedings of the 41st International Conference on Software Engineering . IEEE / ACM, Montreal, QC, Canada, 783–794
Jian Zhang, Xu Wang, Hongyu Zhang, Hailong Sun, Kaixuan Wang, and Xudong Liu. 2019 · 2019
Earlier work this paper cites.
Devign: Effective Vulnerability Identification by Learning Comprehensive Program Semantics via Graph Neural Networks. In Proceedings of the 33rd Conference on Neural Information Processing Systems . OpenReview.net, Vancouver, BC, Canada, 10197–10207
Yaqin Zhou, Shangqing Liu, Jing Kai Siow, Xiaoning Du, and Yang Liu. 2019 · 2019
Earlier work this paper cites.
Transferable Clean-Label Poisoning Attacks on Deep Neural Nets. In Proceedings of the 36th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 97) . PMLR, Long Beach, California, USA, 7614–7623
Chen Zhu, W. Ronny Huang, Hengduo Li, Gavin Taylor, Christoph Studer, and Tom Goldstein. 2019 · 2019
Earlier work this paper cites.
A Transformer-based Approach for Source Code Summarization. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Online, 4998–5007
Wasi Uddin Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2020 · 2020
Earlier work this paper cites.
Structural Language Models of Code. In Proceedings of the 37th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 119) . PMLR, Virtual Event, 245–256
Uri Alon, Roy Sadaka, Omer Levy, and Eran Yahav. 2020 · 2020
Earlier work this paper cites.
Adversarial Robustness for Code. In Proceedings of the 37th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 119) . PMLR, Virtual Event, 896–907
Pavol Bielik and Martin T. Vechev. 2020 · 2020
Earlier work this paper cites.
A comprehensive study on challenges in deploying deep learning based software. In Proceedings of the 28th Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . ACM, Virtual Event, USA, 750–762
Zhenpeng Chen, Yanbin Cao, Yuanqiang Liu, Haoyu Wang, Tao Xie, and Xuanzhe Liu. 2020 · 2020
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages. In Proceedings of the Findings of the Association for Computational Linguistics (Findings of ACL, Vol. EMNLP 2020) . Association for Computational Linguistics, Online Event, 1536–1547
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, and Ming Zhou. 2020 · 2020
Cited alongside, same era.
A multi-perspective architecture for semantic code search
Rajarshi Haldar, Lingfei Wu, Jinjun Xiong, and Julia Hockenmaier. 2020 · 2020
Cited alongside, same era.
Improved Automatic Summarization of Subroutines via Attention to File Context. In MSR ’20: 17th International Conference on Mining Software Repositories, , , 2020 . ACM, Seoul, Republic of Korea, 300–310
Sakib Haque, Alexander LeClair, Lingfei Wu, and Collin McMillan. 2020 · 2020
Cited alongside, same era.
Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment. In Proceedings of the 34th AAAI Conference on Artificial Intelligence . AAAI Press, New York, NY, USA, 8018–8025
Occlusion-based Detection of Trojan-triggering Inputs in Large Language Models of Code
Aftab Hussain, Md. Rafiqul Islam Rabin, Toufique Ahmed, Mohammad Amin Alipour, and Bowen Xu. 2023a · 2023
Later among the works it cites.
A Survey of Trojans in Neural Models of Source Code: Taxonomy and Techniques
Aftab Hussain, Md. Rafiqul Islam Rabin, Toufique Ahmed, Bowen Xu, Prem Devanbu, and Mohammad Amin Alipour. 2023b · 2023
Later among the works it cites.
Trojanedcm: A repository for poisoned neural models of source code
Aftab Hussain, Md Rafiqul Islam Rabin, and Mohammad Amin Alipour. 2023c · 2023
Later among the works it cites.
Enhancing Robustness of AI Offensive Code Generators via Data Augmentation
Cristina Improta, Pietro Liguori, Roberto Natella, Bojan Cukic, and Domenico Cotroneo. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Di Jin, Zhijing Jin, Joey Tianyi Zhou, and Peter Szolovits. 2020 · 2020
Cited alongside, same era.
Contextualized perturbation for textual adversarial attack
Dianqi Li, Yizhe Zhang, Hao Peng, Liqun Chen, Chris Brockett, Ming-Ting Sun, and Bill Dolan. 2020b · 2020
Cited alongside, same era.
BERT-ATTACK: Adversarial Attack Against BERT Using BERT. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Online, 6193–6202
Linyang Li, Ruotian Ma, Qipeng Guo, Xiangyang Xue, and Xipeng Qiu. 2020a · 2020
Cited alongside, same era.
Software Vulnerability Detection Using Deep Neural Networks: A Survey
Guanjun Lin, Sheng Wen, Qing-Long Han, Jun Zhang, and Yang Xiang. 2020 · 2020
Cited alongside, same era.
Cognitive Biases in Software Engineering: A Systematic Mapping Study
Rahul Mohanani, Iflaah Salman, Burak Turhan, Pilar Rodríguez, and Paul Ralph. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
CodeBLEU: a Method for Automatic Evaluation of Code Synthesis
Shuo Ren, Daya Guo, Shuai Lu, Long Zhou, Shujie Liu, Duyu Tang, Neel Sundaresan, Ming Zhou, Ambrosio Blanco, and Shuai Ma. 2020 · 2020
Cited alongside, same era.
STRATA: simple, gradient-free attacks for models of code
Jacob M Springer, Bryn Marie Reinstadler, and Una-May O’Reilly. 2020 · 2020
Cited alongside, same era.
IntelliCode compose: code generation using transformer. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . ACM, Virtual Event, USA, 1433–1443
Alexey Svyatkovskiy, Shao Kun Deng, Shengyu Fu, and Neel Sundaresan. 2020 · 2020
Cited alongside, same era.
CodeAttack: Code-Based Adversarial Attacks for Pre-trained Programming Language Models. In Thirty-Seventh AAAI Conference on Artificial Intelligence, AAAI 2023, Thirty-Fifth Conference on Innovative Applications of Artificial Intelligence, IAAI 2023, Thirteenth Symposium on Educational Advances in Artificial Intelligence . AAAI Press, Washington, DC, USA, 14892–14900
Akshita Jha and Chandan K. Reddy. 2023 · 2023
Later among the works it cites.
SEGRESS: Software Engineering Guidelines for REporting Secondary Studies
Barbara A. Kitchenham, Lech Madeyski, and David Budgen. 2023 · 2023
Later among the works it cites.
StarCoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, Jia Li, and Jenny Chim and. 2023a · 2023
Later among the works it cites.
A comparative study of adversarial training methods for neural models of source code
Zhen Li, Xiang Huang, Yangrui Li, and Guenevere Chen. 2023b · 2023
Later among the works it cites.
Deep Learning for Android Malware Defenses: A Systematic Literature Review
Yue Liu, Chakkrit Tantithamthavorn, Li Li, and Yepang Liu. 2023 · 2023
Later among the works it cites.
A Data-free Backdoor Injection Approach in Neural Networks. In Proceedings of the 32nd USENIX Security Symposium . USENIX Association, Anaheim, CA, USA, 2671–2688
Peizhuo Lv, Chang Yue, Ruigang Liang, Yunfei Yang, Shengzhi Zhang, Hualong Ma, and Kai Chen. 2023 · 2023
Later among the works it cites.
Evolutionary Approaches for Adversarial Attacks on Neural Source Code Classifiers
Valeria Mercuri, Martina Saletta, and Claudio Ferretti. 2023 · 2023
Later among the works it cites.
Code Llama: Open Foundation Models for Code
Meta. 2023 · 2023
Later among the works it cites.
DIP: Dead code Insertion based Black-box Attack for Programming Language Model. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Toronto, Canada, 7777–7791
CheolWon Na, YunSeok Choi, and Jee-Hyong Lee. 2023 · 2023
Later among the works it cites.
Adversarial Attacks on Code Models with Discriminative Graph Patterns
Thanh-Dat Nguyen, Yang Zhou, Xuan-Bach Dinh Le, Patanamon Thongtanunam, and David Lo. 2023 · 2023
Later among the works it cites.
CodexLeaks: Privacy Leaks from Code Generation Language Models in GitHub Copilot. In Proceedings of the 32nd USENIX Security Symposium . USENIX Association, Anaheim, CA, USA, 2133–2150
Liang Niu, Muhammad Shujaat Mirza, Zayd Maradni, and Christina Pöpper. 2023 · 2023
Later among the works it cites.
Jinfeng Wen and Zhenpeng Chen and Xin Jin and Xuanzhe Liu
Rise of the Planet of Serverless Computing: A Systematic Review. 2023 · 2023
Later among the works it cites.
Code Interpreter
OpenAI. 2023 · 2023
Later among the works it cites.
Toward a Theory of Causation for Interpreting Neural Code Models
David N Palacio, Nathan Cooper, Alvaro Rodriguez, Kevin Moran, and Denys Poshyvanyk. 2023 · 2023
Later among the works it cites.
BadCS: A Backdoor Attack Framework for Code search
Shiyi Qi, Yuanhang Yang, Shuzheng Gao, Cuiyun Gao, and Zenglin Xu. 2023 · 2023
Later among the works it cites.
Backdooring Neural Code Search. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Toronto, Canada, 9692–9708
Weisong Sun, Yuchen Chen, Guanhong Tao, Chunrong Fang, Xiangyu Zhang, Quanjun Zhang, and Bin Luo. 2023 · 2023
Later among the works it cites.
Code Difference Guided Adversarial Example Generation for Deep Code Models. In Proceedings of the 38th IEEE/ACM International Conference on Automated Software Engineering . IEEE, luxembourg, 850–862
Zhao Tian, Junjie Chen, and Zhi Jin. 2023 · 2023
Later among the works it cites.
LLMSecEval: A Dataset of Natural Language Prompts for Security Evaluations. In Proceedings of the 20th IEEE/ACM International Conference on Mining Software Repositories . IEEE, Melbourne, Australia, 588–592
Catherine Tony, Markus Mutas, Nicolás E. Díaz Ferreyra, and Riccardo Scandariato. 2023 · 2023
Later among the works it cites.
Machine/Deep Learning for Software Engineering: A Systematic Literature Review
Simin Wang, Liguo Huang, Amiao Gao, Jidong Ge, Tengfei Zhang, Haitao Feng, Ishna Satyarth, Ming Li, He Zhang, and Vincent Ng. 2023b · 2023
Later among the works it cites.
CodeT5+: Open Code Large Language Models for Code Understanding and Generation
Yue Wang, Hung Le, Akhilesh Deepak Gotmare, Nghi D. Q. Bui, Junnan Li, and Steven C. H. Hoi. 2023c · 2023
Later among the works it cites.
Vulnerability Detection with Graph Simplification and Enhanced Graph Representation Learning. In Proceedings of the 45th IEEE/ACM International Conference on Software Engineering . IEEE, Melbourne, Australia, 2275–2286
Xin-Cheng Wen, Yupan Chen, Cuiyun Gao, Hongyu Zhang, Jie M. Zhang, and Qing Liao. 2023 · 2023
Later among the works it cites.
Deceptprompt: Exploiting llm-driven code generation via adversarial natural language instructions
Fangzhou Wu, Xiaogeng Liu, and Chaowei Xiao. 2023 · 2023
Later among the works it cites.
Towards Privacy Preserving Cross Project Defect Prediction with Federated Learning. In Proceedings of the IEEE International Conference on Software Analysis, Evolution and Reengineering . IEEE, Taipa, Macao, 485–496
Hiroki Yamamoto, Dong Wang, Gopi Krishnan Rajbahadur, Masanari Kondo, Yasutaka Kamei, and Naoyasu Ubayashi. 2023 · 2023
Later among the works it cites.
AdvBinSD: Poisoning the Binary Code Similarity Detector via Isolated Instruction Sequences. In Intl Conf on Parallel & Distributed Processing with Applications, Big Data & Cloud Computing, Sustainable Computing & Communications, Social Computing & Networking . IEEE, Wuhan, China, 1149–1156
Xiaoyu Yi, Gaolei Li, Ao Ding, Yan Zheng, Yuqing Li, and Jianhua Li. 2023 · 2023
Later among the works it cites.
AdVulCode: Generating Adversarial Vulnerable Code against Deep Learning-Based Vulnerability Detectors
Xueqi Yu, Zhen Li, Xiang Huang, and Shasha Zhao. 2023 · 2023
Later among the works it cites.
degraphcs: Embedding variable-based flow graph for neural code search
Chen Zeng, Yue Yu, Shanshan Li, Xin Xia, Zhiming Wang, Mingyang Geng, Linxiao Bai, Wei Dong, and Xiangke Liao. 2023 · 2023
Later among the works it cites.
Transfer Attacks and Defenses for Large Language Models on Coding Tasks
Chi Zhang, Zifan Wang, Ravi Mangal, Matt Fredrikson, Limin Jia, and Corina S. Pasareanu. 2023c · 2023
Later among the works it cites.
RNNS: Representation Nearest Neighbor Search Black-Box Attack on Code Models
Jie Zhang, Wei Ma, Qiang Hu, Xiaofei Xie, Yves Le Traon, and Yang Liu. 2023a · 2023
Later among the works it cites.
TrojanSQL: SQL Injection against Natural Language Interface to Database. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Singapore, 4344–4359
Jinchuan Zhang, Yan Zhou, Binyuan Hui, Yaxin Liu, Ziming Li, and Songlin Hu. 2023e · 2023
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
How Robust Is a Large Pre-trained Language Model for Code Generation f f A Case on Attacking GPT2. In Proceedings of the IEEE International Conference on Software Analysis, Evolution and Reengineering . IEEE, Taipa, Macao, 708–712
Rui Zhu and Cunming Zhang. 2023 · 2023
Later among the works it cites.
Data Augmentation Approaches for Source Code Models: A Survey
Terry Yue Zhuo, Zhou Yang, Zhensu Sun, Yufei Wang, Li Li, Xiaoning Du, Zhenchang Xing, and David Lo. 2023 · 2023
Later among the works it cites.
TrojanPuzzle: Covertly Poisoning Code-Suggestion Models. In Proceedings of the IEEE Symposium on Security and Privacy . IEEE, San Francisco, CA, USA, 1122–1140
Hojjat Aghakhani, Wei Dai, Andre Manoel, Xavier Fernandes, Anant Kharkar, Christopher Kruegel, Giovanni Vigna, David Evans, Ben Zorn, and Robert Sim. 2024 · 2024
Closest in time.
Traces of Memorisation in Large Language Models for Code. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . ACM, Lisbon, Portugal, 78:1–78:12
Ali Al-Kaswan, Maliheh Izadi, and Arie van Deursen. 2024 · 2024
Closest in time.
Investigating Adversarial Attacks in Software Analytics via Machine Learning Explainability
MD Awal, Mrigank Rochan, and Chanchal K Roy. 2024 · 2024
Closest in time.
Artifacts of this paper
Yuchen Chen and Weisong Sun. 2024 · 2024
Closest in time.
Vulnerabilities in AI Code Generators: Exploring Targeted Data Poisoning Attacks. In Proceedings of the 32nd IEEE/ACM International Conference on Program Comprehension . ACM, Lisbon, Portugal, 280–292
Domenico Cotroneo, Cristina Improta, Pietro Liguori, and Roberto Natella. 2024 · 2024
Closest in time.
Machine Learning for Actionable Warning Identification: A Comprehensive Survey
Xiuting Ge, Chunrong Fang, Xuanye Li, Weisong Sun, Daoyuan Wu, Juan Zhai, Shangwei Lin, Zhihong Zhao, Yang Liu, and Zhenyu Chen. 2024 · 2024
Closest in time.
MalwareTotal: Multi-Faceted and Sequence-Aware Bypass Tactics against Static Malware Detection. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . ACM, Lisbon, Portugal, 172:1–172:12
Shuai He, Cai Fu, Hong Hu, Jiahe Chen, Jianqiang Lv, and Shuai Jiang. 2024 · 2024
Closest in time.
Large Language Models for Software Engineering: A Systematic Literature Review
Xinyi Hou, Yanjie Zhao, Yue Liu, Zhou Yang, Kailong Wang, Li Li, Xiapu Luo, David Lo, John Grundy, and Haoyu Wang. 2024 · 2024
Closest in time.
On Trojan Signatures in Large Language Models of Code
Aftab Hussain, Md. Rafiqul Islam Rabin, and Mohammad Amin Alipour. 2024a · 2024
Closest in time.
Aftab Hussain, Md. Rafiqul Islam Rabin, Navid Ayoobi, and Mohammad Amin Alipour. 2024b · 2024
Closest in time.
Language Models for Code Completion: A Practical Evaluation. In Proceedings of the 46th International Conference on Software Engineering . ACM, Lisbon, Portugal, 79:1–79:13
Maliheh Izadi, Jonathan Katzy, Tim van Dam, Marc Otten, Razvan Mihai Popescu, and Arie van Deursen. 2024 · 2024
Closest in time.
Poison Attack and Poison Detection on Deep Source Code Processing Models
Jia Li, Zhuo Li, Huangzhao Zhang, Ge Li, Zhi Jin, Xing Hu, and Xin Xia. 2024b · 2024
Closest in time.
Bias Behind the Wheel: Fairness Testing of Autonomous Driving Systems
Xinyue Li, Zhenpeng Chen, Jie M Zhang, Federica Sarro, Ying Zhang, and Xuanzhe Liu. 2024a · 2024
Closest in time.
Large Language Model-Based Agents for Software Engineering: A Survey
Junwei Liu, Kaixin Wang, Yixuan Chen, Xin Peng, Zhenpeng Chen, Lingming Zhang, and Yiling Lou. 2024 · 2024
Closest in time.
Poisoned ChatGPT Finds Work for Idle Hands: Exploring Developers’ Coding Practices with Insecure Suggestions from Poisoned AI Models. In Proceedings of the IEEE Symposium on Security and Privacy . IEEE, San Francisco, CA, USA, 1141–1159
Sanghak Oh, Kiho Lee, Seonhye Park, Doowon Kim, and Hyoungshick Kim. 2024 · 2024
Closest in time.
Talking about large language models
Murray Shanahan. 2024 · 2024
Closest in time.
A Survey of Source Code Search: A 3-Dimensional Perspective
Weisong Sun, Chunrong Fang, Yifei Ge, Yuling Hu, Yuchen Chen, Quanjun Zhang, Xiuting Ge, Yang Liu, and Zhenyu Chen. 2024b · 2024
Closest in time.
Towards Trustworthy LLMs for Code: A Data-Centric Synergistic Auditing Framework
Chong Wang, Zhenpeng Chen, Tianlin Li, Yilun Zhao, and Yang Liu. 2024 · 2024
Closest in time.
Semantic Sleuth: Identifying Ponzi Contracts via Large Language Models. In Proceedings of the 39th International Conference on Automated Software Engineering . ACM, Sacramento, CA, USA, 582–593
Cong Wu, Jing Chen, Ziwei Wang, Ruichao Liang, and Ruiying Du. 2024 · 2024
Closest in time.
DeCE: Deceptive Cross-Entropy Loss Designed for Defending Backdoor Attacks
Guang Yang, Yu Zhou, Xiang Chen, Xiangyu Zhang, Terry Yue Zhuo, David Lo, and Taolue Chen. 2024f · 2024
Closest in time.
How important are good method names in neural code generation? a model robustness perspective
Guang Yang, Yu Zhou, Wenhua Yang, Tao Yue, Xiang Chen, and Taolue Chen. 2024g · 2024
Closest in time.
Exploiting the Adversarial Example Vulnerability of Transfer Learning of Source Code
Yulong Yang, Haoran Fan, Chenhao Lin, Qian Li, Zhengyu Zhao, and Chao Shen. 2024a · 2024
Closest in time.
Zhou Yang, Zhensu Sun, Terry Yue Zhuo, Premkumar T. Devanbu, and David Lo. 2024b · 2024
Closest in time.
Stealthy Backdoor Attack for Code Models
Zhou Yang, Bowen Xu, Jie M. Zhang, Hong Jin Kang, Jieke Shi, Junda He, and David Lo. 2024c · 2024
Closest in time.
Gotcha! This Model Uses My Code! Evaluating Membership Leakage Risks in Code Models
Zhou Yang, Zhipeng Zhao, Chenyu Wang, Jieke Shi, Dongsun Kim, DongGyun Han, and David Lo. 2024d · 2024
Closest in time.
LateBA: Latent Backdoor Attack on Deep Bug Search via Infrequent Execution Codes. In Proceedings of the 15th Asia-Pacific Symposium on Internetware . ACM, Macau, SAR, China, 1–10
Xiaoyu Yi, Gaolei Li, Wenkai Huang, Xi Lin, Jianhua Li, and Yuchen Liu. 2024 · 2024
Closest in time.
CodeBERT-Attack: Adversarial attack against source code deep learning models via pre-trained model
Huangzhao Zhang, Shuai Lu, Zhuo Li, Zhi Jin, Lei Ma, Yang Liu, and Ge Li. 2024c · 2024
Closest in time.
A Survey of Learning-based Automated Program Repair
Quanjun Zhang, Chunrong Fang, Yuxiang Ma, Weisong Sun, and Zhenyu Chen. 2024a · 2024
Closest in time.
Evolutionary Multi-objective Optimization for Contextual Adversarial Example Generation
Shasha Zhou, Mingyu Huang, Yanan Sun, and Ke Li. 2024 · 2024
Closest in time.
DeCoMa: Detecting and Purifying Code Dataset Watermarks through Dual Channel Code Abstraction. In Proceedings of the 34th ACM SIGSOFT International Symposium on Software Testing and Analysis . ACM, Trondheim, Norway, 1–23
Yuan Xiao, Yuchen Chen, Shiqing Ma, Haocheng Huang, Chunrong Fang, Yanwei Chen, Weisong Sun, Yunfeng Zhu, Xiaofang Zhang, and Zhenyu Chen. 2025 · 2025
Closest in time.
CARL: Unsupervised Code-Based Adversarial Attacks for Programming Language Models via Reinforcement Learning
Kaichun Yao, Hao Wang, Chuan Qin, Hengshu Zhu, Yanjun Wu, and Libo Zhang. 2025 · 2025
Closest in time.
Esale: Enhancing Code-Summary Alignment Learning for Source Code Summarization
Chunrong Fang, Weisong Sun, Yuchen Chen, Xiao Chen, Zhao Wei, Quanjun Zhang, Yudu You, Bin Luo, Yang Liu, and Zhenyu Chen. 2024 · 2095
Closest in time.