Fetching the paper…
Reading the bibliography…
Large language models for code (LLM4Code), which demonstrate strong performance (e.g., high accuracy) in processing source code, have significantly transformed software engineering.
Testing Neural Program Analyzers
Md Rafiqul Islam Rabin, Ke Wang, and Mohammad Amin Alipour. 2019 · 1908
Earlier work this paper cites.
RoPGen: Towards Robust Code Authorship Attribution via Automatic Coding Style Transformation. In Proceedings of the 44th International Conference on Software Engineering (Pittsburgh, Pennsylvania) (ICSE ’22) . Association for Computing Machinery, New York, NY, USA, 1906–1918
Zhen Li, Guenevere (Qian) Chen, Chen Chen, Yayi Zou, and Shouhuai Xu. 2022a · 1918
Earlier work this paper cites.
Equation of state calculations by fast computing machines
Nicholas Metropolis, Arianna W Rosenbluth, Marshall N Rosenbluth, Augusta H Teller, and Edward Teller. 1953 · 1953
Earlier work this paper cites.
Sequence Model Design for Code Completion in the Modern IDE
Gareth Ari Aye and Gail E. Kaiser. 2020 · 2004
Earlier work this paper cites.
Extraction of Java program fingerprints for software authorship identification
Haibiao Ding and Mansur H. Samadzadeh. 2004 · 2004
Earlier work this paper cites.
Evaluation of Generalizability of Neural Program Analyzers under Semantic-Preserving Transformations
Md Rafiqul Islam Rabin and Mohammad Amin Alipour. 2021 · 2004
Earlier work this paper cites.
Source Code Author Identification Based on N-gram Author Profiles. In Artificial Intelligence Applications and Innovations , Ilias Maglogiannis, Kostas Karpouzis, and Max Bramer (Eds.). Springer US, Boston, MA, 508–515
Georgia Frantzeskou, Efstathios Stamatatos, Stefanos Gritzalis, and Sokratis Katsikas. 2006 · 2006
Earlier work this paper cites.
A Probabilistic Approach to Source Code Authorship Identification. In Fourth International Conference on Information Technology (ITNG’07) . 243–248
Jay Kothari, Maxim Shevertalov, Edward Stehle, and Spiros Mancoridis. 2007 · 2007
Earlier work this paper cites.
Systematic literature studies: Database searches vs. backward snowballing. In Proceedings of the 2012 ACM-IEEE International Symposium on Empirical Software Engineering and Measurement . 29–38
Samireh Jalali and Claes Wohlin. 2012 · 2012
Earlier work this paper cites.
Towards a Big Data Curated Benchmark of Inter-project Code Clones. In 2014 ICSME . 476–480
Jeffrey Svajlenko, Judith F. Islam, Iman Keivanloo, Chanchal K. Roy, and Mohammad Mamun Mia. 2014 · 2014
Earlier work this paper cites.
Guidelines for Snowballing in Systematic Literature Studies and a Replication in Software Engineering. In Proceedings of the 18th International Conference on Evaluation and Assessment in Software Engineering (London, England, United Kingdom) (EASE ’14) . Association for Computing Machinery, New York, NY, USA, Article 38, 10 pages
Claes Wohlin. 2014 · 2014
Earlier work this paper cites.
De-Anonymizing Programmers via Code Stylometry. In Proceedings of the 24th USENIX Conference on Security Symposium (Washington, D.C.) (SEC’15) . USENIX Association, USA, 255–270
Aylin Caliskan-Islam, Richard Harang, Andrew Liu, Arvind Narayanan, Clare Voss, Fabian Yamaguchi, and Rachel Greenstadt. 2015 · 2015
Earlier work this paper cites.
Ian Goodfellow, Jonathon Shlens, and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Convolutional Neural Networks over Tree Structures for Programming Language Processing. In Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence (Phoenix, Arizona) (AAAI’16) . AAAI Press, 1287–1293
Lili Mou, Ge Li, Lu Zhang, Tao Wang, and Zhi Jin. 2016 · 2016
Earlier work this paper cites.
Transferability in machine learning: from phenomena to black-box attacks using adversarial samples
Nicolas Papernot, Patrick McDaniel, and Ian Goodfellow. 2016 · 2016
Earlier work this paper cites.
Probabilistic Model for Code with Decision Trees. In Proceedings of the 2016 ACM SIGPLAN International Conference on Object-Oriented Programming, Systems, Languages, and Applications (Amsterdam, Netherlands) (OOPSLA 2016) . Association for Computing Machinery, New York, NY, USA, 731–747
Veselin Raychev, Pavol Bielik, and Martin Vechev. 2016 · 2016
Earlier work this paper cites.
"Why Should I Trust You?": Explaining the Predictions of Any Classifier. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining (San Francisco, California, USA) (KDD ’16) . Association for Computing Machinery, New York, NY, USA, 1135–1144
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Source code authorship attribution using long short-term memory based networks. In Computer Security - ESORICS 2017 . Springer Verlag, 65–82
Bander Alsulami, Edwin Dauber, Richard Harang, Spiros Mancoridis, and Rachel Greenstadt. 2017 · 2017
Earlier work this paper cites.
Towards Evaluating the Robustness of Neural Networks. In 2017 IEEE Symposium on Security and Privacy (SP) . IEEE Computer Society, Los Alamitos, CA, USA, 39–57
Nicholas Carlini and David Wagner. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Authorship attribution of source code by using back propagation neural network based on particle swarm optimization
Xinyu Yang, Guoai Xu, Qi Li, Yanhui Guo, and Miao Zhang. 2017 · 2017
Earlier work this paper cites.
Detecting Backdoor Attacks on Deep Neural Networks by Activation Clustering
Bryant Chen, Wilka Carvalho, Nathalie Baracaldo, Heiko Ludwig, Benjamin Edwards, Taesung Lee, Ian M. Molloy, and Biplav Srivastava. 2018 · 2018
Earlier work this paper cites.
Improving Text-to-SQL Evaluation Methodology. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Iryna Gurevych and Yusuke Miyao (Eds.). Association for Computational Linguistics, Melbourne, Australia, 351–360
Catherine Finegan-Dollak, Jonathan K. Kummerfeld, Li Zhang, Karthik Ramanathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev. 2018 · 2018
Earlier work this paper cites.
Summarizing Source Code with Transferred API Knowledge. In Proceedings of the 27th International Joint Conference on Artificial Intelligence (Stockholm, Sweden) (IJCAI’18) . AAAI Press, 2269–2275
Xing Hu, Ge Li, Xin Xia, David Lo, Shuai Lu, and Zhi Jin. 2018 · 2018
Earlier work this paper cites.
The Mythos of Model Interpretability: In Machine Learning, the Concept of Interpretability is Both Important and Slippery
Zachary C. Lipton. 2018 · 2018
Earlier work this paper cites.
Stress Test Evaluation for Natural Language Inference. In Proceedings of the 27th International Conference on Computational Linguistics . Association for Computational Linguistics, Santa Fe, New Mexico, USA, 2340–2353
Aakanksha Naik, Abhilasha Ravichander, Norman Sadeh, Carolyn Rose, and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
Neural Program Search: Solving Programming Tasks from Description and Examples
Illia Polosukhin and Alexander Skidanov. 2018 · 2018
Earlier work this paper cites.
Recognizing and Imitating Programmer Style: Adversaries in Program Authorship Attribution
Lucy Simko, Luke Zettlemoyer, and Tadayoshi Kohno. 2018 · 2018
Earlier work this paper cites.
Explanations of model predictions with live and breakDown packages
Mateusz Staniak and Przemyslaw Biecek. 2018 · 2018
Earlier work this paper cites.
Spectral Signatures in Backdoor Attacks. In Advances in Neural Information Processing Systems , S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, and R. Garnett (Eds.), Vol. 31. Curran Associates, Inc
Brandon Tran, Jerry Li, and Aleksander Madry. 2018 · 2018
Earlier work this paper cites.
The Adverse Effects of Code Duplication in Machine Learning Models of Code. In Proceedings of the 2019 ACM SIGPLAN International Symposium on New Ideas, New Paradigms, and Reflections on Programming and Software (Athens, Greece) (Onward! 2019) . Association for Computing Machinery, New York, NY, USA, 143–153
Miltiadis Allamanis. 2019 · 2019
Earlier work this paper cites.
Code2vec: Learning Distributed Representations of Code
Uri Alon, Meital Zilberstein, Omer Levy, and Eran Yahav. 2019b · 2019
Earlier work this paper cites.
Parameter-efficient transfer learning for NLP. In International Conference on Machine Learning . PMLR, 2790–2799
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Earlier work this paper cites.
CodeSearchNet challenge: Evaluating the state of semantic code search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Earlier work this paper cites.
TextBugger: Generating Adversarial Text Against Real-world Applications. In Proceedings 2019 Network and Distributed System Security Symposium . Internet Society
Jinfeng Li, Shouling Ji, Tianyu Du, Bo Li, and Ting Wang. 2019 · 2019
Earlier work this paper cites.
Adversarial Authorship Attribution in Open-Source Projects. In Proceedings of the Ninth ACM Conference on Data and Application Security and Privacy (Richardson, Texas, USA) (CODASPY ’19) . Association for Computing Machinery, New York, NY, USA, 291–302
Alina Matyukhina, Natalia Stakhanova, Mila Dalla Preda, and Celine Perley. 2019 · 2019
Earlier work this paper cites.
Misleading Authorship Attribution of Source Code Using Adversarial Learning. In Proceedings of the 28th USENIX Conference on Security Symposium (Santa Clara, CA, USA) (SEC’19) . USENIX Association, USA, 479–496
Erwin Quiring, Alwin Maier, and Konrad Rieck. 2019 · 2019
Earlier work this paper cites.
Neural Network-Based Detection of Self-Admitted Technical Debt: From Performance to Explainability
Xiaoxue Ren, Zhenchang Xing, Xin Xia, David Lo, Xinyu Wang, and John Grundy. 2019 · 2019
Earlier work this paper cites.
An Empirical Study on Learning Bug-Fixing Patches in the Wild via Neural Machine Translation
Michele Tufano, Cody Watson, Gabriele Bavota, Massimiliano Di Penta, Martin White, and Denys Poshyvanyk. 2019 · 2019
Earlier work this paper cites.
Machine Learning Testing: Survey, Landscapes and Horizons
Jie M. Zhang, Mark Harman, Lei Ma, and Yang Liu. 2022b · 2019
Earlier work this paper cites.
Devign: Effective Vulnerability Identification by Learning Comprehensive Program Semantics via Graph Neural Networks
Yaqin Zhou, Shangqing Liu, Jingkai Siow, Xiaoning Du, and Yang Liu. 2019 · 2019
Earlier work this paper cites.
A Transformer-based Approach for Source Code Summarization. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Online, 4998–5007
Wasi Ahmad, Saikat Chakraborty, Baishakhi Ray, and Kai-Wei Chang. 2020 · 2020
Earlier work this paper cites.
Adversarial robustness for code. In International Conference on Machine Learning . PMLR, 896–907
Pavol Bielik and Martin Vechev. 2020 · 2020
Earlier work this paper cites.
Understanding Promotion-as-a-Service on GitHub. In Proceedings of the 36th Annual Computer Security Applications Conference (ACSAC ’20) . Association for Computing Machinery, New York, NY, USA, 597–610
Kun Du, Hao Yang, Yubao Zhang, Haixin Duan, Haining Wang, Shuang Hao, Zhou Li, and Min Yang. 2020 · 2020
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, 1536–1547
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, and Ming Zhou. 2020 · 2020
Earlier work this paper cites.
Compressing bert: Studying the effects of weight pruning on transfer learning
Mitchell A Gordon, Kevin Duh, and Nicholas Andrews. 2020 · 2020
Earlier work this paper cites.
An Empirical Study of Model-Agnostic Techniques for Defect Prediction Models
Jirayus Jiarpakdee, Chakkrit Kla Tantithamthavorn, Hoa Khanh Dam, and John Grundy. 2022 · 2020
Earlier work this paper cites.
Proceedings of the ACM on Programming Languages 4, OOPSLA (2020)
Noam Yefet, Uri Alon, and Eran Yahav. 2020 · 2020
Earlier work this paper cites.
Generating Adversarial Examples for Holding Robustness of Source Code Processing Models
Huangzhao Zhang, Zhuo Li, Ge Li, Lei Ma, Yang Liu, and Zhi Jin. 2020a · 2020
Earlier work this paper cites.
Interpretable Text-to-SQL Generation with Joint Optimization. In Web Information Systems and Applications: 17th International Conference, WISA 2020, Guangzhou, China, September 23–25, 2020, Proceedings (Guangzhou, China). Springer-Verlag, Berlin, Heidelberg, 341–351
Mingdong Zhu, Xianfang Wang, and Yang Zhang. 2020 · 2020
Earlier work this paper cites.
Explainable Just-in-Time Bug Prediction: Are We There Yet?. In Proceedings of the 43rd International Conference on Software Engineering: Companion Proceedings (Virtual Event, Spain) (ICSE ’21) . IEEE Press, 129–131
Reem Aleithan. 2021 · 2021
Earlier work this paper cites.
Adversarial Robustness of Program Synthesis Models. In Advances in Programming Languages and Neurosymbolic Systems Workshop
Mrinal Anand, Pratik Kayal, and Mayank Singh. 2021 · 2021
Earlier work this paper cites.
Assessing Robustness of ML-Based Program Analysis Tools using Metamorphic Program Transformations. In ASE 2021 . 1377–1381
Leonhard Applis, Annibale Panichella, and Arie van Deursen. 2021 · 2021
Earlier work this paper cites.
Program Synthesis with Large Language Models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, and Charles Sutton. 2021 · 2021
Earlier work this paper cites.
Extracting Training Data from Large Language Models. In USENIX Security Symposium
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, Alina Oprea, and Colin Raffel. 2021 · 2021
Earlier work this paper cites.
Deep Learning Based Vulnerability Detection: Are We There Yet?
Saikat Chakraborty, Rahul Krishna, Yangruibo Ding, and Baishakhi Ray. 2022b · 2021
Earlier work this paper cites.
Evaluating the robustness of source code plagiarism detection tools to pervasive plagiarism-hiding modifications
Hayden Cheers, Yuqing Lin, and Shamus P Smith. 2021 · 2021
Earlier work this paper cites.
Stealing Deep Reinforcement Learning Models for Fun and Profit. In Proceedings of the 2021 ACM Asia Conference on Computer and Communications Security (Virtual Event, Hong Kong) (ASIA CCS ’21) . Association for Computing Machinery, New York, NY, USA, 307–319
Kangjie Chen, Shangwei Guo, Tianwei Zhang, Xiaofei Xie, and Yang Liu. 2021a · 2021
Earlier work this paper cites.
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, and Heewoo Jun et al. 2021b · 2021
Earlier work this paper cites.
Deceiving neural source code classifiers: finding adversarial examples with grammatical evolution. In Proceedings of the Genetic and Evolutionary Computation Conference Companion (Lille, France) (GECCO ’21) . Association for Computing Machinery, New York, NY, USA, 1889–1897
Claudio Ferretti and Martina Saletta. 2021 · 2021
Earlier work this paper cites.
QuillBot as an online tool: Students’ alternative in paraphrasing and rewriting of English writing
Tira Nur Fitria. 2021 · 2021
Earlier work this paper cites.
GraphCodeBERT: Pre-training Code Representations with Data Flow. In 9th International Conference on Learning Representations, ICLR 2021, Virtual Event, Austria, May 3-7, 2021
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, Shujie Liu, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu andz Michele Tufano, Shao Kun Deng, Colin B. Clement, Dawn Drain, Neel Sundaresan, Jian Yin, Daxin Jiang, and Ming Zhou. 2021 · 2021
Earlier work this paper cites.
The Growing Cost of Deep Learning for Source Code
Vincent J. Hellendoorn and Anand Ashok Sawant. 2021 · 2021
Earlier work this paper cites.
Measuring Coding Challenge Competence With APPS
Dan Hendrycks, Steven Basart, Saurav Kadavath, Mantas Mazeika, Akul Arora, Ethan Guo, Collin Burns, Samir Puranik, Horace He, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
Prefix-Tuning: Optimizing Continuous Prompts for Generation. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , Chengqing Zong, Fei Xia, Wenjie Li, and Roberto Navigli (Eds.). Association for Computational Linguistics, Online, 4582–4597
Xiang Lisa Li and Percy Liang. 2021 · 2021
Earlier work this paper cites.
SySeVR: A Framework for Using Deep Learning to Detect Software Vulnerabilities
Zhen Li, Deqing Zou, Shouhuai Xu, Hai Jin, Yawei Zhu, and Zhaoxuan Chen. 2022e · 2021
Earlier work this paper cites.
EVIL: Exploiting Software via Natural Language. In 2021 IEEE 32nd International Symposium on Software Reliability Engineering (ISSRE) . IEEE Computer Society, Los Alamitos, CA, USA, 321–332
Pietro Liguori, Erfan Al-Hossami, Vittorio Orbinato, Roberto Natella, Samira Shaikh, Domenico Cotroneo, and Bojan Cukic. 2021 · 2021
Earlier work this paper cites.
A Practical Black-Box Attack on Source Code Authorship Identification Classifiers
Qianjun Liu, Shouling Ji, Changchang Liu, and Chunming Wu. 2021 · 2021
Earlier work this paper cites.
CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Shuai Lu, Daya Guo, Shuo Ren, Junjie Huang, Alexey Svyatkovskiy, Ambrosio Blanco, Colin B. Clement, Dawn Drain, Daxin Jiang, Duyu Tang, Ge Li, Lidong Zhou, Linjun Shou, Long Zhou, Michele Tufano, Ming Gong, Ming Zhou, Nan Duan, Neel Sundaresan, Shao Kun Deng, Shengyu Fu, and Shujie Liu. 2021 · 2021
Earlier work this paper cites.
Adversarial Attacks to API Recommender Systems: Time to Wake Up and Smell the Coffee?. In 2021 36th IEEE/ACM International Conference on Automated Software Engineering (ASE) . 253–265
Phuong T. Nguyen, Claudio Di Sipio, Juri Di Rocco, Massimiliano Di Penta, and Davide Di Ruscio. 2021 · 2021
Earlier work this paper cites.
Thinking Like a Developer? Comparing the Attention of Humans with Neural Models of Code. In 2021 36th IEEE/ACM International Conference on Automated Software Engineering (ASE) . 867–879
Matteo Paltenghi and Michael Pradel. 2021 · 2021
Earlier work this paper cites.
PyExplainer: Explaining the Predictions of Just-in-Time Defect Models. In Proceedings of the 36th IEEE/ACM International Conference on Automated Software Engineering (Melbourne, Australia) (ASE ’21) . IEEE Press, 407–418
Chanathip Pornprasit, Chakkrit Tantithamthavorn, Jirayus Jiarpakdee, Michael Fu, and Patanamon Thongtanunam. 2022 · 2021
Earlier work this paper cites.
A Search-Based Testing Framework for Deep Neural Networks of Source Code Embedding. In 14th IEEE Conference on Software Testing, Verification and Validation, ICST 2021, Porto de Galinhas, Brazil, April 12-16, 2021 . IEEE
Maryam Vahdat Pour, Zhuo Li, Lei Ma, and Hadi Hemmati. 2021 · 2021
Earlier work this paper cites.
ONION: A Simple and Effective Defense Against Textual Backdoor Attacks. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Online and Punta Cana, Dominican Republic, 9558–9566
Fanchao Qi, Yangyi Chen, Mukai Li, Yuan Yao, Zhiyuan Liu, and Maosong Sun. 2021 · 2021
Earlier work this paper cites.
On the generalizability of Neural Program Models with respect to semantic-preserving program transformations
Md Rafiqul Islam Rabin, Nghi D.Q. Bui, Ke Wang, Yijun Yu, Lingxiao Jiang, and Mohammad Amin Alipour. 2021a · 2021
Earlier work this paper cites.
On the generalizability of Neural Program Models with respect to semantic-preserving program transformations
Md Rafiqul Islam Rabin, Nghi DQ Bui, Ke Wang, Yijun Yu, Lingxiao Jiang, and Mohammad Amin Alipour. 2021b · 2021
Earlier work this paper cites.
Understanding neural code intelligence through program simplification. In Proceedings of the 29th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering (Athens, Greece) (ESEC/FSE 2021) . Association for Computing Machinery, New York, NY, USA, 441–452
Md Rafiqul Islam Rabin, Vincent J. Hellendoorn, and Mohammad Amin Alipour. 2021c · 2021
Earlier work this paper cites.
You Autocomplete Me: Poisoning Vulnerabilities in Neural Code Completion. In 30th USENIX Security Symposium (USENIX Security 21) . USENIX Association, 1559–1575
Roei Schuster, Congzheng Song, Eran Tromer, and Vitaly Shmatikov. 2021 · 2021
Cited alongside, same era.
Explainable Software Defect Prediction: Are We There Yet?
Jiho Shin, Reem Aleithan, Jaechang Nam, Junjie Wang, and Song Wang. 2021 · 2021
Cited alongside, same era.
STRATA: Simple, Gradient-Free Attacks for Models of Code
Jacob M. Springer, Bryn Marie Reinstadler, and Una-May O’Reilly. 2021 · 2021
Cited alongside, same era.
Generating Adversarial Computer Programs using Optimized Obfuscations
Shashank Srikant, Sijia Liu, Tamara Mitrovska, Shiyu Chang, Quanfu Fan, Gaoyuan Zhang, and Una-May O’Reilly. 2021 · 2021
Cited alongside, same era.
Fast and memory-efficient neural code completion. In 2021 IEEE/ACM 18th International Conference on Mining Software Repositories (MSR) . IEEE, 329–340
Alexey Svyatkovskiy, Sebastian Lee, Anna Hadjitofi, Maik Riechert, Juliana Vicente Franco, and Miltiadis Allamanis. 2021 · 2021
Where to Look When Repairing Code? Comparing the Attention of Neural Models and Developers
Dominik Huber, Matteo Paltenghi, and Michael Pradel. 2023 · 2023
Later among the works it cites.
Enhancing Robustness of AI Offensive Code Generators via Data Augmentation
Cristina Improta, Pietro Liguori, Roberto Natella, Bojan Cukic, and Domenico Cotroneo. 2023 · 2023
Later among the works it cites.
CodeAttack: code-based adversarial attacks for pre-trained programming language models. In Proceedings of the Thirty-Seventh AAAI Conference on Artificial Intelligence and Thirty-Fifth Conference on Innovative Applications of Artificial Intelligence and Thirteenth Symposium on Educational Advances in Artificial Intelligence (AAAI’23/IAAI’23/EAAI’23) . AAAI Press, Article 1670, 9 pages
Akshita Jha and Chandan K. Reddy. 2023 · 2023
Later among the works it cites.
Benchmarking and Explaining Large Language Model-based Code Generation: A Causality-Centric Approach
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Generating Adversarial Examples of Source Code Classification Models via Q-Learning-Based Markov Decision Process. In 2021 IEEE 21st International Conference on Software Quality, Reliability and Security (QRS) . 807–818
Junfeng Tian, Chenxin Wang, Zhen Li, and Yu Wen. 2021 · 2021
Cited alongside, same era.
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing, EMNLP 2021
Yue Wang, Weishi Wang, Shafiq Joty, and Steven C.H. Hoi. 2021b · 2021
Cited alongside, same era.
An Empirical Study of Model-Agnostic Interpretation Technique for Just-in-Time Software Defect Prediction. In Collaborative Computing: Networking, Applications and Worksharing , Honghao Gao and Xinheng Wang (Eds.). Springer International Publishing, Cham, 420–438
Xingguang Yang, Huiqun Yu, Guisheng Fan, Zijie Huang, Kang Yang, and Ziyi Zhou. 2021 · 2021
Cited alongside, same era.
Interpretable Program Synthesis. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems (Yokohama, Japan) (CHI ’21) . Association for Computing Machinery, New York, NY, USA, Article 105, 16 pages
Tianyi Zhang, Zhiyang Chen, Yuanli Zhu, Priyan Vaithilingam, Xinyu Wang, and Elena L. Glassman. 2021 · 2021
Cited alongside, same era.
Interpreting Deep Learning-Based Vulnerability Detector Predictions Based on Heuristic Searching
Deqing Zou, Yawei Zhu, Shouhuai Xu, Zhen Li, Hai Jin, and Hengkai Ye. 2021 · 2021
Cited alongside, same era.
Parameter-Efficient Finetuning of Transformers for Source Code
Shamil Ayupov and Nadezhda Chirkova. 2022 · 2022
Cited alongside, same era.
EW-Tune: A Framework for Privately Fine-Tuning Large Language Models with Differential Privacy. In 2022 IEEE International Conference on Data Mining Workshops (ICDMW) . IEEE
Rouzbeh Behnia, Mohammadreza Reza Ebrahimi, Jason Pacheco, and Balaji Padmanabhan. 2022 · 2022
Cited alongside, same era.
Zhenlan Ji, Pingchuan Ma, Zongjie Li, and Shuai Wang. 2023 · 2023
Later among the works it cites.
ClawSAT: Towards Both Robust and Accurate Code Models. In 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE Computer Society, Los Alamitos, CA, USA, 212–223
Jinghan Jia, Shashank Srikant, Tamara Mitrovska, Chuang Gan, Shiyu Chang, Sijia Liu, and Una-May O’Reilly. 2023 · 2023
Later among the works it cites.
Connecting the .dotfiles: Checked-In Secret Exposure with Extra (Lateral Movement) Steps. In 2023 IEEE/ACM 20th International Conference on Mining Software Repositories (MSR) . 322–333
Gerhard Jungwirth, Aakanksha Saha, Michael Schröder, Tobias Fiebig, Martina Lindorfer, and Jürgen Cito. 2023 · 2023
Later among the works it cites.
LORD: Low Rank Decomposition Of Monolingual Code LLMs For One-Shot Compression
Ayush Kaushal, Tejas Vaidhya, and Irina Rish. 2023 · 2023
Later among the works it cites.
Studying the Effect of AI Code Generators on Supporting Novice Learners in Introductory Programming. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (Hamburg, Germany) (CHI ’23) . Association for Computing Machinery, New York, NY, USA, Article 455, 23 pages
Majeed Kazemitabaar, Justin Chow, Carl Ka To Ma, Barbara J. Ericson, David Weintrop, and Tovi Grossman. 2023 · 2023
Later among the works it cites.
Bonan Kou, Shengmai Chen, Zhijie Wang, Lei Ma, and Tianyi Zhang. 2023 · 2023
Later among the works it cites.
Multi-target Backdoor Attacks for Code Pre-trained Models. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Toronto, Canada, 7236–7254
Yanzhou Li, Shangqing Liu, Kangjie Chen, Xiaofei Xie, Tianwei Zhang, and Yang Liu. 2023c · 2023
Later among the works it cites.
Do Pretrained Language Models Indeed Understand Software Engineering Tasks?
Yao Li, Tao Zhang, Xiapu Luo, Haipeng Cai, Sen Fang, and Dawei Yuan. 2023g · 2023
Later among the works it cites.
Protecting Intellectual Property of Large Language Model-Based Code Generation APIs via Watermarks. In Proceedings of the 2023 ACM SIGSAC Conference on Computer and Communications Security (CCS ’23) . Association for Computing Machinery, New York, NY, USA, 2336–2350
Zongjie Li, Chaozheng Wang, Shuai Wang, and Cuiyun Gao. 2023f · 2023
Later among the works it cites.
Robin: A Novel Method to Produce Robust Interpreters for Deep Learning-Based Code Classifiers. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE Computer Society, Los Alamitos, CA, USA, 27–39
Zhen Li, Ruqian Zhang, Deqing Zou, Ning Wang, Yating Li, Shouhuai Xu, Chen Chen, and Hai Jin. 2023h · 2023
Later among the works it cites.
Can NMT Understand Me? Towards Perturbation-Based Evaluation of NMT Models for Code Generation. In Proceedings of the 1st International Workshop on Natural Language-Based Software Engineering (Pittsburgh, Pennsylvania) (NLBSE ’22) . Association for Computing Machinery, New York, NY, USA, 59–66
Pietro Liguori, Cristina Improta, Simona De Vivo, Roberto Natella, Bojan Cukic, and Domenico Cotroneo. 2023 · 2023
Later among the works it cites.
An Empirical Study of Parameter-Efficient Fine-Tuning Methods for Pre-Trained Code Models. In 2023 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE Computer Society, Los Alamitos, CA, USA, 397–408
Jiaxing Liu, Chaofeng Sha, and Xin Peng. 2023a · 2023
Later among the works it cites.
On the Reliability and Explainability of Automated Code Generation Approaches
Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, and Li Li. 2023b · 2023
Later among the works it cites.
Trustworthy and Synergistic Artificial Intelligence for Software Engineering: Vision and Roadmaps
David Lo. 2023 · 2023
Later among the works it cites.
LLaMA-Reviewer: Advancing Code Review Automation with Large Language Models through Parameter-Efficient Fine-Tuning. In 2023 IEEE 34th International Symposium on Software Reliability Engineering (ISSRE) . IEEE Computer Society, Los Alamitos, CA, USA, 647–658
Junyi Lu, Lei Yu, Xiaojia Li, Li Yang, and Chun Zuo. 2023 · 2023
Later among the works it cites.
On the Robustness of Code Generation Techniques: An Empirical Study on GitHub Copilot. In Proceedings of the 45th International Conference on Software Engineering (Melbourne, Victoria, Australia) (ICSE ’23) . IEEE Press, 2149–2160
Antonio Mastropaolo, Luca Pascarella, Emanuela Guglielmi, Matteo Ciniselli, Simone Scalabrino, Rocco Oliveto, and Gabriele Bavota. 2023 · 2023
Later among the works it cites.
Evolutionary Approaches for Adversarial Attacks on Neural Source Code Classifiers
Valeria Mercuri, Martina Saletta, and Claudio Ferretti. 2023 · 2023
Later among the works it cites.
Explaining Transformer-based Code Models: What Do They Learn? When They Do Not Work?. In 2023 IEEE 23rd International Working Conference on Source Code Analysis and Manipulation (SCAM) . IEEE Computer Society, Los Alamitos, CA, USA, 96–106
Ahmad Haji Mohammadkhani, Chakkrit Tantithamthavorn, and Hadi Hemmatif. 2023b · 2023
Later among the works it cites.
When to Show a Suggestion? Integrating Human Feedback in AI-Assisted Programming
Hussein Mozannar, Gagan Bansal, Adam Fourney, and Eric Horvitz. 2023 · 2023
Later among the works it cites.
DIP: Dead code Insertion based Black-box Attack for Programming Language Model. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Toronto, Canada, 7777–7791
CheolWon Na, YunSeok Choi, and Jee-Hyong Lee. 2023 · 2023
Later among the works it cites.
Adversarial Attacks on Code Models with Discriminative Graph Patterns
Thanh-Dat Nguyen, Zhou Yang, Xuan Bach D. Le, Patanamon, Thongtanunam, and David Lo. 2023 · 2023
Later among the works it cites.
Generative Artificial Intelligence for Software Engineering – A Research Agenda
Anh Nguyen-Duc, Beatriz Cabrero-Daniel, Adam Przybylek, Chetan Arora, Dron Khanna, Tomas Herda, Usman Rafiq, Jorge Melegati, Eduardo Guerra, Kai-Kristian Kemell, Mika Saari, Zheying Zhang, Huy Le, Tho Quan, and Pekka Abrahamsson. 2023 · 2023
Later among the works it cites.
CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis. In The Eleventh International Conference on Learning Representations
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2023 · 2023
Later among the works it cites.
Sanghak Oh, Kiho Lee, Seonhye Park, Doowon Kim, and Hyoungshick Kim. 2023 · 2023
Later among the works it cites.
Evaluating and Explaining Large Language Models for Code Using Syntactic Structures
David N Palacio, Alejandro Velasco, Daniel Rodriguez-Cardenas, Kevin Moran, and Denys Poshyvanyk. 2023 · 2023
Later among the works it cites.
The Impact of AI on Developer Productivity: Evidence from GitHub Copilot
Sida Peng, Eirini Kalliamvakou, Peter Cihon, and Mert Demirer. 2023 · 2023
Later among the works it cites.
“It’s Weird That It Knows What I Want”: Usability and Interactions with Copilot for Novice Programmers
James Prather, Brent N. Reeves, Paul Denny, Brett A. Becker, Juho Leinonen, Andrew Luxton-Reilly, Garrett Powell, James Finnie-Ansley, and Eddie Antonio Santos. 2023 · 2023
Later among the works it cites.
BadCS: A Backdoor Attack Framework for Code search
Shiyi Qi, Yuanhang Yang, Shuzhzeng Gao, Cuiyun Gao, and Zenglin Xu. 2023 · 2023
Later among the works it cites.
Benchmarking Causal Study to Interpret Large Language Models for Source Code. In 2023 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE Computer Society, Los Alamitos, CA, USA, 329–334
Daniel Rodriguez-Cardenas, David N. Palacio, Dipin Khati, Henry Burke, and Denys Poshyvanyk. 2023 · 2023
Later among the works it cites.
Naturalness of Attention: Revisiting Attention in Code Language Models
Mootez Saad and Tushar Sharma. 2023 · 2023
Later among the works it cites.
Lost at C: A User Study on the Security Implications of Large Language Model Code Assistants. In 32nd USENIX Security Symposium (USENIX Security 23) . USENIX Association, Anaheim, CA, 2205–2222
Gustavo Sandoval, Hammond Pearce, Teo Nys, Ramesh Karri, Siddharth Garg, and Brendan Dolan-Gavitt. 2023 · 2023
Later among the works it cites.
Pitfalls in Language Models for Code Intelligence: A Taxonomy and Survey
Xinyu She, Yue Liu, Yanjie Zhao, Yiling He, Li Li, Chakkrit Tantithamthavorn, Zhan Qin, and Haoyu Wang. 2023 · 2023
Later among the works it cites.
Towards Efficient Fine-Tuning of Pre-trained Code Models: An Experimental Study and Beyond. In Proceedings of the 32nd ACM SIGSOFT International Symposium on Software Testing and Analysis (ISSTA 2023) . Association for Computing Machinery, New York, NY, USA, 39–51
Ensheng Shi, Yanlin Wang, Hongyu Zhang, Lun Du, Shi Han, Dongmei Zhang, and Hongbin Sun. 2023a · 2023
Later among the works it cites.
Smaller, Faster, Greener: Compressing Pre-trained Code Models via Surrogate-Assisted Optimization
Jieke Shi, Zhou Yang, Hong Jin Kang, Bowen Xu, Junda He, and David Lo. 2023b · 2023
Later among the works it cites.
Exploring the Robustness of Large Language Models for Solving Programming Problems
Atsushi Shirafuji, Yutaka Watanobe, Takumi Ito, Makoto Morishita, Yuki Nakamura, Yusuke Oda, and Jun Suzuki. 2023 · 2023
Later among the works it cites.
Paraphrasing Techniques for Maritime QA system
Fatemeh Shiri, Terry Yue Zhuo, Zhuang Li, Van Nguyen, Shirui Pan, Weiqing Wang, Reza Haffari, and Yuan-Fang Li. 2023 · 2023
Later among the works it cites.
Milo: Attacking Deep Pre-trained Model for Programming Languages Tasks with Anti-analysis Code Obfuscation. In COMPSAC . 586–594
Leo Song and Steven H.H. Ding. 2023 · 2023
Later among the works it cites.
Backdooring Neural Code Search. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Toronto, Canada, 9692–9708
Weisong Sun, Yuchen Chen, Guanhong Tao, Chunrong Fang, Xiangyu Zhang, Quanjun Zhang, and Bin Luo. 2023a · 2023
Later among the works it cites.
CodeMark: Imperceptible Watermarking for Code Datasets against Neural Code Completion Models. In Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (ESEC/FSE 2023) . Association for Computing Machinery, New York, NY, USA, 1561–1572
Zhensu Sun, Xiaoning Du, Fu Song, and Li Li. 2023b · 2023
Later among the works it cites.
Don’t Complete It! Preventing Unhelpful Code Completion for Productive and Sustainable Neural Code Completion Systems. In 2023 IEEE/ACM 45th International Conference on Software Engineering: Companion Proceedings (ICSE-Companion) . 324–325
Zhensu Sun, Xiaoning Du, Fu Song, Shangwen Wang, Mingze Ni, and Li Li. 2023c · 2023
Later among the works it cites.
Towards More Effective AI-Assisted Programming: A Systematic Design Exploration to Improve Visual Studio IntelliCode’s User Experience. In 2023 IEEE/ACM 45th International Conference on Software Engineering: Software Engineering in Practice (ICSE-SEIP) . 185–195
Priyan Vaithilingam, Elena L. Glassman, Peter Groenwegen, Sumit Gulwani, Austin Z. Henley, Rohan Malpani, David Pugh, Arjun Radhakrishna, Gustavo Soares, Joey Wang, and Aaron Yim. 2023 · 2023
Later among the works it cites.
Helena Vasconcelos, Gagan Bansal, Adam Fourney, Q. Vera Liao, and Jennifer Wortman Vaughan. 2023 · 2023
Later among the works it cites.
ReCode: Robustness Evaluation of Code Generation Models. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Toronto, Canada, 13818–13843
Shiqi Wang, Zheng Li, Haifeng Qian, Chenghao Yang, Zijian Wang, Mingyue Shang, Varun Kumar, Samson Tan, Baishakhi Ray, Parminder Bhatia, Ramesh Nallapati, Murali Krishna Ramanathan, Dan Roth, and Bing Xiang. 2023c · 2023
Later among the works it cites.
Demystifying What Code Summarization Models Learned
Yu Wang and Ke Wang. 2023 · 2023
Later among the works it cites.
An Explanation Method for Models of Code
Yu Wang, Ke Wang, and Linzhang Wang. 2023d · 2023
Later among the works it cites.
Towards Greener Yet Powerful Code Generation via Quantization: An Empirical Study (ESEC/FSE 2023) . 224–236
Xiaokai Wei, Sujan Kumar Gonugondla, Shiqi Wang, Wasi Ahmad, Baishakhi Ray, Haifeng Qian, Xiaopeng Li, Varun Kumar, Zijian Wang, Yuchen Tian, Qing Sun, Ben Athiwaratkun, Mingyue Shang, Murali Krishna Ramanathan, Parminder Bhatia, and Bing Xiang. 2023 · 2023
Later among the works it cites.
DeceptPrompt: Exploiting LLM-driven Code Generation via Adversarial Natural Language Instructions
Fangzhou Wu, Xiaogeng Liu, and Chaowei Xiao. 2023 · 2023
Later among the works it cites.
DevGPT: Studying Developer-ChatGPT Conversations
Tao Xiao, Christoph Treude, Hideaki Hata, and Kenichi Matsumoto. 2023 · 2023
Later among the works it cites.
Towards Privacy Preserving Cross Project Defect Prediction with Federated Learning. In 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . 485–496
Hiroki Yamamoto, Dong Wang, Gopi Krishnan Rajbahadur, Masanari Kondo, Yasutaka Kamei, and Naoyasu Ubayashi. 2023 · 2023
Later among the works it cites.
COCO: Testing Code Generation Systems via Concretized Instructions
Ming Yan, Junjie Chen, Jie M. Zhang, Xuejie Cao, Chen Yang, and Mark Harman. 2023 · 2023
Later among the works it cites.
How Important Are Good Method Names in Neural Code Generation? A Model Robustness Perspective
Guang Yang, Yu Zhou, Wenhua Yang, Tao Yue, Xiang Chen, and Taolue Chen. 2023c · 2023
Later among the works it cites.
AdVulCode: Generating Adversarial Vulnerable Code against Deep Learning-Based Vulnerability Detectors
Xueqi Yu, Zhen Li, Xiang Huang, and Shasha Zhao. 2023 · 2023
Later among the works it cites.
Code Membership Inference for Detecting Unauthorized Data Use in Code Pre-trained Language Models
Sheng Zhang and Hui Li. 2023 · 2023
Later among the works it cites.
Challenging Machine Learning-Based Clone Detectors via Semantic-Preserving Code Transformations
Weiwei Zhang, Shengjian Guo, Hongyu Zhang, Yulei Sui, Yinxing Xue, and Yun Xu. 2023a · 2023
Later among the works it cites.
A Survey of Large Language Models for Code: Evolution, Benchmarking, and Future Trends
Zibin Zheng, Kaiwen Ning, Yanlin Wang, Jingwen Zhang, Dewu Zheng, Mingxi Ye, and Jiachi Chen. 2023 · 2023
Later among the works it cites.
On the Concerns of Developers When Using GitHub Copilot
Xiyu Zhou, Peng Liang, Beiqi Zhang, Zengyang Li, Aakash Ahmad, Mojtaba Shahin, and Muhammad Waseem. 2023 · 2023
Later among the works it cites.
How Robust Is a Large Pre-trained Language Model for Code Generationf A Case on Attacking GPT2. In 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . 708–712
Rui Zhu and Cunming Zhang. 2023 · 2023
Later among the works it cites.
CigaR: Cost-efficient Program Repair with LLMs
Dávid Hidvégi, Khashayar Etemadi, Sofia Bobadilla, and Martin Monperrus. 2024 · 2024
Closest in time.
Do Large Code Models Understand Programming Concepts? A Black-box Approach
Ashish Hooda, Mihai Christodorescu, Miltiadis Allamanis, Aaron Wilson, Kassem Fawaz, and Somesh Jha. 2024 · 2024
Closest in time.
Can ChatGPT Support Developers? An Empirical Evaluation of Large Language Models for Code Generation
Kailun Jin, Chung-Yu Wang, Hung Viet Pham, and Hadi Hemmati. 2024 · 2024
Closest in time.
Evaluating Program Repair with Semantic-Preserving Transformations: A Naturalness Assessment
Thanh Le-Cong, Dat Nguyen, Bach Le, and Toby Murray. 2024 · 2024
Closest in time.
Who Wrote this Code? Watermarking for Code Generation
Taehyun Lee, Seokhee Hong, Jaewoo Ahn, Ilgee Hong, Hwaran Lee, Sangdoo Yun, Jamin Shin, and Gunhee Kim. 2024 · 2024
Closest in time.
Resilient Watermarking for LLM-Generated Codes
Boquan Li, Mengdi Zhang, Peixin Zhang, Jun Sun, and Xingmei Wang. 2024 · 2024
Closest in time.
A Large-Scale Survey on the Usability of AI Programming Assistants: Successes and Challenges. In 2024 IEEE/ACM 46th International Conference on Software Engineering (ICSE) . IEEE Computer Society, Los Alamitos, CA, USA, 605–617
Jenny T. Liang, Chenyang Yang, and Brad A. Myers. 2024 · 2024
Closest in time.
On the Reliability and Explainability of Language Models for Program Generation
Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, and Li Li. 2024b · 2024
Closest in time.
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
Vahid Majdinasab, Amin Nikanjam, and Foutse Khomh. 2024 · 2024
Closest in time.
How Beginning Programmers and Code LLMs (Mis)read Each Other
Sydney Nguyen, Hannah McLean Babe, Yangtian Zi, Arjun Guha, Carolyn Jane Anderson, and Molly Q Feldman. 2024 · 2024
Closest in time.
Iman Saberi, Fatemeh Fard, and Fuxiang Chen. 2024 · 2024
Closest in time.
Distilled GPT for source code summarization
Chia-Yi Su and Collin McMillan. 2024 · 2024
Closest in time.
Zhensu Sun, Xiaoning Du, Fu Song, Shangwen Wang, and Li Li. 2024 · 2024
Closest in time.
Exploring Parameter-Efficient Fine-Tuning Techniques for Code Generation with Large Language Models
Martin Weyssow, Xin Zhou, Kisub Kim, David Lo, and Houari Sahraoui. 2024 · 2024
Closest in time.
Stealthy Backdoor Attack for Code Models
Zhou Yang, Bowen Xu, Jie M. Zhang, Hong Jin Kang, Jieke Shi, Junda He, and David Lo. [n. d.] · 2024
Closest in time.
Android in the Zoo: Chain-of-Action-Thought for GUI Agents
Jiwen Zhang, Jihao Wu, Yihua Teng, Minghui Liao, Nuo Xu, Xiao Xiao, Zhongyu Wei, and Duyu Tang. 2024 · 2024
Closest in time.
Astraios: Parameter-Efficient Instruction Tuning Code Large Language Models
Terry Yue Zhuo, Armel Zebaze, Nitchakarn Suppattarachai, Leandro von Werra, Harm de Vries, Qian Liu, and Niklas Muennighoff. 2024 · 2024
Closest in time.