Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have strong capabilities in code comprehension, but fine-tuning costs and semantic alignment issues limit their project-specific optimization; conversely, code models such CodeBERT are easy to fine-tune, but it is often difficult to learn vulnerability semantics from complex code languages.
Bidirectional recurrent neural networks
Mike Schuster and Kuldip K Paliwal. 1997 · 1997
Earlier work this paper cites.
SMOTE: synthetic minority over-sampling technique
Nitesh V Chawla, Kevin W Bowyer, Lawrence O Hall, and W Philip Kegelmeyer. 2002 · 2002
Earlier work this paper cites.
Guide for security-focused configuration management of information systems
Arnold Johnson, Kelley Dempsey, Ron Ross, Sarbari Gupta, Dennis Bailey, et al · 2011
Earlier work this paper cites.
A survey of emerging threats in cybersecurity
Julian Jang-Jaccard and Surya Nepal. 2014 · 2014
Earlier work this paper cites.
Pattern-Based Vulnerability Discovery
Fabian Yamaguchi. 2015 · 2015
Earlier work this paper cites.
Gated Graph Sequence Neural Networks. In Proceedings of ICLR’16
Yujia Li, Richard Zemel, Marc Brockschmidt, and Daniel Tarlow. 2016 · 2016
Earlier work this paper cites.
Automatic feature learning for vulnerability prediction
Hoa Khanh Dam, Truyen Tran, Trang Pham, Shien Wee Ng, John Grundy, and Aditya Ghose. 2017 · 2017
Earlier work this paper cites.
POSTER: Vulnerability discovery with function representation learning from unlabeled projects. In Proceedings of the 2017 ACM SIGSAC conference on computer and communications security . 2539–2541
Guanjun Lin, Jun Zhang, Wei Luo, Lei Pan, and Yang Xiang. 2017 · 2017
Earlier work this paper cites.
Pattern-based methods for vulnerability discovery
Fabian Yamaguchi. 2017 · 2017
Earlier work this paper cites.
Vuldeepecker: A deep learning-based system for vulnerability detection
Zhen Li, Deqing Zou, Shouhuai Xu, Xinyu Ou, Hai Jin, Sujuan Wang, Zhijun Deng, and Yuyi Zhong. 2018 · 2018
Earlier work this paper cites.
Spt-code: Sequence-to-sequence pre-training for learning source code representations. In Proceedings of the 44th International Conference on Software Engineering . 2006–2018
Changan Niu, Chuanyi Li, Vincent Ng, Jidong Ge, Liguo Huang, and Bin Luo. 2022 · 2018
Earlier work this paper cites.
Automated vulnerability detection in source code using deep representation learning. In 2018 17th IEEE international conference on machine learning and applications (ICMLA) . IEEE, 757–762
Rebecca Russell, Louis Kim, Lei Hamilton, Tomo Lazovich, Jacob Harer, Onur Ozdemir, Paul Ellingwood, and Marc McConley. 2018 · 2018
Earlier work this paper cites.
VulSniper: Focus Your Attention to Shoot Fine-Grained Vulnerabilities.. In IJCAI . 4665–4671
Xu Duan, Jingzheng Wu, Shouling Ji, Zhiqing Rui, Tianyue Luo, Mutian Yang, and Yanjun Wu. 2019 · 2019
Earlier work this paper cites.
Metric learning for adversarial robustness
Chengzhi Mao, Ziyuan Zhong, Junfeng Yang, Carl Vondrick, and Baishakhi Ray. 2019 · 2019
Earlier work this paper cites.
Yaqin Zhou, Shangqing Liu, Jingkai Siow, Xiaoning Du, and Yang Liu. 2019 · 2019
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages. In Findings of the Association for Computational Linguistics: EMNLP 2020 . 1536–1547
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Cited alongside, same era.
GraphCodeBERT: Pre-training Code Representations with Data Flow. In International Conference on Learning Representations
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, LIU Shujie, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu, et al · 2020
Cited alongside, same era.
Learning and evaluating contextual embedding of source code. In International conference on machine learning . PMLR, 5110–5121
Aditya Kanade, Petros Maniatis, Gogul Balakrishnan, and Kensen Shi. 2020 · 2020
Cited alongside, same era.
Syntax-BERT: Improving Pre-trained Transformers with Syntax Trees. In Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume . 3011–3020
Jiangang Bai, Yujing Wang, Yiren Chen, Yaming Yang, Jing Bai, Jing Yu, and Yunhai Tong. 2021 · 2021
Cited alongside, same era.
Linevul: A transformer-based line-level vulnerability prediction. In Proceedings of the 19th International Conference on Mining Software Repositories . 608–620
Michael Fu and Chakkrit Tantithamthavorn. 2022 · 2022
Later among the works it cites.
UniXcoder: Unified Cross-Modal Pre-training for Code Representation. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 7212–7225
Daya Guo, Shuai Lu, Nan Duan, Yanlin Wang, Ming Zhou, and Jian Yin. 2022 · 2022
Later among the works it cites.
LineVD: Statement-level vulnerability detection using graph neural networks. In Proceedings of the 19th International Conference on Mining Software Repositories . 596–607
David Hin, Andrey Kan, Huaming Chen, and M Ali Babar. 2022 · 2022
Later among the works it cites.
ReGVD: Revisiting graph neural networks for vulnerability detection. In Proceedings of the ACM/IEEE 44th International Conference on Software Engineering: Companion Proceedings . 178–182
Van-Anh Nguyen, Dai Quoc Nguyen, Van Nguyen, Trung Le, Quan Hung Tran, and Dinh Phung. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep learning based vulnerability detection: Are we there yet
Saikat Chakraborty, Rahul Krishna, Yangruibo Ding, and Baishakhi Ray. 2021 · 2021
Cited alongside, same era.
Deepwukong: Statically detecting software vulnerabilities using deep graph neural network
Xiao Cheng, Haoyu Wang, Jiayi Hua, Guoai Xu, and Yulei Sui. 2021 · 2021
Cited alongside, same era.
DOBF: A deobfuscation pre-training objective for programming languages
Marie-Anne Lachaux, Baptiste Roziere, Marc Szafraniec, and Guillaume Lample. 2021 · 2021
Cited alongside, same era.
Sysevr: A framework for using deep learning to detect software vulnerabilities
Zhen Li, Deqing Zou, Shouhuai Xu, Hai Jin, Yawei Zhu, and Zhaoxuan Chen. 2021b · 2021
Cited alongside, same era.
Traceability transformed: Generating more accurate links with pre-trained bert models. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 324–335
Jinfeng Lin, Yalin Liu, Qingkai Zeng, Meng Jiang, and Jane Cleland-Huang. 2021 · 2021
Cited alongside, same era.
DeepTective: Detection of PHP vulnerabilities using hybrid graph neural networks. In Proceedings of the 36th annual ACM symposium on applied computing . 1687–1690
Rishi Rabheru, Hazim Hanif, and Sergio Maffeis. 2021 · 2021
Cited alongside, same era.
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . 8696–8708
Yue Wang, Weishi Wang, Shafiq Joty, and Steven CH Hoi. 2021 · 2021
Cited alongside, same era.
Vu1SPG: Vulnerability detection based on slice property graph representation learning. In 2021 IEEE 32nd International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 457–467
Weining Zheng, Yuan Jiang, and Xiaohong Su. 2021 · 2021
Cited alongside, same era.
Transformer-based language models for software vulnerability detection. In Proceedings of the 38th Annual Computer Security Applications Conference . 481–496
Chandra Thapa, Seung Ick Jang, Muhammad Ejaz Ahmed, Seyit Camtepe, Josef Pieprzyk, and Surya Nepal. 2022 · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Later among the works it cites.
VulCNN: An image-inspired scalable vulnerability detection system. In Proceedings of the 44th International Conference on Software Engineering . 2365–2376
Yueming Wu, Deqing Zou, Shihan Dou, Wei Yang, Duo Xu, and Hai Jin. 2022 · 2022
Later among the works it cites.
ChatGPT for Vulnerability Detection, Classification, and Repair: How Far Are We?
Michael Fu, Chakkrit Tantithamthavorn, Van Nguyen, and Trung Le. 2023 · 2023
Later among the works it cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al · 2023
Later among the works it cites.
An empirical study of deep learning models for vulnerability detection. In 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2237–2248
Benjamin Steenhoek, Md Mahbubur Rahman, Richard Jiles, and Wei Le. 2023a · 2023
Later among the works it cites.
Do Language Models Learn Semantics of Code? A Case Study in Vulnerability Detection
Benjamin Steenhoek, Md Mahbubur Rahman, Shaila Sharmin, and Wei Le. 2023b · 2023
Later among the works it cites.
Vulnerability Detection by Learning from Syntax-Based Execution Paths of Code
Junwei Zhang, Zhongxin Liu, Xing Hu, Xin Xia, and Shanping Li. 2023 · 2023
Later among the works it cites.
Traced: Execution-aware pre-training for source code. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering ICSE 2024 . 1–12
Yangruibo Ding, Benjamin Steenhoek, Kexin Pei, Gail Kaiser, Wei Le, and Baishakhi Ray. [n. d.] · 2024
Closest in time.
Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Language Models
Qingkai Min, Qipeng Guo, Xiangkun Hu, Songfang Huang, Zheng Zhang, and Yue Zhang. 2024 · 2024
Closest in time.