Fetching the paper…
Reading the bibliography…
Detecting vulnerabilities is vital for software security, yet deep learning-based vulnerability detectors (DLVD) face a data shortage, which limits their effectiveness.
Training with Noise is Equivalent to Tikhonov Regularization
Chris M. Bishop. 1995 · 1995
Earlier work this paper cites.
ANTLR: A predicated-LL(k) parser generator
T. J. Parr and R. W. Quong. 1995 · 1995
Earlier work this paper cites.
Noise modelling and evaluating learning from examples
Ray J. Hickey. 1996 · 1996
Earlier work this paper cites.
SMOTE: synthetic minority over-sampling technique
Nitesh V. Chawla, Kevin W. Bowyer, Lawrence O. Hall, and W. Philip Kegelmeyer. 2002 · 2002
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, and Ming Zhou. 2020 · 2002
Earlier work this paper cites.
The Probabilistic Relevance Framework: BM25 and Beyond
Stephen Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
Evaluation: From Precision, Recall and F-Measure to ROC, Informedness, Markedness & Correlation
David Powers. 2011 · 2011
Earlier work this paper cites.
Classification in the Presence of Label Noise: A Survey
Benoit Frenay and Michel Verleysen. 2014 · 2013
Earlier work this paper cites.
C-brain: a deep learning accelerator that tames the diversity of CNNs through adaptive data-level parallelization. In Proceedings of the 53rd Annual Design Automation Conference (Austin, Texas) (DAC ’16) . Association for Computing Machinery, New York, NY, USA, Article 123, 6 pages
Lili Song, Ying Wang, Yinhe Han, Xin Zhao, Bosheng Liu, and Xiaowei Li. 2016 · 2016
Earlier work this paper cites.
Improved Regularization of Convolutional Neural Networks with Cutout
Terrance DeVries and Graham W. Taylor. 2017 · 2017
Earlier work this paper cites.
VulDeePecker: A Deep Learning-Based System for Vulnerability Detection. In 25th Annual Network and Distributed System Security Symposium, NDSS 2018, San Diego, California, USA, February 18-21, 2018 . The Internet Society
Zhen Li, Deqing Zou, Shouhuai Xu, Xinyu Ou, Hai Jin, Sujuan Wang, Zhijun Deng, and Yuyi Zhong. 2018 · 2018
Earlier work this paper cites.
mixup: Beyond Empirical Risk Minimization. In International Conference on Learning Representations
Hongyi Zhang, Moustapha Cisse, Yann N. Dauphin, and David Lopez-Paz. 2018 · 2018
Earlier work this paper cites.
Diversity in Machine Learning
Zhiqiang Gong, Ping Zhong, and Weidong Hu. 2019 · 2019
Earlier work this paper cites.
Augmenting Data with Mixup for Sentence Classification: An Empirical Study
Hongyu Guo, Yongyi Mao, and Richong Zhang. 2019 · 2019
Earlier work this paper cites.
Deep Learning-Based Vulnerable Function Detection: A Benchmark. In Information and Communications Security: 21st International Conference, ICICS 2019, Beijing, China, December 15–17, 2019, Revised Selected Papers (Beijing, China). Springer-Verlag, Berlin, Heidelberg, 219–232
Guanjun Lin, Wei Xiao, Jun Zhang, and Yang Xiang. 2019 · 2019
Earlier work this paper cites.
Impact of Discretization Noise of the Dependent Variable on Machine Learning Classifiers in Software Engineering
Gopi Krishnan Rajbahadur, Shaowei Wang, Yasutaka Kamei, and Ahmed E. Hassan. 2021 · 2019
Earlier work this paper cites.
Yaqin Zhou, Shangqing Liu, Jingkai Siow, Xiaoning Du, and Yang Liu. 2019 · 2019
Earlier work this paper cites.
Deep Learning Based Vulnerability Detection: Are We There Yet?
Saikat Chakraborty, Rahul Krishna, Yangruibo Ding, and Baishakhi Ray. 2020 · 2020
Earlier work this paper cites.
A C/C++ Code Vulnerability Dataset with Code Changes and CVE Summaries. In Proceedings of the 17th International Conference on Mining Software Repositories (Seoul, Republic of Korea) (MSR ’20) . Association for Computing Machinery, New York, NY, USA, 508–512
Jiahao Fan, Yi Li, Shaohua Wang, and Tien N. Nguyen. 2020 · 2020
Earlier work this paper cites.
Self-Supervised Contrastive Learning for Code Retrieval and Summarization via Semantic-Preserving Transformations. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (Virtual Event, Canada) (SIGIR ’21) . Association for Computing Machinery, New York, NY, USA, 511–521
Nghi D. Q. Bui, Yijun Yu, and Lingxiao Jiang. 2021 · 2021
Cited alongside, same era.
Vulnerability detection with fine-grained interpretations. In Proceedings of the 29th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering (Athens, Greece) (ESEC/FSE 2021) . Association for Computing Machinery, New York, NY, USA, 292–303
Yi Li, Shaohua Wang, and Tien N. Nguyen. 2021 · 2021
Cited alongside, same era.
The Impact of Feature Importance Methods on the Interpretation of Defect Classifiers
Gopi Krishnan Rajbahadur, Shaowei Wang, Gustavo A. Oliva, Yasutaka Kamei, and Ahmed E. Hassan. 2022 · 2021
Cited alongside, same era.
Learning Structural Edits via Incremental Tree Transformations. In International Conference on Learning Representations
VULGEN: Realistic Vulnerability Generation Via Pattern Mining and Deep Learning. In Proceedings of the 45th International Conference on Software Engineering (Melbourne, Victoria, Australia) (ICSE ’23) . IEEE Press, 2527–2539
Yu Nong, Yuzhe Ou, Michael Pradel, Feng Chen, and Haipeng Cai. 2023 · 2023
Later among the works it cites.
Introducing ChatGPT
OpenAI. 2022 · 2023
Later among the works it cites.
An Analysis of the Automatic Bug Fixing Performance of ChatGPT. In 2023 IEEE/ACM International Workshop on Automated Program Repair (APR) . 23–30
Dominik Sobania, Martin Briesch, Carol Hanna, and Justyna Petke. 2023 · 2023
Later among the works it cites.
DeepVD: Toward Class-Separation Features for Neural Network Vulnerability Detection. In Proceedings of the 45th International Conference on Software Engineering (Melbourne, Victoria, Australia) (ICSE ’23) . IEEE Press, 2249–2261
Wenbo Wang, Tien N. Nguyen, Shaohua Wang, Yi Li, Jiyuan Zhang, and Aashish Yadavally. 2023b · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ziyu Yao, Frank F. Xu, Pengcheng Yin, Huan Sun, and Graham Neubig. 2021 · 2021
Cited alongside, same era.
A Framework of Vulnerable Code Dataset Generation by Open-Source Injection. In 2021 IEEE International Conference on Artificial Intelligence and Computer Applications (ICAICA) . 1099–1103
Shasha Zhang. 2021 · 2021
Cited alongside, same era.
NatGen: generative pre-training by “naturalizing” source code. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (Singapore, Singapore) (ESEC/FSE 2022) . Association for Computing Machinery, New York, NY, USA, 18–30
Saikat Chakraborty, Toufique Ahmed, Yangruibo Ding, Premkumar T. Devanbu, and Baishakhi Ray. 2022 · 2022
Cited alongside, same era.
LineVul: A Transformer-based Line-Level Vulnerability Prediction. In 2022 IEEE/ACM 19th International Conference on Mining Software Repositories (MSR) . 608–620
Michael Fu and Chakkrit Tantithamthavorn. 2022 · 2022
Cited alongside, same era.
Exploring Representation-level Augmentation for Code Search. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , Yoav Goldberg, Zornitsa Kozareva, and Yue Zhang (Eds.). Association for Computational Linguistics, Abu Dhabi, United Arab Emirates, 4924–4936
Haochen Li, Chunyan Miao, Cyril Leung, Yanxian Huang, Yuan Huang, Hongyu Zhang, and Yanlin Wang. 2022 · 2022
Cited alongside, same era.
Generating realistic vulnerabilities via neural code editing: an empirical study. In ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering (ESEC/FSE) (ESEC/FSE 2022) . Association for Computing Machinery, New York, NY, USA
Yu Nong, Yuzhe Ou, Michael Pradel, Feng Chen, and Haipeng Cai. 2022 · 2022
Cited alongside, same era.
MultIPAs: applying program transformations to introductory programming assignments for data augmentation. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (Singapore, Singapore) (ESEC/FSE 2022) . Association for Computing Machinery, New York, NY, USA, 1657–1661
Pedro Orvalho, Mikoláš Janota, and Vasco Manquinho. 2022 · 2022
Cited alongside, same era.
Data Augmentation by Program Transformation
Shiwen Yu, Ting Wang, and Ji Wang. 2022 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
You Augment Me: Exploring ChatGPT-based Data Augmentation for Semantic Code Search. In 2023 IEEE International Conference on Software Maintenance and Evolution (ICSME) . 14–25
Yanlin Wang, Lianghong Guo, Ensheng Shi, Wenqing Chen, Jiachi Chen, Wanjun Zhong, Menghan Wang, Hui Li, Hongyu Zhang, Ziyu Lyu, and Zibin Zheng. 2023a · 2023
Later among the works it cites.
Does data sampling improve deep learning-based vulnerability detection? Yeas! and Nays!. In Proceedings of the 45th IEEE/ACM International Conference on Software Engineering (ICSE) . IEEE, 2287–2298
Xu Yang, Shaowei Wang, Yi Li, and Shaohua Wang. 2023 · 2023
Later among the works it cites.
https://github.com/VulScribeR/VulScribeR
2024 · 2024
Closest in time.
Code Search is All You Need? Improving Code Suggestions with Code Search. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering (Lisbon, Portugal) (ICSE ’24) . Association for Computing Machinery, New York, NY, USA, Article 73, 13 pages
Junkai Chen, Xing Hu, Zhenhao Li, Cuiyun Gao, Xin Xia, and David Lo. 2024 · 2024
Closest in time.
Llm agents can autonomously exploit one-day vulnerabilities
Richard Fang, Rohan Bindu, Akul Gupta, and Daniel Kang. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming – The Rise of Code Intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Y. Wu, Y. K. Li, Fuli Luo, Yingfei Xiong, and Wenfeng Liang. 2024 · 2024
Closest in time.
LLM-Powered Code Vulnerability Repair with Reinforcement Learning and Semantic Reward
Nafis Tanveer Islam, Joseph Khoury, Andrew Seong, Gonzalo De La Torre Parra, Elias Bou-Harb, and Peyman Najafirad. 2024 · 2024
Closest in time.
Enhancing Code Vulnerability Detection via Vulnerability-Preserving Data Augmentation. In Proceedings of the 25th ACM SIGPLAN/SIGBED International Conference on Languages, Compilers, and Tools for Embedded Systems . 166–177
Shangqing Liu, Wei Ma, Jian Wang, Xiaofei Xie, Ruitao Feng, and Yang Liu. 2024 · 2024
Closest in time.
GRACE: Empowering LLM-based software vulnerability detection with graph structure and in-context learning
Guilong Lu, Xiaolin Ju, Xiang Chen, Wenlong Pei, and Zhilong Cai. 2024 · 2024
Closest in time.
Llmparser: An exploratory study on using large language models for log parsing. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering . 1–13
Zeyang Ma, An Ran Chen, Dong Jae Kim, Tse-Hsun Chen, and Shaowei Wang. 2024 · 2024
Closest in time.
Mutation-based data augmentation for software defect prediction
Rui Mao, Li Zhang, and Xiaofang Zhang. 2024 · 2024
Closest in time.
VGX: Large-Scale Sample Generation for Boosting Learning-Based Software Vulnerability Analyses. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering (Lisbon, Portugal) (ICSE ’24) . Association for Computing Machinery, New York, NY, USA, Article 149, 13 pages
Yu Nong, Richard Fang, Guangbei Yi, Kunsong Zhao, Xiapu Luo, Feng Chen, and Haipeng Cai. 2024 · 2024
Closest in time.
Code with CodeQwen1.5
Qwen Team. 2024 · 2024
Closest in time.
Natural Is the Best: Model-Agnostic Code Simplification for Pre-trained Large Language Models
Yan Wang, Xiaoning Li, Tien N Nguyen, Shaohua Wang, Chao Ni, and Ling Ding. 2024 · 2024
Closest in time.
Chatgpt prompt patterns for improving code quality, refactoring, requirements elicitation, and software design
Jules White, Sam Hays, Quchen Fu, Jesse Spencer-Smith, and Douglas C Schmidt. 2024 · 2024
Closest in time.