Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have demonstrated impressive results on natural language tasks, and security researchers are beginning to employ them in both offensive and defensive systems.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 1877–1901
1901
Earlier work this paper cites.
R. Bellman, “A markovian decision process,” in Journal of Mathematics and Mechanics , vol. 6, 1957, p. 679–684
1957
Earlier work this paper cites.
M. Bishop, “About penetration testing,” IEEE Security & Privacy , vol. 5, no. 6, pp. 84–87, 2007
2007
Earlier work this paper cites.
G. Gu, J. Zhang, and W. Lee, “Botsniffer: Detecting botnet command and control channels in network traffic,” 2008
2008
Earlier work this paper cites.
A. Finn, Mastering Hyper-V Deployment . John Wiley & Sons, 2010
2010
Earlier work this paper cites.
D. Kennedy, J. O’gorman, D. Kearns, and M. Aharoni, Metasploit: the penetration tester’s guide . No Starch Press, 2011
2011
Earlier work this paper cites.
G. Jacob, R. Hund, C. Kruegel, and T. Holz, “ { \{ JACKSTRAWS } \} : Picking command and control connections from bot traffic,” in 20th USENIX Security Symposium (USENIX Security 11) , 2011
2011
Earlier work this paper cites.
S. K. Cha, T. Avgerinos, A. Rebert, and D. Brumley, “Unleashing mayhem on binary code,” in 2012 IEEE Symposium on Security and Privacy . IEEE, 2012, pp. 380–394
2012
Earlier work this paper cites.
L. Bilge, D. Balzarotti, W. Robertson, E. Kirda, and C. Kruegel, “Disclosure: detecting botnet command and control servers through large-scale netflow analysis,” in Proceedings of the 28th Annual Computer Security Applications Conference , 2012, pp. 129–138
2012
Earlier work this paper cites.
DARPA, “Darpa’s cyber grand challenge (cgc) (archived),” https://www.darpa.mil/program/cyber-grand-challenge , 2013
2013
Earlier work this paper cites.
T. Avgerinos, S. K. Cha, A. Rebert, E. J. Schwartz, M. Woo, and D. Brumley, “Automatic exploit generation,” Communications of the ACM , vol. 57, no. 2, pp. 74–84, 2014
2014
Earlier work this paper cites.
X. Qiu, S. Wang, Q. Jia, C. Xia, and Q. Xia, “An automated method of penetration testing,” in 2014 IEEE Computers, Communications and IT Applications Conference , 2014, pp. 211–216
2014
Earlier work this paper cites.
J. Zhao, W. Shang, M. Wan, and P. Zeng, “Penetration testing automation assessment method based on rule tree,” in 2015 IEEE International Conference on Cyber Technology in Automation, Control, and Intelligent Systems (CYBER) , 2015, pp. 1829–1833
2015
Earlier work this paper cites.
Y. Shoshitaishvili, R. Wang, C. Salls, N. Stephens, M. Polino, A. Dutcher, J. Grosen, S. Feng, C. Hauser, C. Kruegel et al. , “Sok:(state of) the art of war: Offensive techniques in binary analysis,” in 2016 IEEE symposium on security and privacy (SP) . IEEE, 2016, pp. 138–157
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
G. Falco, A. Viswanathan, C. Caldera, and H. Shrobe, “A master attack methodology for an ai-based automated attack planner for smart cities,” IEEE Access , vol. 6, pp. 48 360–48 373, 2018
2018
Earlier work this paper cites.
B. E. Strom, A. Applebaum, D. P. Miller, K. C. Nickels, A. G. Pennington, and C. B. Thomas, “Mitre att&ck: Design and philosophy,” in Technical report . The MITRE Corporation, 2018
2018
Earlier work this paper cites.
G. Costantino, A. La Marra, F. Martinelli, and I. Matteucci, “Candy: A social engineering attack to leak information from infotainment system,” in 2018 IEEE 87th Vehicular Technology Conference (VTC Spring) . IEEE, 2018, pp. 1–5
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Lockheed Martin, “Cyber kill chain,” https://www.lockheedmartin.com/en-us/capabilities/cyber/cyber-kill-chain.html , 2019
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” Advances in Neural Information Processing Systems , vol. 33, pp. 9459–9474, 2020
2020
Earlier work this paper cites.
A. Fioraldi, D. Maier, H. Eißfeldt, and M. Heuse, “ { \{ AFL++ } \} : Combining incremental steps of fuzzing research,” in 14th USENIX Workshop on Offensive Technologies (WOOT 20) , 2020
2020
Earlier work this paper cites.
Z. Hu, R. Beuran, and Y. Tan, “Automated penetration testing using deep reinforcement learning,” in 2020 IEEE European Symposium on Security and Privacy Workshops (EuroS&PW) . IEEE, 2020, pp. 2–10
2020
Earlier work this paper cites.
S. Y. Enoch, Z. Huang, C. Y. Moon, D. Lee, M. K. Ahn, and D. S. Kim, “Harmer: Cyber-attacks automation and evaluation,” IEEE Access , vol. 8, pp. 129 397–129 414, 2020
2020
Earlier work this paper cites.
S. Liao, C. Zhou, Y. Zhao, Z. Zhang, C. Zhang, Y. Gao, and G. Zhong, “A comprehensive detection approach of nmap: Principles, rules and experiments,” in 2020 international conference on cyber-enabled distributed computing and knowledge discovery (CyberC) . IEEE, 2020, pp. 64–71
2020
Earlier work this paper cites.
Z. Jiang, J. Araki, H. Ding, and G. Neubig, “How can we know when language models know? on the calibration of language models for question answering,” Transactions of the Association for Computational Linguistics , vol. 9, pp. 962–977, 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Malik and E. Azeem, “The secrets to mimikatz-the credential dumper,” International Journal for Electronic Crime Investigation , vol. 5, no. 4, pp. 27–34, 2021
2021
Earlier work this paper cites.
J. Li, T. Tang, W. X. Zhao, J.-Y. Nie, and J.-R. Wen, “Pretrained language models for text generation: A survey,” 2022
2022
Earlier work this paper cites.
A. Fioraldi, D. C. Maier, D. Zhang, and D. Balzarotti, “Libafl: A framework to build modular and reusable fuzzers,” in Proceedings of the 2022 ACM SIGSAC Conference on Computer and Communications Security , 2022, pp. 1051–1065
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” Advances in Neural Information Processing Systems , vol. 35, pp. 24 824–24 837, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
L. Fan, G. Wang, Y. Jiang, A. Mandlekar, Y. Yang, H. Zhu, A. Tang, D.-A. Huang, Y. Zhu, and A. Anandkumar, “Minedojo: Building open-ended embodied agents with internet-scale knowledge,” in Thirty-sixth Conference on Neural Information Processing Systems Datasets and Benchmarks Track , 2022. [Online]. Available: https://openreview.net/forum?id=rc8o_j8I8PX
2022
Cited alongside, same era.
A. Sobieszek and T. Price, “Playing games with ais: the limits of gpt-3 and similar large language models,” Minds and Machines , vol. 32, no. 2, pp. 341–364, 2022
2023
Later among the works it cites.
M. Botacin, “Gpthreats-3: Is automatic malware generation a threat?” in 2023 IEEE Security and Privacy Workshops (SPW) . IEEE, 2023, pp. 238–254
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
S. Kublik and S. Saboo, GPT-3 . O’Reilly Media, Incorporated, 2022
2022
Cited alongside, same era.
A. Thudi, H. Jia, I. Shumailov, and N. Papernot, “On the necessity of auditable algorithmic definitions for machine unlearning,” in 31st USENIX Security Symposium (USENIX Security 22) , 2022, pp. 4007–4022
2022
Cited alongside, same era.
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, H. Nori, H. Palangi, M. T. Ribeiro, and Y. Zhang, “Sparks of artificial general intelligence: Early experiments with gpt-4,” 2023
2023
Cited alongside, same era.
M. Schreiner, “Gpt-4 architecture, datasets, costs and more leaked,” THE DECODER , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Yang, A. Prabhakar, S. Yao, K. Pei, and K. R. Narasimhan, “Language agents as hackers: Evaluating cybersecurity skills with capture the flag,” in Multi-Agent Security Workshop@ NeurIPS’23 , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
B. Chen, A. Paliwal, and Q. Yan, “Jailbreaker in jail: Moving target defense for large language models,” in Proceedings of the 10th ACM Workshop on Moving Target Defense , 2023, pp. 29–32
2023
Later among the works it cites.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing,” ACM Computing Surveys , vol. 55, no. 9, pp. 1–35, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
“thinkgpt,” https://github.com/jina-ai/thinkgpt , 2023
2023
Later among the works it cites.
O. Topsakal and T. C. Akinci, “Creating large language model applications utilizing langchain: A primer on developing llm apps fast,” in Proceedings of the International Conference on Applied Engineering and Natural Sciences, Konya, Turkey , 2023, pp. 10–12
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Chen, Q. Fu, Y. Yuan, Z. Wen, G. Fan, D. Liu, D. Zhang, Z. Li, and Y. Xiao, “Hallucination detection: Robustly discerning reliable answers in large language models,” in Proceedings of the 32nd ACM International Conference on Information and Knowledge Management , 2023, pp. 245–255
2023
Later among the works it cites.
J. Li, X. Cheng, W. X. Zhao, J.-Y. Nie, and J.-R. Wen, “Helma: A large-scale hallucination evaluation benchmark for large language models,” 2023
2023
Later among the works it cites.
S. McLean, G. J. Read, J. Thompson, C. Baber, N. A. Stanton, and P. M. Salmon, “The risks associated with artificial general intelligence: A systematic review,” Journal of Experimental & Theoretical Artificial Intelligence , vol. 35, no. 5, pp. 649–663, 2023
2023
Later among the works it cites.
A. Sarabi, T. Yin, and M. Liu, “An llm-based framework for fingerprinting internet-connected devices,” in Proceedings of the 2023 ACM on Internet Measurement Conference , 2023, pp. 478–484
2023
Later among the works it cites.
“What is microsoft security copilot?” https://learn.microsoft.com/en-us/security-copilot/microsoft-security-copilot , Oct. 2023, accessed: 2024-01-24
2024
Closest in time.
R. Meng, M. Mirchev, M. Böhme, and A. Roychoudhury, “Large language model guided protocol fuzzing,” in Proceedings of the 31st Annual Network and Distributed System Security Symposium (NDSS) , 2024
2024
Closest in time.
G. Deng, Y. Liu, Y. Li, K. Wang, Y. Zhang, Z. Li, H. Wang, T. Zhang, and Y. Liu, “Masterkey: Automated jailbreak across multiple large language model chatbots,” in The Network and Distributed System Security Symposium (NDSS) , vol. 2023, 2024
2024
Closest in time.
2024
Closest in time.
Y. Yang, Q. Zhang, C. Li, D. S. Marta, N. Batool, and J. Folkesson, “Human-centric autonomous systems with llms for user command reasoning,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2024, pp. 988–994
2024
Closest in time.
A. Happe and J. Cito, “Getting pwn’d by ai: Penetration testing with large language models,” in Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2023, pp. 2082–2086
2086
Closest in time.