Fetching the paper…
Reading the bibliography…
Generative AI agents, software systems powered by Large Language Models (LLMs), are emerging as a promising approach to automate cybersecurity tasks.
The Matter of Heartbleed
Zakir Durumeric, Frank Li, James Kasten, Johanna Amann, Jethro Beekman, Mathias Payer, Nicolas Weaver, David Adrian, Vern Paxson, Michael Bailey, and J. Alex Halderman · 2014
Earlier work this paper cites.
Penetration Testing Automation Assessment Method Based on Rule Tree
Jianming Zhao, Wenli Shang, Ming Wan, and Peng Zeng · 2015
Earlier work this paper cites.
SambaCry Is Coming , 2017
Mikhail Kuzin, Yaroslav Shmelev, and Dimitry Galov · 2017
Earlier work this paper cites.
I Lead, You Help but Only with Enough Details: Understanding User Experience of Co-Creation with Artificial Intelligence
Changhoon Oh, Jungwoo Song, Jinhan Choi, Seonghyeon Kim, Sungwoo Lee, and Bongwon Suh · 2018
Earlier work this paper cites.
A Systematic Literature Review and Meta-Analysis on Artificial Intelligence in Penetration Testing and Vulnerability Assessment
Dean Richard McKinnel, Tooska Dargahi, Ali Dehghantanha, and Kim-Kwang Raymond Choo · 2019
Earlier work this paper cites.
Automated Penetration Testing Using Deep Reinforcement Learning
Zhenguo Hu, Razvan Beuran, and Yasuo Tan · 2020
Earlier work this paper cites.
Hydra , 2021
Marc van Hauser Heuse · 2021
Earlier work this paper cites.
CVE-2021-3156: Heap-Based Buffer Overflow in Sudo (Baron Samedit) , 2021
Himanshu Kathpal · 2021
Earlier work this paper cites.
Docker: Accelerated Container Application Development , 2022
2022
Earlier work this paper cites.
Spring Framework Zero-Day Remote Code Execution (Spring4Shell) Vulnerability , 2022
Bharat Jogi · 2022
Earlier work this paper cites.
The Race to the Vulnerable: Measuring the Log4j Shell Incident , 2022
Raphael Hiesgen, Marcin Nawrocki, Thomas C. Schmidt, and Matthias Wählisch · 2022
Earlier work this paper cites.
Impact and Research Challenges of Penetrating Testing and Vulnerability Assessment on Network Threat
Areej Fatima, Tahir Abbas Khan, Tamer Mohamed Abdellatif, Sidra Zulfiqar, Muhammad Asif, Waseem Safi, Hussam Al Hamadi, and Amer Hani Al-Kassem · 2023
Earlier work this paper cites.
INNES: An Intelligent Network Penetration Testing Model Based on Deep Reinforcement Learning
Qianyu Li, Miao Hu, Hao Hao, Min Zhang, and Yang Li · 2023
Earlier work this paper cites.
Generative Agents: Interactive Simulacra of Human Behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2023
Earlier work this paper cites.
LLM Powered Autonomous Agents , 2023
Lilian Weng · 2023
Earlier work this paper cites.
Getting Pwn’d by AI: Penetration Testing with Large Language Models
Andreas Happe and Jürgen Cito · 2023
Earlier work this paper cites.
AgentBench: Evaluating LLMs as Agents , 2023
Xiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu, Xuanyu Lei, Hanyu Lai, Yu Gu, Hangliang Ding, Kaiwen Men, Kejuan Yang, Shudan Zhang, Xiang Deng, Aohan Zeng, Zhengxiao Du, Chenhui Zhang, Sheng Shen, Tianjun Zhang, Yu Su, Huan Sun, Minlie Huang, Yuxiao Dong, and Jie Tang · 2023
Cited alongside, same era.
Lei Huang, Weijiang Yu, Weitao Ma, Weihong Zhong, Zhangyin Feng, Haotian Wang, Qianglong Chen, Weihua Peng, Xiaocheng Feng, Bing Qin, and Ting Liu · 2023
Cited alongside, same era.
ReAct: Synergizing Reasoning and Acting in Language Models , 2023
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2023
Cited alongside, same era.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models , 2023
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou · 2023
Cited alongside, same era.
System 2 Attention (Is Something You Might Need Too) , 2023
Jason Weston and Sainbayar Sukhbaatar · 2023
ChatGPT , 2024
OpenAI · 2024
Closest in time.
AgentQuest: A Modular Benchmark Framework to Measure Progress and Improve LLM Agents
Luca Gioacchini, Giuseppe Siracusano, Davide Sanvito, Kiril Gashteovski, David Friede, Roberto Bifulco, and Carolin Lawrence · 2024
Closest in time.
Nmap: The Network Mapper - Free Security Scanner , 2024
2024
Closest in time.
National Institute of Standards and Technology , 2024
2024
Closest in time.
Cognitive Architectures for Language Agents , 2024
Theodore R. Sumers, Shunyu Yao, Karthik Narasimhan, and Thomas L. Griffiths · 2024
Closest in time.
AutoGPT , 2024
Significant Gravitas · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Large Language Models Can Be Easily Distracted by Irrelevant Context , 2023
Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed Chi, Nathanael Schärli, and Denny Zhou · 2023
Cited alongside, same era.
Improving Zero-shot Reader by Reducing Distractions from Irrelevant Documents in Open-Domain Question Answering
Sukmin Cho, Jeongyeon Seo, Soyeong Jeong, and Jong Park · 2023
Cited alongside, same era.
Metasploit |Penetration Testing Software, Pen Testing Security , 2024
2024
Cited alongside, same era.
OWASP Nettacker |OWASP Foundation , 2024
2024
Cited alongside, same era.
A Survey on Large Language Model Based Autonomous Agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, Wayne Xin Zhao, Zhewei Wei, and Ji-Rong Wen · 2024
Cited alongside, same era.
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool , 2024
Gelei Deng, Yi Liu, Víctor Mayoral-Vilches, Peng Liu, Yuekang Li, Yuan Xu, Tianwei Zhang, Yang Liu, Martin Pinzger, and Stefan Rass · 2024
Cited alongside, same era.
Generative AI for Pentesting: The Good, the Bad, the Ugly
Eric Hilario, Sami Azam, Jawahar Sundaram, Khwaja Imran Mohammed, and Bharanidharan Shanmugam · 2024
Cited alongside, same era.
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs , 2024
Angelica Chen, Jason Phang, Alicia Parrish, Vishakh Padmakumar, Chen Zhao, Samuel R. Bowman, and Kyunghyun Cho · 2024
Closest in time.
Is Cognition and Action Consistent or Not: Investigating Large Language Model’s Personality , 2024
Yiming Ai, Zhiwei He, Ziyin Zhang, Wenhong Zhu, Hongkun Hao, Kai Yu, Lingjun Chen, and Rui Wang · 2024
Closest in time.
Human-LLM Collaborative Annotation Through Effective Verification of LLM Labels
Xinru Wang, Hannah Kim, Sajjadur Rahman, Kushan Mitra, and Zhengjie Miao · 2024
Closest in time.
Beyond Static AI Evaluations: Advancing Human Interaction Evaluations for LLM Harms and Risks , 2024
Lujain Ibrahim, Saffron Huang, Lama Ahmad, and Markus Anderljung · 2024
Closest in time.
Welcome To Instructor - Instructor , 2024
Jason Liu · 2024
Closest in time.
Welcome to Pydantic - Pydantic , 2024
Samuel Colvin · 2024
Closest in time.
Introducing OpenAI o1 , 2024
OpenAI · 2024
Closest in time.
OpenAI o1-Mini , 2024
OpenAI · 2024
Closest in time.
Reasoning Models - OpenAI , 2024
OpenAI · 2024
Closest in time.