Fetching the paper…
Reading the bibliography…
The prospect of artificial intelligence (AI) competing in the adversarial landscape of cyber security has long been considered one of the most impactful, challenging, and potentially dangerous applications of AI.
“Reflections on trusting trust”
Ken Thompson · 1984
Earlier work this paper cites.
“The Cyber Kill Chain”, https://www.lockheedmartin.com/en-us/capabilities/cyber/cyber-kill-chain.html , 2011
2011
Earlier work this paper cites.
“Artificial intelligence: a modern approach”
Stuart Russell and Peter Norvig · 2016
Earlier work this paper cites.
Greg Brockman et al · 2016
Earlier work this paper cites.
“Mastering the game of go without human knowledge”
David Silver et al · 2017
Earlier work this paper cites.
“Weird Machines, Exploitability, and Provable Unexploitability”
Thomas Dullien · 2017
Earlier work this paper cites.
“Mitre att&ck: Design and philosophy”
Blake Strom et al · 2018
Earlier work this paper cites.
“On the measure of intelligence”
François Chollet · 2019
Earlier work this paper cites.
“How much knowledge can you pack into the parameters of a language model?”
Adam Roberts, Colin Raffel and Noam Shazeer · 2020
Earlier work this paper cites.
“Chain-of-thought prompting elicits reasoning in large language models”
Jason Wei et al · 2022
Earlier work this paper cites.
“Pentestgpt: An llm-empowered automatic penetration testing tool”
Gelei Deng et al · 2023
Earlier work this paper cites.
“Language agents as hackers: Evaluating cybersecurity skills with capture the flag”
John Yang et al · 2023
Earlier work this paper cites.
Wesley Tann et al · 2023
Earlier work this paper cites.
“Purple llama cyberseceval: A secure coding benchmark for language models”
Manish Bhatt et al · 2023
Earlier work this paper cites.
“A comprehensive benchmark for evaluating cybersecurity knowledge of foundation models”, 2023
Guancheng Li et al · 2023
Earlier work this paper cites.
Zefang Liu · 2023
Earlier work this paper cites.
URL: https://aicyberchallenge.com/about/
“DARPA Artificial Intelligence Cyber Challenge (AIxCC)”, 2023 · 2023
Earlier work this paper cites.
“DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines”
Omar Khattab et al · 2023
Earlier work this paper cites.
“The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning”, 2024
Nathaniel Li et al · 2024
Earlier work this paper cites.
Shengye Wan et al · 2024
Cited alongside, same era.
“Evaluating frontier models for dangerous capabilities”
Mary Phuong et al · 2024
Cited alongside, same era.
“PenHeal: A Two-Stage LLM Framework for Automated Pentesting and Optimal Remediation”
Junjie Huang and Quanyan Zhu · 2024
Cited alongside, same era.
“Autoattacker: A large language model guided system to implement automatic cyber-attacks”
Jiacen Xu et al · 2024
Cited alongside, same era.
“Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risk of Language Models”, 2024
“Swe-agent: Agent-computer interfaces enable automated software engineering”
John Yang et al · 2024
Later among the works it cites.
“Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risk of Language Models”
Andy Zhang et al · 2024
Later among the works it cites.
“MITRE Caldera: A Scalable, Automated Adversary Emulation Platform”, 2024
MITRE · 2024
Later among the works it cites.
“The Shift from Models to Compound AI Systems”, https://bair.berkeley.edu/blog/2024/02/18/compound-ai-systems/ , 2024
Matei Zaharia et al · 2024
Later among the works it cites.
“LLMs can’t plan, but can help planning in LLM-modulo frameworks”
Subbarao Kambhampati et al · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Andy. Zhang et al · 2024
Cited alongside, same era.
“EnIGMA: Enhanced Interactive Generative Model Agent for CTF Challenges”
Talor Abramovich et al · 2024
Cited alongside, same era.
Andrey Anurin et al · 2024
Cited alongside, same era.
“An empirical evaluation of LLMs for solving offensive security challenges”
Minghao Shao et al · 2024
Cited alongside, same era.
“LLMs as hackers: Autonomous linux privilege escalation attacks”
A Happe, A Kaplan and J Cito · 2024
Cited alongside, same era.
“Got Root? A Linux Priv-Esc Benchmark”
Andreas Happe and Jürgen Cito · 2024
Cited alongside, same era.
“A survey on evaluation of large language models”
Yupeng Chang et al · 2024
Cited alongside, same era.
“CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge”
Norbert Tihanyi et al · 2024
Cited alongside, same era.
“On Memorization of Large Language Models in Logical Reasoning”
Chulin Xie et al · 2024
Later among the works it cites.
“Inspect AI: Framework for Large Language Model Evaluations”, 2024
UK AI Safety Institute · 2024
Later among the works it cites.
“SpecterOps - BloodHound CE”, https://github.com/SpecterOps/BloodHound , 2024
2024
Later among the works it cites.
“SharpHound”, https://github.com/BloodHoundAD/SharpHound , 2024
2024
Later among the works it cites.
“bloodhound.py”, https://github.com/dirkjanm/BloodHound.py , 2024
2024
Later among the works it cites.
“Neo4j”, https://neo4j.com/ , 2024
2024
Later among the works it cites.
“Mirage: cyber deception against autonomous cyber attacks in emulation and simulation”
Michael Kouremetis et al · 2024
Later among the works it cites.
“mimikatz”, 2024
Benjamin Delpy · 2024
Later among the works it cites.
“HuggingFace: Text Generation Inference”, https://github.com/huggingface/text-generation-inference , 2024
2024
Later among the works it cites.
URL: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard
“HuggingFace: Open LLM Leaderboard”, 2024 · 2024
Later among the works it cites.
“OpenAI Python API library”, 2024
OpenAI · 2024
Later among the works it cites.
URL: https://www.mitre.org/news-insights/news-release/mitre-establish-new-ai-experimentation-and-prototyping-capability-us
“MITRE to Establish New AI Experimentation and Prototyping Capability for U.S. Government Agencies”, 2024 · 2024
Later among the works it cites.
“On the Feasibility of Using LLMs to Execute Multistage Network Attacks”, 2025
Brian Singer et al · 2025
Closest in time.