Fetching the paper…
Reading the bibliography…
Penetration-testing is crucial for identifying system vulnerabilities, with privilege-escalation being a critical subtask to gain elevated access to protected resources.
Geer D, Harthorne J (2002) Penetration testing: a duet. In: 18th Annual Computer Security Applications Conference, 2002. Proceedings., pp 185–195, DOI
2002
Earlier work this paper cites.
Bishop M (2007) About penetration testing. IEEE Security & Privacy 5(6):84–87, DOI
2007
Earlier work this paper cites.
Sommer R, Paxson V (2010) Outside the closed world: On using machine learning for network intrusion detection. In: 2010 IEEE symposium on security and privacy, IEEE, pp 305–316
2010
Earlier work this paper cites.
Weidman G (2014) Penetration testing: a hands-on introduction to hacking. No starch press
2014
Earlier work this paper cites.
Shah S, Mehtre BM (2015) An overview of vulnerability assessment and penetration testing techniques. Journal of Computer Virology and Hacking Techniques 11:27–49
2015
Earlier work this paper cites.
Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I (2017) Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Harang R, Ducau FN (2018) Measuring the speed of the red queen’s race. BlackHat: Las Vegas, NV, USA
2018
Earlier work this paper cites.
Shebli HMZA, Beheshti BD (2018) A study on penetration testing process and tools. In: 2018 IEEE Long Island Systems, Applications and Technology Conference (LISAT), pp 1–7, DOI
2018
Earlier work this paper cites.
Singh A, Jaswal N, Agarwal M, Teixeira D (2018) Metasploit Penetration Testing Cookbook: Evade antiviruses, bypass firewalls, and exploit complex environments with the most widely used penetration testing framework. Packt Publishing Ltd
2018
Earlier work this paper cites.
Strom BE, Applebaum A, Miller DP, Nickels KC, Pennington AG, Thomas CB (2018) Mitre att&ck: Design and philosophy. In: Technical report, The MITRE Corporation
2018
Earlier work this paper cites.
Lewis P, Perez E, Piktus A, Petroni F, Karpukhin V, Goyal N, Küttler H, Lewis M, Yih Wt, Rocktäschel T, Riedel S, Kiela D (2020a) Retrieval-augmented generation for knowledge-intensive nlp tasks. In: Larochelle H, Ranzato M, Hadsell R, Balcan M, Lin H (eds) Advances in Neural Information Processing Systems, Curran Associates, Inc., vol 33, pp 9459–9474, URL
2020
Earlier work this paper cites.
Sarker IH, Kayes A, Badsha S, Alqahtani H, Watters P, Ng A (2020) Cybersecurity data science: an overview from machine learning perspective. Journal of Big data 7:1–29
2020
Earlier work this paper cites.
Bender EM, Gebru T, McMillan-Major A, Shmitchell S (2021) On the dangers of stochastic parrots: Can language models be too big? In: Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, pp 610–623
2021
Earlier work this paper cites.
Andreas J (2022) Language models as agent models. arXiv preprint arXiv:221201681
2022
Earlier work this paper cites.
Dong Q, Li L, Dai D, Zheng C, Wu Z, Chang B, Sun X, Xu J, Sui Z (2022) A survey for in-context learning. arXiv preprint arXiv:230100234
2022
Earlier work this paper cites.
He X, Yang D, Feng W, Fu TJ, Akula A, Jampani V, Narayana P, Basu S, Wang WY, Wang XE (2022) Cpl: Counterfactual prompt learning for vision and language models. arXiv preprint arXiv:221010362
2022
Earlier work this paper cites.
Wei J, Tay Y, Bommasani R, Raffel C, Zoph B, Borgeaud S, Yogatama D, Bosma M, Zhou D, Metzler D, Chi EH, Hashimoto T, Vinyals O, Liang P, Dean J, Fedus W (2022) Emergent abilities of large language models. arXiv preprint arXiv:220607682 URL
2022
Earlier work this paper cites.
Abdelnabi S, Greshake K, Mishra S, Endres C, Holz T, Fritz M (2023) Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection. URL
2023
Earlier work this paper cites.
Bubeck S, Chandrasekaran V, Eldan R, Gehrke J, Horvitz E, Kamar E, Lee P, Lee YT, Li Y, Lundberg S, Nori H, Palangi H, Ribeiro MT, Zhang Y (2023) Sparks of artificial general intelligence: Early experiments with gpt-4
2023
Earlier work this paper cites.
Community O (2023) What is the actual cutoff date for gpt-4?
2023
Earlier work this paper cites.
Dagan G, Keller F, Lascarides A (2023) Dynamic planning with a llm. arXiv preprint arXiv:230806391
2023
Earlier work this paper cites.
Dai D, Sun Y, Dong L, Hao Y, Ma S, Sui Z, Wei F (2023) Why can gpt learn in-context? language models implicitly perform gradient descent as meta-optimizers. In: ICLR 2023 Workshop on Mathematical and Empirical Understanding of Foundation Models
2023
Earlier work this paper cites.
Deng G, Liu Y, Mayoral-Vilches V, Liu P, Li Y, Xu Y, Zhang T, Liu Y, Pinzger M, Rass S (2023) Pentestgpt: An llm-empowered automatic penetration testing tool. arXiv preprint arXiv:230806782
2023
Cited alongside, same era.
Dutta TS (2023) Hackers released new black hat ai tools xxxgpt and wolf gpt
2023
Cited alongside, same era.
Gatlan S (2023) The dark side of generative ai: Five malicious llms found on the dark web
2023
Cited alongside, same era.
Gupta M, Akiri C, Aryal K, Parker E, Praharaj L (2023) From chatgpt to threatgpt: Impact of generative ai in cybersecurity and privacy. IEEE Access
2023
Cited alongside, same era.
Happe A, Cito J (2023a) Getting pwn’d by ai: Penetration testing with large language models. In: Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC/FSE 2023, DOI
Huang J, Zhu Q (2024) Penheal: A two-stage llm framework for automated pentesting and optimal remediation. arXiv preprint arXiv:240717788
2024
Closest in time.
2024
Closest in time.
Kowira EM, Suki NN, Nathan Y (2024) Automated privilege escalation enumeration and execution script for linux. In: AIP Conference Proceedings, AIP Publishing, vol 2802
2024
Closest in time.
2024
Closest in time.
Li X, Cao Y, Ma Y, Sun A (2024) Long context vs. rag for llms: An evaluation and revisits. URL
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
Happe A, Cito J (2023b) Understanding hackers’ work: An empirical study of offensive security practitioners. In: Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC/FSE 2023
2023
Cited alongside, same era.
Jin Y, Jang E, Cui J, Chung JW, Lee Y, Shin S (2023) Darkbert: A language model for the dark side of the internet. arXiv preprint arXiv:230508596
2023
Cited alongside, same era.
Kong A, Zhao S, Chen H, Li Q, Qin Y, Sun R, Zhou X, Wang E, Dong X (2023) Better zero-shot reasoning with role-play prompting. arXiv preprint arXiv:230807702
2023
Cited alongside, same era.
Kosinski M (2023) Theory of mind might have spontaneously emerged in large language models
2023
Cited alongside, same era.
Liu Y, Deng G, Li Y, Wang K, Zhang T, Liu Y, Wang H, Zheng Y, Liu Y (2023) Prompt injection attack against llm-integrated applications
2023
Cited alongside, same era.
Mascellino A (2023) Ai tool wormgpt enables convincing fake emails for bec attacks
2023
Cited alongside, same era.
Montalbano E (2023) Darkbert: Gpt-based malware trains up on the entire dark web
2023
Cited alongside, same era.
2024
Closest in time.
Mavikumbure HS, Cobilean V, Wickramasinghe CS, Drake D, Manic M (2024) Generative ai in cyber security of cyber physical systems: Benefits and threats. In: 2024 16th International Conference on Human System Interaction (HSI), pp 1–8, DOI
2024
Closest in time.
2024
Closest in time.
Renze M, Guven E (2024) Self-reflection in llm agents: Effects on problem-solving performance. arXiv preprint arXiv:240506682
2024
Closest in time.
Yao Y, Duan J, Xu K, Cai Y, Sun Z, Zhang Y (2024) A survey on large language model (llm) security and privacy: The good, the bad, and the ugly. High-Confidence Computing 4(2):100211, DOI
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Das BC, Amini MH, Wu Y (2025) Security and privacy challenges of large language models: A survey. URL
2025
Closest in time.
2025
Closest in time.
mrb3n, Cry0l1t3 (2025) Linux privilege escalation
2025
Closest in time.
munky9001 (2011) Db_autopwn deprecated! about time
2025
Closest in time.
OWASP (2013) Owasp web security testing guide
2025
Closest in time.
OWASP (2021) Owasp top 10:2021
2025
Closest in time.
OWASP (2025) Owasp application security verification standard (asvs)
2025
Closest in time.
Pinna E, Cardaci A (2025) Gtfobins
2025
Closest in time.
Polop C (2025) Hacktricks: Linux privilege escalation
2025
Closest in time.
Shahar S, Tib3rius (2018) Linux privesc
2025
Closest in time.