Fetching the paper…
Reading the bibliography…
LLMs have becoming increasingly powerful, both in their benign and malicious uses.
Common vulnerabilities and exposures
Common Vulnerabilities · 2005
Earlier work this paper cites.
A classification of sql-injection attacks and countermeasures
William G Halfond, Jeremy Viegas, Alessandro Orso, et al · 2006
Earlier work this paper cites.
A research agenda acknowledging the persistence of passwords
Cormac Herley and Paul Van Oorschot · 2011
Earlier work this paper cites.
Metasploit: the penetration tester’s guide
David Kennedy, Jim O’gorman, Devon Kearns, and Mati Aharoni · 2011
Earlier work this paper cites.
Practical malware analysis: the hands-on guide to dissecting malicious software
Michael Sikorski and Andrew Honig · 2012
Earlier work this paper cites.
Owasp zed attack proxy
Simon Bennetts · 2013
Earlier work this paper cites.
The basics of hacking and penetration testing: ethical hacking and penetration testing made easy
Patrick Engebretson · 2013
Earlier work this paper cites.
Path sensitive static analysis of web applications for remote code execution vulnerability detection
Yunhui Zheng and Xiangyu Zhang · 2013
Earlier work this paper cites.
A survey of emerging threats in cybersecurity
Julian Jang-Jaccard and Surya Nepal · 2014
Earlier work this paper cites.
Burp Suite Essentials
Akash Mahajan · 2014
Earlier work this paper cites.
A hacking of more than $50 million dashes hopes in the world of virtual currency
Nathaniel Popper · 2016
Earlier work this paper cites.
Acidrain: Concurrency-related attacks on database-backed web applications
Todd Warszawski and Peter Bailis · 2017
Earlier work this paper cites.
Automated vulnerability detection in source code using deep representation learning
Rebecca Russell, Louis Kim, Lei Hamilton, Tomo Lazovich, Jacob Harer, Onur Ozdemir, Paul Ellingwood, and Marc McConley · 2018
Earlier work this paper cites.
Data exfiltration: A review of external attack vectors and countermeasures
Faheem Ullah, Matthew Edwards, Rajiv Ramdhany, Ruzanna Chitchyan, M Ali Babar, and Awais Rashid · 2018
Earlier work this paper cites.
Machine learning in cybersecurity: A review
Anand Handa, Ashu Sharma, and Sandeep K Shukla · 2019
Earlier work this paper cites.
Exploiting the remote server access support of coap protocol
Annie Gilda Roselin, Priyadarsi Nanda, Surya Nepal, Xiangjian He, and Jarod Wright · 2019
Earlier work this paper cites.
The social and psychological impact of cyberattacks
Maria Bada and Jason RC Nurse · 2020
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Cited alongside, same era.
Will ai make cyber swords or shields?
Andrew Lohn and Krystal Jackson · 2022
Cited alongside, same era.
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, et al · 2022
Cited alongside, same era.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Introduction to llm agents
Tanay Varshney · 2023
Later among the works it cites.
Openchat: Advancing open-source language models with mixed-quality data
Guan Wang, Sijie Cheng, Xianyuan Zhan, Xiangang Li, Sen Song, and Yang Liu · 2023
Later among the works it cites.
Shadow alignment: The ease of subverting safely-aligned language models
Xianjun Yang, Xiao Wang, Qi Zhang, Linda Petzold, William Yang Wang, Xun Zhao, and Dahua Lin · 2023
Later among the works it cites.
Benchmarking and defending against indirect prompt injection attacks on large language models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Emergent autonomous scientific research capabilities of large language models
Daniil A Boiko, Robert MacKnight, and Gabe Gomes · 2023
Cited alongside, same era.
Augmenting large language models with chemistry tools
Andres M Bran, Sam Cox, Oliver Schilter, Carlo Baldassari, Andrew White, and Philippe Schwaller · 2023
Cited alongside, same era.
Getting pwn’d by ai: Penetration testing with large language models
Andreas Happe and Jürgen Cito · 2023
Cited alongside, same era.
Agentcoder: Multi-agent-based code generation with iterative testing and optimisation
Dong Huang, Qingwen Bu, Jie M Zhang, Michael Luck, and Heming Cui · 2023
Cited alongside, same era.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al · 2023
Cited alongside, same era.
Swe-bench: Can language models resolve real-world github issues?
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan · 2023
Cited alongside, same era.
Exploiting programmatic behavior of llms: Dual-use through standard security attacks
Daniel Kang, Xuechen Li, Ion Stoica, Carlos Guestrin, Matei Zaharia, and Tatsunori Hashimoto · 2023
Cited alongside, same era.
Jingwei Yi, Yueqi Xie, Bin Zhu, Keegan Hines, Emre Kiciman, Guangzhong Sun, Xing Xie, and Fangzhao Wu · 2023
Later among the works it cites.
Removing rlhf protections in gpt-4 via fine-tuning
Qiusi Zhan, Richard Fang, Rohan Bindu, Akul Gupta, Tatsunori Hashimoto, and Daniel Kang · 2023
Later among the works it cites.
Universal and transferable adversarial attacks on aligned language models
Andy Zou, Zifan Wang, J Zico Kolter, and Matt Fredrikson · 2023
Later among the works it cites.
Llm agents can autonomously hack websites, 2024
Richard Fang, Rohan Bindu, Akul Gupta, Qiusi Zhan, and Daniel Kang · 2024
Closest in time.
Generative ai for pentesting: the good, the bad, the ugly
Eric Hilario, Sami Azam, Jawahar Sundaram, Khwaja Imran Mohammed, and Bharanidharan Shanmugam · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al · 2024
Closest in time.
Evaluating frontier models for dangerous capabilities
Mary Phuong, Matthew Aitchison, Elliot Catt, Sarah Cogan, Alexandre Kaskasoli, Victoria Krakovna, David Lindner, Matthew Rahtz, Yannis Assael, Sarah Hodkinson, et al · 2024
Closest in time.
Nous hermes 2 - yi-34b, 2024
Nous Research · 2024
Closest in time.
Are emergent abilities of large language models a mirage?
Rylan Schaeffer, Brando Miranda, and Sanmi Koyejo · 2024
Closest in time.
Openhermes 2.5 - mistral 7b, 2024
Teknium · 2024
Closest in time.
Tdag: A multi-agent framework based on dynamic task decomposition and agent generation
Yaoxiang Wang, Zhiyong Wu, Junfeng Yao, and Jinsong Su · 2024
Closest in time.
Injecagent: Benchmarking indirect prompt injections in tool-integrated large language model agents
Qiusi Zhan, Zhixiang Liang, Zifan Ying, and Daniel Kang · 2024
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2024
Closest in time.