Fetching the paper…
Reading the bibliography…
Integrating LLM and reinforcement learning (RL) agent effectively to achieve complementary performance is critical in high stake tasks like cybersecurity operations.
Dual processes in reasoning?
Peter C Wason and J St BT Evans · 1974
Earlier work this paper cites.
Thinking, fast and slow
Daniel Kahneman · 2011
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Trusting artificial intelligence in cybersecurity is a double-edged sword
Mariarosaria Taddeo, Tom McCutcheon, and Luciano Floridi · 2019
Earlier work this paper cites.
Cyborg: A gym for the development of autonomous cyber agents
Maxwell Standen, Martin Lucas, David Bowman, Toby J Richer, Junae Kim, and Damian Marriott · 2021
Earlier work this paper cites.
Cyberbattlesim
Microsoft Defender Research Team · 2021
Earlier work this paper cites.
Eager: Asking and answering questions for automatic reward shaping in language-guided rl
Thomas Carta, Pierre-Yves Oudeyer, Olivier Sigaud, and Sylvain Lamprier · 2022
Earlier work this paper cites.
The secret life of software vulnerabilities: A large-scale empirical study
Emanuele Iannone, Roberta Guadagni, Filomena Ferrucci, Andrea De Lucia, and Fabio Palomba · 2022
Earlier work this paper cites.
Retroformer: Pushing the limits of end-to-end retrosynthesis transformer
Yue Wan, Chang-Yu Hsieh, Ben Liao, and Shengyu Zhang · 2022
Earlier work this paper cites.
Tarek Ali and Panos Kostakos · 2023
Earlier work this paper cites.
Gpthreats-3: Is automatic malware generation a threat?
Marcus Botacin · 2023
Earlier work this paper cites.
Do as i can, not as i say: Grounding language in robotic affordances
Anthony Brohan, Yevgen Chebotar, Chelsea Finn, Karol Hausman, Alexander Herzog, Daniel Ho, Julian Ibarz, Alex Irpan, Eric Jang, Ryan Julian, et al · 2023
Earlier work this paper cites.
Can llm-generated misinformation be detected?
Canyu Chen and Kai Shu · 2023
Earlier work this paper cites.
Collaborating with language models for embodied reasoning
Ishita Dasgupta, Christine Kaeser-Chen, Kenneth Marino, Arun Ahuja, Sheila Babayan, Felix Hill, and Rob Fergus · 2023
Earlier work this paper cites.
Pentestgpt: An llm-empowered automatic penetration testing tool
Gelei Deng, Yi Liu, Víctor Mayoral-Vilches, Peng Liu, Yuekang Li, Yuan Xu, Tianwei Zhang, Yang Liu, Martin Pinzger, and Stefan Rass · 2023
Cited alongside, same era.
Self-collaboration code generation via chatgpt
Yihong Dong, Xue Jiang, Zhi Jin, and Ge Li · 2023
Cited alongside, same era.
Guiding pretraining in reinforcement learning with large language models
Yuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas, Trevor Darrell, Pieter Abbeel, Abhishek Gupta, and Jacob Andreas · 2023
Cited alongside, same era.
Decoding the threat landscape: Chatgpt, fraudgpt, and wormgpt in social engineering attacks
Polra Victor Falade · 2023
Cited alongside, same era.
Metagpt: Meta programming for multi-agent collaborative framework
Examining zero-shot vulnerability repair with large language models
Hammond Pearce, Benjamin Tan, Baleegh Ahmad, Ramesh Karri, and Brendan Dolan-Gavitt · 2023
Later among the works it cites.
Loggpt: Exploring chatgpt for log-based anomaly detection
Jiaxing Qi, Shaohan Huang, Zhongzhi Luan, Carol Fung, Hailong Yang, and Depei Qian · 2023
Later among the works it cites.
Communicative agents for software development
Chen Qian, Xin Cong, Cheng Yang, Weize Chen, Yusheng Su, Juyuan Xu, Zhiyuan Liu, and Maosong Sun · 2023
Later among the works it cites.
Lost at c: A user study on the security implications of large language model code assistants
Gustavo Sandoval, Hammond Pearce, Teo Nys, Ramesh Karri, Siddharth Garg, and Brendan Dolan-Gavitt · 2023
Later among the works it cites.
Multi-agent collaboration: Harnessing the power of intelligent llm agents
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sirui Hong, Xiawu Zheng, Jonathan Chen, Yuheng Cheng, Jinlin Wang, Ceyao Zhang, Zili Wang, Steven Ka Shing Yau, Zijuan Lin, Liyang Zhou, et al · 2023
Cited alongside, same era.
Enabling intelligent interactions between an agent and an llm: A reinforcement learning approach
Bin Hu, Chenyang Zhao, Pu Zhang, Zihao Zhou, Yuanhang Yang, Zenglin Xu, and Bin Liu · 2023
Cited alongside, same era.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung · 2023
Cited alongside, same era.
Reward design with language models
Minae Kwon, Sang Michael Xie, Kalesha Bullard, and Dorsa Sadigh · 2023
Cited alongside, same era.
Camel: Communicative agents for ”mind” exploration of large language model society
Guohao Li, Hasan Abed Al Kader Hammoud, Hani Itani, Dmitrii Khizbullin, and Bernard Ghanem · 2023
Cited alongside, same era.
Swiftsage: A generative agent with fast and slow thinking for complex interactive tasks
Bill Yuchen Lin, Yicheng Fu, Karina Yang, Faeze Brahman, Shiyu Huang, Chandra Bhagavatula, Prithviraj Ammanabrolu, Yejin Choi, and Xiang Ren · 2023
Cited alongside, same era.
Eureka: Human-level reward design via coding large language models
Yecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang, Osbert Bastani, Dinesh Jayaraman, Yuke Zhu, Linxi Fan, and Anima Anandkumar · 2023
Cited alongside, same era.
Harnessing gpt-4 for generation of cybersecurity grc policies: A focus on ransomware attack mitigation
Timothy McIntosh, Tong Liu, Teo Susnjak, Hooman Alavizadeh, Alex Ng, Raza Nowrozy, and Paul Watters · 2023
Cited alongside, same era.
Yashar Talebirad and Amirhossein Nadiri · 2023
Later among the works it cites.
Automated cyber defence: A review, 2023
Sanyam Vyas, John Hannay, Andrew Bolton, and Professor Pete Burnap · 2023
Later among the works it cites.
A survey on large language model based autonomous agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, et al · 2023
Later among the works it cites.
Humanoid agents: Platform for simulating human-like generative agents
Zhilin Wang, Yu Ying Chiu, and Yu Cheung Chiu · 2023
Later among the works it cites.
Jarvis-1: Open-world multi-task agents with memory-augmented multimodal language models
Zihao Wang, Shaofei Cai, Anji Liu, Yonggang Jin, Jinbing Hou, Bowei Zhang, Haowei Lin, Zhaofeng He, Zilong Zheng, Yaodong Yang, et al · 2023
Later among the works it cites.
Autogen: Enabling next-gen llm applications via multi-agent conversation framework
Qingyun Wu, Gagan Bansal, Jieyu Zhang, Yiran Wu, Shaokun Zhang, Erkang Zhu, Beibin Li, Li Jiang, Xiaoyun Zhang, and Chi Wang · 2023
Later among the works it cites.
Universal fuzzing via large language models
Chunqiu Steven Xia, Matteo Paltenghi, Jia Le Tian, Michael Pradel, and Lingming Zhang · 2023
Later among the works it cites.
Exploring large language models for communication games: An empirical study on werewolf
Yuzhuang Xu, Shuo Wang, Peng Li, Fuwen Luo, Xiaolong Wang, Weidong Liu, and Yang Liu · 2023
Later among the works it cites.
Agent SCA: Advanced Physical Side Channel Analysis Agent with LLMs
Ferhat Yaman · 2023
Later among the works it cites.
A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Yifan Yao, Jinhao Duan, Kaidi Xu, Yuanfang Cai, Eric Sun, and Yue Zhang · 2023
Later among the works it cites.
Exploring collaboration mechanisms for llm agents: A social psychology view
Jintian Zhang, Xin Xu, and Shumin Deng · 2023
Later among the works it cites.