Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have shown remarkable potential across various domains, including cybersecurity.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, D. Amodei, Language Models are Few-Shot Learners, in: Advances in Neural Information Processing Systems, Vol. 33, Curran Associates, Inc., 2020, pp. 1877–1901
1901
Earlier work this paper cites.
doi:10.1007/BF00992698
C. J. C. H. Watkins, P. Dayan, Q-learning, Machine Learning 8 (3) (1992) 279–292 · 1992
Earlier work this paper cites.
doi:10.1038/nature14236
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al., Human-level control through deep reinforcement learning, Nature 518 (7540) (2015) 529–533 · 2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. u. Kaiser, I. Polosukhin, Attention is all you need, in: Advances in Neural Information Processing Systems, Vol. 30, Curran Associates, Inc., 2017
2017
Earlier work this paper cites.
J. Wei, M. Bosma, V. Zhao, K. Guu, A. W. Yu, B. Lester, N. Du, A. M. Dai, Q. V. Le, Finetuned language models are zero-shot learners, in: International Conference on Learning Representations, 2022
2022
Earlier work this paper cites.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, J. Schulman, J. Hilton, F. Kelton, L. Miller, M. Simens, A. Askell, P. Welinder, P. F. Christiano, J. Leike, R. Lowe, Training language models to follow instructions with human feedback, in: Advances in Neural Information Processing Systems, Vol. 35, Curran Associates, Inc., 2022, pp. 27730–27744
2022
Earlier work this paper cites.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, W. Chen, LoRA: Low-rank adaptation of large language models, in: International Conference on Learning Representations, 2022
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. Chi, Q. V. Le, D. Zhou, Chain-of-Thought Prompting Elicits Reasoning in Large Language Models, Advances in Neural Information Processing Systems 35 (2022) 24824–24837
2022
Earlier work this paper cites.
D. Hafner, Benchmarking the spectrum of agent capabilities, in: International Conference on Learning Representations, 2022
2022
Earlier work this paper cites.
doi:10.1007/978-3-031-19842-7_21
Y. Kant, A. Ramachandran, S. Yenamandra, I. Gilitschenski, D. Batra, A. Szot, H. Agrawal, Housekeep: Tidying Virtual Households Using Commonsense Reasoning, in: Computer Vision – ECCV 2022, Lecture Notes in Computer Science, Springer Nature Switzerland, Cham, 2022, pp. 355–373 · 2022
Earlier work this paper cites.
K. Valmeekam, A. Olmo, S. Sreedharan, S. Kambhampati, Large language models still can’t plan (a benchmark for LLMs on planning and reasoning about change), in: NeurIPS 2022 Foundation Models for Decision Making Workshop, 2022
2022
Earlier work this paper cites.
T. Silver, V. Hariprasad, R. S. Shuttleworth, N. Kumar, T. Lozano-Pérez, L. P. Kaelbling, PDDL planning with pretrained large language models, in: NeurIPS 2022 Foundation Models for Decision Making Workshop, 2022
2022
Earlier work this paper cites.
doi:10.1007/978-3-031-29269-9_1
A. Kott, Autonomous Intelligent Cyber-defense Agent: Introduction and Overview, in: Autonomous Intelligent Cyber Defense Agent (AICA), Vol. 87, Springer International Publishing, Cham, 2023, pp. 1–15 · 2023
Earlier work this paper cites.
doi:10.1007/s10844-022-00738-0
M. C. Ghanem, T. M. Chen, E. G. Nepomuceno, Hierarchical reinforcement learning for efficient and effective automated penetration testing of large networks, Journal of Intelligent Information Systems 60 (2) (2023) 281–303 · 2023
Earlier work this paper cites.
doi:10.1145/3586183.3606763
J. S. Park, J. O’Brien, C. J. Cai, M. R. Morris, P. Liang, M. S. Bernstein, Generative Agents: Interactive Simulacra of Human Behavior, in: Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology, UIST ’23, Association for Computing Machinery, New York, NY, USA, 2023, pp. 1–22 · 2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Du, O. Watkins, Z. Wang, C. Colas, T. Darrell, P. Abbeel, A. Gupta, J. Andreas, Guiding Pretraining in Reinforcement Learning with Large Language Models, in: Proceedings of the 40th International Conference on Machine Learning, Honolulu, USA, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Later among the works it cites.
T. Dettmers, A. Pagnoni, A. Holtzman, L. Zettlemoyer, Qlora: Efficient finetuning of quantized llms, in: Advances in Neural Information Processing Systems, Vol. 36, Curran Associates, Inc., 2023, pp. 10088–10115
2023
Later among the works it cites.
L. Tunstall, E. Beeching, N. Lambert, N. Rajani, S. Huang, K. Rasul, A. M. Rush, T. Wolf, The alignment handbook, https://github.com/huggingface/alignment-handbook (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, T. B. Hashimoto, Alpaca: a strong, replicable instruction-following model; 2023, URL https://crfm. stanford. edu/2023/03/13/alpaca. html (2023)
2023
Cited alongside, same era.
2023
Cited alongside, same era.
S. Yao, J. Zhao, D. Yu, N. Du, I. Shafran, K. R. Narasimhan, Y. Cao, React: Synergizing reasoning and acting in language models, in: The Eleventh International Conference on Learning Representations, 2023
2023
Cited alongside, same era.
N. Shinn, F. Cassano, A. Gopinath, K. Narasimhan, S. Yao, Reflexion: language agents with verbal reinforcement learning, in: Advances in Neural Information Processing Systems, Vol. 36, Curran Associates, Inc., 2023, pp. 8634–8652
2023
Cited alongside, same era.
S. Hao, Y. Gu, H. Ma, J. Hong, Z. Wang, D. Z. Wang, Z. Hu, Reasoning with language model is planning with world model, in: NeurIPS 2023 Workshop on Generalization in Planning, 2023
2023
Cited alongside, same era.
Y. Wu, S. Y. Min, S. Prabhumoye, Y. Bisk, R. R. Salakhutdinov, A. Azaria, T. M. Mitchell, Y. Li, Spring: Studying papers and reasoning to play games, in: Advances in Neural Information Processing Systems, Vol. 36, Curran Associates, Inc., 2023, pp. 22383–22687
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
G. Wang, Y. Xie, Y. Jiang, A. Mandlekar, C. Xiao, Y. Zhu, L. Fan, A. Anandkumar, Voyager: An open-ended embodied agent with large language models, Transactions on Machine Learning Research (2024)
2024
Closest in time.
L. Wang, C. Ma, X. Feng, Z. Zhang, H. Yang, J. Zhang, Z. Chen, J. Tang, X. Chen, Y. Lin, et al., A survey on large language model based autonomous agents, Frontiers of Computer Science 18 (6) (2024) 186345
2024
Closest in time.
doi:10.5220/0012391800003636
M. Rigaki., O. Lukáš., C. Catania., S. Garcia., Out of the cage: How stochastic parrots win in cyber security environments, in: Proceedings of the 16th International Conference on Agents and Artificial Intelligence - Volume 3: ICAART, INSTICC, SciTePress, Roma, Italy, 2024, pp. 774–781 · 2024
Closest in time.
H. W. Chung, L. Hou, S. Longpre, B. Zoph, Y. Tay, W. Fedus, Y. Li, X. Wang, M. Dehghani, S. Brahma, A. Webson, S. S. Gu, Z. Dai, M. Suzgun, X. Chen, A. Chowdhery, A. Castro-Ros, M. Pellat, K. Robinson, D. Valter, S. Narang, G. Mishra, A. Yu, V. Zhao, Y. Huang, A. Dai, H. Yu, S. Petrov, E. H. Chi, J. Dean, J. Devlin, A. Roberts, D. Zhou, Q. V. Le, J. Wei, Scaling instruction-finetuned language models, Journal of Machine Learning Research 25 (70) (2024) 1–53
2024
Closest in time.
Z. Wang, S. Cai, G. Chen, A. Liu, X. Ma, Y. Liang, T. CraftJarvis, Describe, explain, plan and select: interactive planning with large language models enables open-world multi-task agents, in: Proceedings of the 37th International Conference on Neural Information Processing Systems, NIPS ’23, Curran Associates Inc., Red Hook, NY, USA, 2024
2024
Closest in time.
doi:10.1609/aaai.v38i18.30006
T. Silver, S. Dan, K. Srinivas, J. B. Tenenbaum, L. Kaelbling, M. Katz, Generalized planning in pddl domains with pretrained large language models, Proceedings of the AAAI Conference on Artificial Intelligence 38 (18) (2024) 20256–20264 · 2024
Closest in time.
doi:https://doi.org/10.1016/j.cose.2024.103804
R. Raman, P. Calyam, K. Achuthan, Chatgpt or bard: Who is a better certified ethical hacker?, Computers & Security 140 (2024) 103804 · 2024
Closest in time.
doi:10.57967/hf/3057
Stratosphere Research Laboratory, AIC, FEL, CTU, Netsecdata (revision 913b966) (2024) · 2024
Closest in time.
doi:10.1145/3611643.3613083
A. Happe, J. Cito, Getting pwn’d by ai: Penetration testing with large language models, in: Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering, ESEC/FSE 2023, Association for Computing Machinery, New York, NY, USA, 2023, p. 2082–2086 · 2086
Closest in time.