Fetching the paper…
Reading the bibliography…
The ability to plan a course of action that achieves a desired state of affairs has long been considered a core competence of intelligent agents and has been an integral part of AI research since its inception.
Course of action generation for cyber security using classical planning
Mark S Boddy, Johnathan Gohde, Thomas Haigh, and Steven A Harp · 2005
Earlier work this paper cites.
The fast downward planning system
Malte Helmert · 2006
Earlier work this paper cites.
GPT3-to-plan: Extracting plans from text using GPT-3
Alberto Olmo, Sarath Sreedharan, and Subbarao Kambhampati · 2021
Earlier work this paper cites.
Llm+ p: Empowering large language models with optimal planning proficiency
Bo Liu, Yuqian Jiang, Xiaohan Zhang, Qiang Liu, Shiqi Zhang, Joydeep Biswas, and Peter Stone · 2023
Earlier work this paper cites.
On the planning abilities of large language models–a critical investigation
Karthik Valmeekam, Matthew Marquez, Sarath Sreedharan, and Subbarao Kambhampati · 2023
Earlier work this paper cites.
Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change
Karthik Valmeekam, Matthew Marquez, Alberto Olmo, Sarath Sreedharan, and Subbarao Kambhampati · 2024
Earlier work this paper cites.
Openai o1 system card
OpenAI · 2024
Earlier work this paper cites.
My (pure) speculation about what openai o1 might be doing, 2024
Subbarao Kambhampati · 2024
Earlier work this paper cites.
Introducing openai o1-preview, 2024
OpenAI · 2024
Cited alongside, same era.
Ban warnings fly as users dare to probe the “thoughts” of openai’s latest model
Benj Edwards · 2024
Cited alongside, same era.
Chain of thoughtlessness: An analysis of cot in planning
Kaya Stechly, Karthik Valmeekam, and Subbarao Kambhampati · 2024
Cited alongside, same era.
Can large language models reason and plan?
Subbarao Kambhampati · 2024
Cited alongside, same era.
Llms can’t plan, but can help planning in llm-modulo frameworks
Subbarao Kambhampati, Karthik Valmeekam, Lin Guan, Kaya Stechly, Mudit Verma, Siddhant Bhambri, Lucas Saldyt, and Anil Murthy · 2024
Cited alongside, same era.
Solving olympiad geometry without human demonstrations
o1 gets it right almost always, 2024
Noam Brown · 2024
Closest in time.
Thought of search: Planning with language models through the lens of efficiency, 2024
Michael Katz, Harsha Kokel, Kavitha Srinivas, and Shirin Sohrabi · 2024
Closest in time.
Sayash Kapoor, Benedikt Stroebl, Zachary S. Siegel, Nitya Nadgir, and Arvind Narayanan · 2024
Closest in time.
Mathematical discoveries from program search with large language models
Bernardino Romera-Paredes, Mohammadamin Barekatain, Alexander Novikov, Matej Balog, M Pawan Kumar, Emilien Dupont, Francisco JR Ruiz, Jordan S Ellenberg, Pengming Wang, Omar Fawzi, et al · 2024
Closest in time.
log. probs returned by openai’s api are *incredibly* unstable, 2023
Tan Zhi Xuan · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Trieu H Trinh, Yuhuai Wu, Quoc V Le, He He, and Thang Luong · 2024
Cited alongside, same era.
Guides: Reasoning, 2024
OpenAI · 2024
Cited alongside, same era.
Jian Xie, Kai Zhang, Jiangjie Chen, Tinghui Zhu, Renze Lou, Yuandong Tian, Yanghua Xiao, and Yu Su · 2024
Closest in time.
Natural plan: Benchmarking llms on natural language planning
Huaixiu Steven Zheng, Swaroop Mishra, Hugh Zhang, Xinyun Chen, Minmin Chen, Azade Nova, Le Hou, Heng-Tze Cheng, Quoc V Le, Ed H Chi, et al · 2024
Closest in time.