Fetching the paper…
Reading the bibliography…
LLM-based autonomous agents often fail to execute complex web tasks that require dynamic interaction due to the inherent uncertainty and complexity of these environments.
A Formal Basis for the Heuristic Determination of Minimum Cost Paths
Hart, P. E.; Nilsson, N. J.; and Raphael, B. 1968 · 1968
Earlier work this paper cites.
Uncertainty-based competition between prefrontal and dorsolateral striatal systems for behavioral control
Daw, N. D.; Niv, Y.; and Dayan, P. 2005 · 2005
Earlier work this paper cites.
Efficient selectivity and backup operators in Monte-Carlo tree search
Coulom, R. 2006 · 2006
Earlier work this paper cites.
Bandit based monte-carlo planning
Kocsis, L.; and Szepesvári, C. 2006 · 2006
Earlier work this paper cites.
A survey of monte carlo tree search methods
Browne, C. B.; Powley, E.; Whitehouse, D.; Lucas, S. M.; Cowling, P. I.; Rohlfshagen, P.; Tavener, S.; Perez, D.; Samothrakis, S.; and Colton, S. 2012 · 2012
Earlier work this paper cites.
Tactical cooperative planning for autonomous highway driving using Monte-Carlo Tree Search
Lenz, D.; Kessler, T.; and Knoll, A. 2016 · 2016
Earlier work this paper cites.
Mastering the game of Go without human knowledge
Silver, D.; Schrittwieser, J.; Simonyan, K.; Antonoglou, I.; Huang, A.; Guez, A.; Hubert, T.; baker, L.; Lai, M.; Bolton, A.; Chen, Y.; Lillicrap, T. P.; Hui, F.; Sifre, L.; van den Driessche, G.; Graepel, T.; and Hassabis, D. 2017 · 2017
Earlier work this paper cites.
Reinforcement Learning on Web Interfaces using Workflow-Guided Exploration
Liu, E. Z.; Guu, K.; Pasupat, P.; Shi, T.; and Liang, P. 2018 · 2018
Earlier work this paper cites.
Monte-carlo planning and learning with language action value estimates
Jang, Y.; Seo, S.; Lee, J.; and Kim, K.-E. 2021 · 2021
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; Zhou, D.; et al. 2022 · 2022
Earlier work this paper cites.
Achiam, J.; Adler, S.; Agarwal, S.; Ahmad, L.; Akkaya, I.; Aleman, F. L.; Almeida, D.; Altenschmidt, J.; Altman, S.; Anadkat, S.; et al. 2023 · 2023
Earlier work this paper cites.
Reasoning with Language Model is Planning with World Model
Hao, S.; Gu, Y.; Ma, H.; Hong, J.; Wang, Z.; Wang, D.; and Hu, Z. 2023a · 2023
Earlier work this paper cites.
Faithful Question Answering with Monte-Carlo Planning
Hong, R.; Zhang, H.; Zhao, H.; Yu, D.; and Zhang, C. 2023 · 2023
Earlier work this paper cites.
A Zero-Shot Language Agent for Computer Control with Structured Reflection
Li, T.; Li, G.; Deng, Z.; Wang, B.; and Li, Y. 2023 · 2023
Cited alongside, same era.
Laser: Llm agent with state-space exploration for web navigation
Ma, K.; Zhang, H.; Wang, H.; Pan, X.; and Yu, D. 2023 · 2023
Cited alongside, same era.
Adapt: As-needed decomposition and planning with language models
Prasad, A.; Koller, A.; Hartmann, M.; Clark, P.; Sabharwal, A.; Bansal, M.; and Khot, T. 2023 · 2023
Cited alongside, same era.
Webwise: Web interface control and sequential exploration with large language models
Tao, H.; TV, S.; Shlapentokh-Rothman, M.; Hoiem, D.; and Ji, H. 2023 · 2023
Cited alongside, same era.
Language models can solve computer tasks
Kim, G.; Baldi, P.; and McAleer, S. 2024 · 2024
Closest in time.
Tree Search for Language Model Agents
Koh, J. Y.; McAleer, S.; Fried, D.; and Salakhutdinov, R. 2024 · 2024
Closest in time.
AutoWebGLM: Bootstrap And Reinforce A Large Language Model-based Web Navigating Agent
Lai, H.; Liu, X.; Iong, I. L.; Yao, S.; Chen, Y.; Shen, P.; Yu, H.; Zhang, H.; Zhang, X.; Dong, Y.; et al. 2024 · 2024
Closest in time.
Autonomous evaluation and refinement of digital agents
Pan, J.; Zhang, Y.; Tomlin, N.; Zhou, Y.; Levine, S.; and Suhr, A. 2024 · 2024
Closest in time.
Reflexion: Language agents with verbal reinforcement learning
Shinn, N.; Cassano, F.; Gopinath, A.; Narasimhan, K.; and Yao, S. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Team, G.; Anil, R.; Borgeaud, S.; Wu, Y.; Alayrac, J.-B.; Yu, J.; Soricut, R.; Schalkwyk, J.; Dai, A. M.; Hauth, A.; et al. 2023 · 2023
Cited alongside, same era.
PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization
Wang, X.; Li, C.; Wang, Z.; Bai, F.; Luo, H.; Zhang, J.; Jojic, N.; Xing, E. P.; and Hu, Z. 2023 · 2023
Cited alongside, same era.
The dawn of lmms: Preliminary explorations with gpt-4v (ision)
Yang, Z.; Li, L.; Lin, K.; Wang, J.; Lin, C.-C.; Liu, Z.; and Wang, L. 2023 · 2023
Cited alongside, same era.
Synapse: Trajectory-as-exemplar prompting with memory for computer control
Zheng, L.; Wang, R.; Wang, X.; and An, B. 2023 · 2023
Cited alongside, same era.
The claude 3 model family: Opus, sonnet, haiku
Anthropic, A. 2024 · 2024
Cited alongside, same era.
THOUGHTSCULPT: Reasoning with Intermediate Revision and Search
Chi, Y.; Yang, K.; and Klein, D. 2024 · 2024
Cited alongside, same era.
Mind2web: Towards a generalist agent for the web
Deng, X.; Gu, Y.; Zheng, B.; Chen, S.; Stevens, S.; Wang, B.; Sun, H.; and Su, Y. 2024 · 2024
Cited alongside, same era.
Fu, Y.; Kim, D.-K.; Kim, J.; Sohn, S.; Logeswaran, L.; Bae, K.; and Lee, H. 2024 · 2024
Cited alongside, same era.
SteP: Stacked LLM Policies for Web Actions
Sodhi, P.; Branavan, S. R. K.; Artzi, Y.; and McDonald, R. 2024 · 2024
Closest in time.
Adaplanner: Adaptive planning from feedback with language models
Sun, H.; Zhuang, Y.; Kong, L.; Dai, B.; and Zhang, C. 2024 · 2024
Closest in time.
Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
Tian, Y.; Peng, B.; Song, L.; Jin, L.; Yu, D.; Mi, H.; and Yu, D. 2024 · 2024
Closest in time.
A survey on large language model based autonomous agents
Wang, L.; Ma, C.; Feng, X.; Zhang, Z.; Yang, H.; Zhang, J.; Chen, Z.; Tang, J.; Chen, X.; Lin, Y.; et al. 2024 · 2024
Closest in time.
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
Xie, Y.; Goyal, A.; Zheng, W.; Kan, M.-Y.; Lillicrap, T. P.; Kawaguchi, K.; and Shieh, M. 2024 · 2024
Closest in time.
Zhang, D.; Huang, X.; Zhou, D.; Li, Y.; and Ouyang, W. 2024 · 2024
Closest in time.
Large language models as commonsense knowledge for large-scale task planning
Zhao, Z.; Lee, W. S.; and Hsu, D. 2024 · 2024
Closest in time.