Fetching the paper…
Reading the bibliography…
LLM-driven autonomous agents have emerged as a promising direction in recent years.
Preliminary discussion of the logical design of an electronic computing instrument
Burks, A. W., Goldstine, H. H., and von Neumann, J · 1946
Earlier work this paper cites.
Computer: A History of the Information Machine
Campbell-Kelly, M. and Aspray, W · 1996
Earlier work this paper cites.
Crafting papers on machine learning
Langley, P · 2000
Earlier work this paper cites.
Principles of Computer System Design: An Introduction
Saltzer, J. H. and Kaashoek, M. F · 2009
Earlier work this paper cites.
Big.little processing
ARM · 2011
Earlier work this paper cites.
Computer Architecture: A Quantitative Approach
Hennessy, J. L. and Patterson, D. A · 2011
Earlier work this paper cites.
Direct memory access: Overview and applications
Khawaja, K. K. A. M. and Khan, A. A · 2014
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Lewis, P., Perez, E., Piktus, A., Petroni, F., Karpukhin, V., Goyal, N., Küttler, H., Lewis, M., Yih, W.-t., Rocktäschel, T., et al · 2020
Earlier work this paper cites.
Keep calm and explore: Language models for action generation in text-based games
Yao, S., Rao, R., Hausknecht, M., and Narasimhan, K · 2020
Earlier work this paper cites.
Webgpt: Browser-assisted question-answering with human feedback
Nakano, R., Hilton, J., Balaji, S., Wu, J., Ouyang, L., Kim, C., Hesse, C., Jain, S., Kosaraju, V., Saunders, W., et al · 2021
Earlier work this paper cites.
Do as i can, not as i say: Grounding language in robotic affordances
Ahn, M., Brohan, A., Brown, N., Chebotar, Y., Cortes, O., David, B., Finn, C., Fu, C., Gopalakrishnan, K., Hausman, K., et al · 2022
Earlier work this paper cites.
A survey on in-context learning
Dong, Q., Li, L., Dai, D., Zheng, C., Ma, J., Li, R., Xia, H., Xu, J., Wu, Z., Liu, T., et al · 2022
Earlier work this paper cites.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Fan, L., Wang, G., Jiang, Y., Mandlekar, A., Yang, Y., Zhu, H., Tang, A., Huang, D.-A., Zhu, Y., and Anandkumar, A · 2022
Earlier work this paper cites.
Self-consistency improves chain of thought reasoning in language models
Wang, X., Wei, J., Schuurmans, D., Le, Q., Chi, E., Narang, S., Chowdhery, A., and Zhou, D · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Xia, F., Chi, E., Le, Q. V., Zhou, D., et al · 2022
Earlier work this paper cites.
React: Synergizing reasoning and acting in language models
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K., and Cao, Y · 2022
Earlier work this paper cites.
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., et al · 2023
Cited alongside, same era.
Chemcrow: Augmenting large-language models with chemistry tools
Bran, A. M., Cox, S., Schilter, O., Baldassari, C., White, A. D., and Schwaller, P · 2023
Cited alongside, same era.
Llm+ p: Empowering large language models with optimal planning proficiency
Liu, B., Jiang, Y., Zhang, X., Liu, Q., Zhang, S., Biswas, J., and Stone, P · 2023
Cited alongside, same era.
Memochat: Tuning llms to use memos for consistent long-range open-domain conversation
Lu, J., An, S., Lin, M., Pergola, G., He, Y., Yin, D., Sun, X., and Wu, Y · 2023
Cited alongside, same era.
Cogagent: A visual language model for gui agents
Hong, W., Wang, W., Lv, Q., Xu, J., Yu, W., Ji, J., Wang, Y., Wang, Z., Dong, Y., Ding, M., et al · 2024
Later among the works it cites.
An embodied generalist agent in 3d world
Huang, J., Yong, S., Ma, X., Linghu, X., Li, P., Wang, Y., Li, Q., Zhu, S.-C., Jia, B., and Huang, S · 2024
Later among the works it cites.
Adaptive collaboration strategy for llms in medical decision making
Kim, Y., Park, C., Jeong, H., Chan, Y. S., Xu, X., McDuff, D., Breazeal, C., and Park, H. W · 2024
Later among the works it cites.
Showui: One vision-language-action model for gui visual agent
Lin, K. Q., Li, L., Gao, D., Yang, Z., Wu, S., Bai, Z., Lei, W., Wang, L., and Shou, M. Z · 2024
Later among the works it cites.
Agent q: Advanced reasoning and learning for autonomous ai agents
Putta, P., Mills, E., Garg, N., Motwani, S., Finn, C., Garg, D., and Rafailov, R · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Packer, C., Wooders, S., Lin, K., Fang, V., Patil, S. G., Stoica, I., and Gonzalez, J. E · 2023
Cited alongside, same era.
Generative agents: Interactive simulacra of human behavior
Park, J. S., O’Brien, J., Cai, C. J., Morris, M. R., Liang, P., and Bernstein, M. S · 2023
Cited alongside, same era.
Gorilla: Large language model connected with massive apis
Patil, S. G., Zhang, T., Wang, X., and Gonzalez, J. E · 2023
Cited alongside, same era.
Communicative agents for software development
Qian, C., Cong, X., Yang, C., Chen, W., Su, Y., Xu, J., Liu, Z., and Sun, M · 2023
Cited alongside, same era.
Toolllm: Facilitating large language models to master 16000+ real-world apis
Qin, Y., Liang, S., Ye, Y., Zhu, K., Yan, L., Lu, Y., Lin, Y., Cong, X., Tang, X., Qian, B., et al · 2023
Cited alongside, same era.
Voyager: An open-ended embodied agent with large language models
Wang, G., Xie, Y., Jiang, Y., Mandlekar, A., Xiao, C., Zhu, Y., Fan, L., and Anandkumar, A · 2023
Cited alongside, same era.
Agenttuning: Enabling generalized agent abilities for llms
Zeng, A., Liu, M., Lu, R., Wang, B., Liu, X., Dong, Y., and Tang, J · 2023
Cited alongside, same era.
Appagent: Multimodal agents as smartphone users
Zhang, C., Yang, Z., Liu, J., Han, Y., Chen, X., Huang, Z., Fu, B., and Yu, G · 2023
Cited alongside, same era.
Vlm agents generate their own memories: Distilling experience into embodied programs
Sarch, G., Jang, L., Tarr, M. J., Cohen, W. W., Marino, K., and Fragkiadaki, K · 2024
Later among the works it cites.
Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Shen, Y., Song, K., Tan, X., Li, D., Lu, W., and Zhuang, Y · 2024
Later among the works it cites.
Learning to use tools via cooperative and interactive agents
Shi, Z., Gao, S., Chen, X., Feng, Y., Yan, L., Shi, H., Yin, D., Ren, P., Verberne, S., and Ren, Z · 2024
Later among the works it cites.
Reflexion: Language agents with verbal reinforcement learning
Shinn, N., Cassano, F., Gopinath, A., Narasimhan, K., and Yao, S · 2024
Later among the works it cites.
Os-copilot: Towards generalist computer agents with self-improvement
Wu, Z., Han, C., Ding, Z., Weng, Z., Liu, Z., Yao, S., Yu, T., and Kong, L · 2024
Later among the works it cites.
Large multimodal agents: A survey
Xie, J., Chen, Z., Zhang, R., Wan, X., and Li, G · 2024
Later among the works it cites.
Swe-agent: Agent-computer interfaces enable automated software engineering
Yang, J., Jimenez, C. E., Wettig, A., Lieret, K., Yao, S., Narasimhan, K., and Press, O · 2024
Later among the works it cites.
Tree of thoughts: Deliberate problem solving with large language models
Yao, S., Yu, D., Zhao, J., Shafran, I., Griffiths, T., Cao, Y., and Narasimhan, K · 2024
Later among the works it cites.
Agent lumos: Unified and modular training for open-source language agents
Yin, D., Brahman, F., Ravichander, A., Chandu, K., Chang, K.-W., Choi, Y., and Lin, B. Y · 2024
Later among the works it cites.
Easytool: Enhancing llm-based agents with concise tool instruction
Yuan, S., Song, K., Chen, J., Tan, X., Shen, Y., Kan, R., Li, D., and Yang, D · 2024
Later among the works it cites.
Videoagent: A memory-augmented multimodal agent for video understanding
Fan, Y., Ma, X., Wu, R., Du, Y., Li, J., Gao, Z., and Li, Q · 2025
Closest in time.