Fetching the paper…
Reading the bibliography…
Large Language Model (LLM) agents have shown significant autonomous capabilities in dynamically searching and incorporating relevant tools or Model Context Protocol (MCP) servers for individual queries.
Protip: Progressive tool retrieval improves planning
Raviteja Anantha, Bortik Bandyopadhyay, Anirudh Kashi, Sayantan Mahinder, Andrew W. Hill, and Srinivas Chappidi. 2023 · 2023
Earlier work this paper cites.
Api-bank: A comprehensive benchmark for tool-augmented llms
Minghao Li, Yingxiu Zhao, Bowen Yu, Feifan Song, Hangyu Li, Haiyang Yu, Zhoujun Li, Fei Huang, and Yongbin Li. 2023 · 2023
Earlier work this paper cites.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph C. O’Brien, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2023 · 2023
Earlier work this paper cites.
Gorilla: Large language model connected with massive apis
Shishir G. Patil, Tianjun Zhang, Xin Wang, and Joseph E. Gonzalez. 2023 · 2023
Earlier work this paper cites.
Toolllm: Facilitating large language models to master 16000+ real-world apis
Yujia Qin, Shihao Liang, Yining Ye, Kunlun Zhu, Lan Yan, Yaxi Lu, Yankai Lin, Xin Cong, Xiangru Tang, Bill Qian, Sihan Zhao, Lauren Hong, Runchu Tian, Ruobing Xie, Jie Zhou, Mark Gerstein, Dahai Li, Zhiyuan Liu, and Maosong Sun. 2023 · 2023
Earlier work this paper cites.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, He Ye, and Yanlin Wang. 2023 · 2023
Earlier work this paper cites.
Re-invoke: Tool invocation rewriting for zero-shot tool retrieval
Yanfei Chen, Jinsung Yoon, Devendra Singh Sachan, Qingze Wang, Vincent Cohen-Addad, Mohammadhossein Bateni, Chen-Yu Lee, and Tomas Pfister. 2024 · 2024
Earlier work this paper cites.
Anytool: Self-reflective, hierarchical agents for large-scale api calls
Yu Du, Fangyun Wei, and Hongyang Zhang. 2024 · 2024
Earlier work this paper cites.
Toolace: Winning the points of llm function calling
Weiwen Liu, Xu Huang, Xingshan Zeng, Xinlong Hao, Shuai Yu, Dexun Li, Shuai Wang, Weinan Gan, Zhengying Liu, Yuanqing Yu, Zezhong Wang, Yuxian Wang, Wu Ning, Yutai Hou, Bin Wang, Chuhan Wu, Xinzhi Wang, Yong Liu, Yasheng Wang, and 8 others. 2024 · 2024
Earlier work this paper cites.
Toolshed: Scale tool-equipped agents with advanced rag-tool fusion and tool knowledge bases
Elias Lumer, Vamse Kumar Subbiah, James A. Burke, Pradeep Honaganahalli Basavaraju, and Austin Huber. 2024 · 2024
Earlier work this paper cites.
Evaluating very long-term conversational memory of llm agents
Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov, Mohit Bansal, Francesco Barbieri, and Yuwei Fang. 2024 · 2024
Earlier work this paper cites.
Function calling
OpenAI. 2024 · 2024
Earlier work this paper cites.
Memgpt: Towards llms as operating systems
Charles Packer, Sarah Wooders, Kevin Lin, Vivian Fang, Shishir G. Patil, Ion Stoica, and Joseph E. Gonzalez. 2024 · 2024
Earlier work this paper cites.
Graph retrieval-augmented generation: A survey
Boci Peng, Yun Zhu, Yongchao Liu, Xiaohe Bo, Haizhou Shi, Chuntao Hong, Yan Zhang, and Siliang Tang. 2024 · 2024
Earlier work this paper cites.
On context utilization in summarization with large language models
Mathieu Ravaut, Aixin Sun, Nancy F. Chen, and Shafiq Joty. 2024 · 2024
Earlier work this paper cites.
Seal-tools: Self-instruct tool learning dataset for agent tuning and detailed benchmark
Mengsong Wu, Tong Zhu, Han Han, Chuanyuan Tan, Xiang Zhang, and Wenliang Chen. 2024 · 2024
Earlier work this paper cites.
Craft: Customizing llms by creating and retrieving from specialized toolsets
Lifan Yuan, Yangyi Chen, Xingyao Wang, Yi R. Fung, Hao Peng, and Heng Ji. 2024 · 2024
Earlier work this paper cites.
Toolrerank: Adaptive and hierarchy-aware reranking for tool retrieval
Yuanhang Zheng, Peng Li, Wei Liu, Yang Liu, Jian Luan, and Bin Wang. 2024 · 2024
Cited alongside, same era.
Architecting agent memory: Principles, patterns, and best practices
Richmond Alake. 2025 · 2025
Cited alongside, same era.
Anthropic
Anthropic. 2025 · 2025
Cited alongside, same era.
How to fix your context
Drew Breunig. 2025 · 2025
Cited alongside, same era.
Toolspectrum : Towards personalized tool utilization for large language models
Zihao Cheng, Hongru Wang, Zeming Liu, Yuhang Guo, Yuanfang Guo, Yunhong Wang, and Haifeng Wang. 2025 · 2025
Cited alongside, same era.
Provence: efficient and robust context pruning for retrieval-augmented generation
Meta llama
Meta Platforms. 2025 · 2025
Closest in time.
Tools documentation
Model Context Protocol. 2025 · 2025
Closest in time.
Openai model provider long -
OpenAI. 2025b · 2025
Closest in time.
On memory construction and retrieval for personalized conversational agents
Zhuoshi Pan, Qianhui Wu, Huiqiang Jiang, Xufang Luo, Hao Cheng, Dongsheng Li, Yuqing Yang, Chin-Yew Lin, H. Vicky Zhao, Lili Qiu, and Jianfeng Gao. 2025 · 2025
Closest in time.
Perplexity ai: Persistent memory in conversational ai
Perplexity. 2025 · 2025
Closest in time.
The new skill in ai is not prompting, it’s context engineering
Philipp Schmid. 2025 · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nadezhda Chirkova, Thibault Formal, Vassilina Nikoulina, and Stéphane Clinchant. 2025 · 2025
Cited alongside, same era.
Deepeval: The open-source llm evaluation framework
Confident AI. 2025 · 2025
Cited alongside, same era.
Mcp-zero: Active tool discovery for autonomous llm agents
Xiang Fei, Xiawu Zheng, and Hao Feng. 2025 · 2025
Cited alongside, same era.
Google model provider long -
Google. 2025b · 2025
Cited alongside, same era.
Evaluating personalized tool-augmented llms from the perspectives of personalization and proactivity
Yupu Hao, Pengfei Cao, Zhuoran Jin, Huanxuan Liao, Yubo Chen, Kang Liu, and Jun Zhao. 2025 · 2025
Cited alongside, same era.
Context rot: How increasing input tokens impacts llm performance
Kelly Hong, Anton Troynikov, and Jeff Huber. 2025 · 2025
Cited alongside, same era.
Keynote: Software is changing (again)
Andrej Karpathy. 2025 · 2025
Cited alongside, same era.
Lianlei Shan, Shixian Luo, Zezhou Zhu, Yu Yuan, and Yong Wu. 2025 · 2025
Closest in time.
Longrope2: Near-lossless llm context window scaling
Ning Shang, Li Lyna Zhang, Siyuan Wang, Gaokai Zhang, Gilsinia Lopez, Fan Yang, Weizhu Chen, and Mao Yang. 2025 · 2025
Closest in time.
Agentic retrieval-augmented generation: A survey on agentic rag
Aditi Singh, Abul Ehtesham, Saket Kumar, and Tala Talaei Khoei. 2025 · 2025
Closest in time.
Llm leaderboard
Vellum.ai. 2025 · 2025
Closest in time.
Recursively summarizing enables long-term dialogue memory in large language models
Qingyue Wang, Yanhe Fu, Yanan Cao, Shuai Wang, Zhiliang Tian, and Liang Ding. 2025 · 2025
Closest in time.
A-mem: Agentic memory for llm agents
Wujiang Xu, Kai Mei, Hang Gao, Juntao Tan, Zujie Liang, and Yongfeng Zhang. 2025 · 2025
Closest in time.
Zep: A context engineering platform for ai agents
Zep. 2025 · 2025
Closest in time.
Personaagent: When large language model agents meet personalization at test time
Weizhi Zhang, Xinyang Zhang, Chenwei Zhang, Liangwei Yang, Jingbo Shang, Zhepei Wei, Henry Peng Zou, Zijie Huang, Zhengyang Wang, Yifan Gao, Xiaoman Pan, Lian Xiong, Jingguo Liu, Philip S. Yu, and Xian Li. 2025 · 2025
Closest in time.
Divide-then-aggregate: An efficient tool learning method via parallel tool invocation
Dongsheng Zhu, Weixian Shi, Zhengliang Shi, Zhaochun Ren, Shuaiqiang Wang, Lingyong Yan, and Dawei Yin. 2025 · 2025
Closest in time.
Yuchen Zhuang, Jingfeng Yang, Haoming Jiang, Xin Liu, Kewei Cheng, Sanket Lokegaonkar, Yifan Gao, Qing Ping, Tianyi Liu, Binxuan Huang, Zheng Li, Zhengyang Wang, Pei Chen, Ruijie Wang, Rongzhi Zhang, Nasser Zalmout, Priyanka Nigam, Bing Yin, and Chao Zhang. 2025 · 2025
Closest in time.