Fetching the paper…
Reading the bibliography…
Agentic workflows have become the dominant paradigm for building complex AI systems, orchestrating specialized components, such as planning, reasoning, action execution, and reflection, to tackle sophisticated real-world tasks.
Language models are few-shot learners
Brown, T. B.; Mann, B.; Ryder, N.; et al. 2020 · 1901
Earlier work this paper cites.
A Value for N-Person Games
Shapley, L. S. 1952 · 1952
Earlier work this paper cites.
Feature Selection via Coalitional Game Theory
Cohen, S.; Dror, G.; and Ruppin, E. 2007 · 1961
Earlier work this paper cites.
Shapley Value , 210–216
Hart, S. 1989 · 1989
Earlier work this paper cites.
New mathematical properties of the Banzhaf value
Dragan, I. 1996 · 1996
Earlier work this paper cites.
A Unified Approach to Interpreting Model Predictions
Lundberg, S.; and Lee, S.-I. 2017 · 2017
Earlier work this paper cites.
Data Shapley: Equitable Valuation of Data for Machine Learning
Ghorbani, A.; and Zou, J. 2019 · 2019
Earlier work this paper cites.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Ichter, B.; Xia, F.; Chi, E.; Le, Q.; and Zhou, D. 2022 · 2022
Earlier work this paper cites.
Mistral 8B Instruct
AI, M. 2023 · 2023
Earlier work this paper cites.
MetaGPT: Meta Programming for Multi-Agent Collaborative Framework
Hong, S.; Zheng, X.; Chen, J. P.; Cheng, Y.; Zhang, C.; Wang, Z.; Yau, S. K. S.; Lin, Z. H.; Zhou, L.; Ran, C.; Xiao, L.; and Wu, C. 2023 · 2023
Earlier work this paper cites.
AgentBench: Evaluating LLMs as Agents
Liu, X.; Zhou, H.; Zhang, Z.; Peng, D.; et al. 2023 · 2023
Earlier work this paper cites.
GPT-4 Turbo
OpenAI. 2023 · 2023
Cited alongside, same era.
Reflexion: language agents with verbal reinforcement learning
Shinn, N.; Cassano, F.; Gopinath, A.; Narasimhan, K.; and Yao, S. 2023 · 2023
Cited alongside, same era.
ReAct: Synergizing Reasoning and Acting in Language Models
Yao, S.; Zhao, J.; Yu, D.; Du, N.; Shafran, I.; Narasimhan, K.; and Cao, Y. 2023 · 2023
Cited alongside, same era.
Claude 3.5 Sonnet
Anthropic. 2024 · 2024
Cited alongside, same era.
Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Chiang, W.-L.; Zheng, L.; Sheng, Y.; Angelopoulos, A. N.; Li, T.; Li, D.; Zhu, B.; Zhang, H.; Jordan, M.; Gonzalez, J. E.; and Stoica, I. 2024 · 2024
Cited alongside, same era.
StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models
OpenAI, J. A.; Adler, S.; Agarwal, S.; et al. 2024 · 2024
Later among the works it cites.
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models
Renze, M.; and Guven, E. 2024 · 2024
Later among the works it cites.
AgentQuest: A Multi-Phase Task Planning and Execution Benchmark for Autonomous Agents
Yang, Y.; Wu, H.; Zhao, C.; Liu, M.; et al. 2024 · 2024
Later among the works it cites.
MMAU: A Holistic Benchmark of Agent Capabilities Across Diverse Domains
Yin, G.; Bai, H.; Ma, S.; Nan, F.; Sun, Y.; Xu, Z.; Ma, S.; Lu, J.; Kong, X.; Zhang, A.; Yap, D. A.; zhang, Y.; Ahnert, K.; Kamath, V.; Berglund, M.; Walsh, D.; Gindele, T.; Wiest, J.; Lai, Z.; Wang, X.; Shan, J.; Cao, M.; Pang, R.; and Wang, Z. 2024 · 2024
Later among the works it cites.
OmniACT: A Dataset and Benchmark for Enabling Multi-Modal Task Completion in Large Language Models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Guo, Z.; Cheng, S.; Wang, H.; Liang, S.; Qin, Y.; Li, P.; Liu, Z.; Sun, M.; and Liu, Y. 2024 · 2024
Cited alongside, same era.
Jiang, A. Q.; Sablayrolles, A.; Roux, A.; Mensch, A.; Savary, B.; Bamford, C.; Chaplot, D. S.; de las Casas, D.; Hanna, E. B.; Bressand, F.; Lengyel, G.; Bour, G.; Lample, G.; Lavaud, L. R.; Saulnier, L.; Lachaux, M.-A.; Stock, P.; Subramanian, S.; Yang, S.; Antoniak, S.; Scao, T. L.; Gervet, T.; Lavril, T.; Wang, T.; Lacroix, T.; and Sayed, W. E. 2024 · 2024
Cited alongside, same era.
Approximating the Shapley Value without Marginal Contributions
Kolpaczki, P.; Bengs, V.; Muschalik, M.; and Hüllermeier, E. 2024 · 2024
Cited alongside, same era.
Decision-Oriented Dialogue for Human-AI Collaboration
Lin, J.; Tomlin, N.; Andreas, J.; and Eisner, J. 2024 · 2024
Cited alongside, same era.
shapiq: Shapley Interactions for Machine Learning
Muschalik, M.; Baniecki, H.; Fumagalli, F.; Kolpaczki, P.; Hammer, B.; and Hüllermeier, E. 2024 · 2024
Cited alongside, same era.
GPT-4o mini
OpenAI. 2024 · 2024
Cited alongside, same era.
Doubao-pro-4k
AI, D. 2024a
Cited in the paper.
Zhang, W.; Wu, J.; Wang, T.; Hu, Z.; et al. 2024 · 2024
Later among the works it cites.
Agentic AI: Autonomous Intelligence for Complex Goals—A Comprehensive Survey
Acharya, D. B.; Kuppan, K.; and Divya, B. 2025 · 2025
Closest in time.
AutoGPT: An Open-Source Platform for Autonomous AI Agents
Gravitas, S. 2023 · 2025
Closest in time.
Qwen; :; Yang, A.; Yang, B.; Zhang, B.; Hui, B.; Zheng, B.; Yu, B.; Li, C.; Liu, D.; Huang, F.; Wei, H.; Lin, H.; Yang, J.; Tu, J.; Zhang, J.; Yang, J.; Yang, J.; Zhou, J.; Lin, J.; Dang, K.; Lu, K.; Bao, K.; Yang, K.; Yu, L.; Li, M.; Xue, M.; Zhang, P.; Zhu, Q.; Men, R.; Lin, R.; Li, T.; Tang, T.; Xia, T.; Ren, X.; Ren, X.; Fan, Y.; Su, Y.; Zhang, Y.; Wan, Y.; Liu, Y.; Cui, Z.; Zhang, Z.; and Qiu, Z. 2025 · 2025
Closest in time.
Ai agents vs. agentic ai: A conceptual taxonomy, applications and challenge
Sapkota, R.; Roumeliotis, K. I.; and Karkee, M. 2025 · 2025
Closest in time.
GLM-4 Air
THUDM. 2024 · 2025
Closest in time.