Fetching the paper…
Reading the bibliography…
The era of intelligent agents is upon us, driven by revolutionary advancements in large language models.
M. Wooldridge and N. R. Jennings, “Intelligent agents: Theory and practice,” The knowledge engineering review , vol. 10, no. 2, pp. 115–152, 1995
1995
Earlier work this paper cites.
O. F. Rana and K. Stout, “What is scalability in multi-agent systems?” in Proceedings of the fourth international conference on Autonomous agents , 2000, pp. 56–63
2000
Earlier work this paper cites.
E. H. Durfee, “Distributed problem solving and planning,” in ECCAI Advanced Course on Artificial Intelligence . Springer, 2001, pp. 118–149
2001
Earlier work this paper cites.
R. Deters, “Scalable multi-agent systems,” in Proceedings of the 2001 joint ACM-ISCOPE conference on Java Grande , 2001, p. 182
2001
Earlier work this paper cites.
K. Keshavjee, J. Bosomworth, J. Copen, J. Lai, B. Kucukyazici, R. Lilani, and A. M. Holbrook, “Best practices in emr implementation: a systematic review,” in AMIA Annual Symposium Proceedings , vol. 2006, 2006, p. 982
2006
Earlier work this paper cites.
S. Robertson, H. Zaragoza et al. , “The probabilistic relevance framework: Bm25 and beyond,” Foundations and Trends® in Information Retrieval , vol. 3, no. 4, pp. 333–389, 2009
2009
Earlier work this paper cites.
C. B. Browne, E. Powley, D. Whitehouse, S. M. Lucas, P. I. Cowling, P. Rohlfshagen, S. Tavener, D. Perez, S. Samothrakis, and S. Colton, “A survey of monte carlo tree search methods,” IEEE Transactions on Computational Intelligence and AI in games , vol. 4, no. 1, pp. 1–43, 2012
2012
Earlier work this paper cites.
2019
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” NeurIPS , vol. 33, pp. 9459–9474, 2020
2020
Earlier work this paper cites.
X. Pan, M. Zhang, S. Ji, and M. Yang, “Privacy risks of general-purpose language models,” in IEEE Symposium on Security and Privacy (SP) . IEEE, 2020, pp. 1314–1331
2020
Earlier work this paper cites.
L. Floridi and M. Chiriatti, “GPT-3: Its Nature, Scope, Limits, and Consequences,” Minds and Machines , vol. 30, pp. 681–694, 2020
2020
Earlier work this paper cites.
M. A. Lemley and B. Casey, “Fair Learning,” Tex. L. Rev. , vol. 99, p. 743, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
E. Strubell, A. Ganesh, and A. McCallum, “Energy and Policy Considerations for Deep Learning in NLP,” in AAAI , vol. 34, no. 09, 2020, pp. 13 693–13 696
2020
Earlier work this paper cites.
Y. Li, S. Ren, P. Wu, S. Chen, C. Feng, and W. Zhang, “Learning distilled collaboration graph for multi-agent perception,” NeurIPS , vol. 34, pp. 29 541–29 552, 2021
2021
Earlier work this paper cites.
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson et al. , “Extracting training data from large language models,” in USENIX , 2021, pp. 2633–2650
2021
Earlier work this paper cites.
S. Hoory, A. Feder, A. Tendler, S. Erell, A. Peled-Cohen, I. Laish, H. Nakhost, U. Stemmer, A. Benjamini, A. Hassidim et al. , “Learning and evaluating a differentially private pre-trained language model,” in EMNLP Findings , 2021, pp. 1178–1189
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell, “On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?” in FAccT , 2021, pp. 610–623
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa, “Large language models are zero-shot reasoners,” NeurIPS , vol. 35, pp. 22 199–22 213, 2022
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” NeurIPS , vol. 35, pp. 24 824–24 837, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
D. Dell’Anna, N. Alechina, F. Dalpiaz, M. Dastani, and B. Logan, “Data-driven revision of conditional norms in multi-agent systems,” Journal of Artificial Intelligence Research , vol. 75, pp. 1549–1593, 2022
2022
Earlier work this paper cites.
R. Nakano, J. Hilton, S. Balaji, J. Wu, L. Ouyang, C. Kim, C. Hesse, S. Jain, V. Kosaraju, W. Saunders, X. Jiang, K. Cobbe, T. Eloundou, G. Krueger, K. Button, M. Knight, B. Chess, and J. Schulman, “Webgpt: Browser-assisted question-answering with human feedback,” 2022
2022
Earlier work this paper cites.
“LlamaIndex,” 11 2022. [Online]. Available: https://github.com/jerryjliu/llama_index
2022
Earlier work this paper cites.
N. Carlini, D. Ippolito, M. Jagielski, K. Lee, F. Tramer, and C. Zhang, “Quantifying memorization across neural language models,” in ICLR , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
N. Kandpal, E. Wallace, and C. Raffel, “Deduplicating training data mitigates privacy risks in language models,” in ICML . PMLR, 2022, pp. 10 697–10 707
2022
Earlier work this paper cites.
D. Ganguli, D. Hernandez, L. Lovitt, A. Askell, Y. Bai, A. Chen, T. Conerly, N. Dassarma, D. Drain, N. Elhage et al. , “Predictability and Surprise in Large Generative Models,” in FAccT , 2022, pp. 1747–1764
2022
Earlier work this paper cites.
Y. Jin, X. Wang, R. Yang, Y. Sun, W. Wang, H. Liao, and X. Xie, “Towards fine-grained reasoning for fake news detection,” in AAAI , vol. 36, no. 5, 2022, pp. 5746–5754
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
T. Sumers, S. Yao, K. Narasimhan, and T. Griffiths, “Cognitive architectures for language agents,” TMLR , 2023
2023
Earlier work this paper cites.
G. Li, H. A. A. K. Hammoud, H. Itani, D. Khizbullin, and B. Ghanem, “Camel: Communicative agents for ”mind” exploration of large language model society,” in NeurIPS , 2023
2023
Earlier work this paper cites.
Q. Wu, G. Bansal, J. Zhang, Y. Wu, B. Li, E. Zhu, L. Jiang, X. Zhang, S. Zhang, J. Liu, A. H. Awadallah, R. W. White, D. Burger, and C. Wang, “Autogen: Enabling next-gen llm applications via multi-agent conversation,” 2023
2023
Earlier work this paper cites.
J. S. Park, J. O’Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein, “Generative agents: Interactive simulacra of human behavior,” in UIST , 2023, pp. 1–22
2023
Earlier work this paper cites.
S. Yao, J. Zhao, D. Yu, N. Du, I. Shafran, K. Narasimhan, and Y. Cao, “React: Synergizing reasoning and acting in language models,” in ICLR , 2023
2023
Earlier work this paper cites.
G. Wang, Y. Xie, Y. Jiang, A. Mandlekar, C. Xiao, Y. Zhu, L. Fan, and A. Anandkumar, “Voyager: An open-ended embodied agent with large language models,” TMLR , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
N. Shinn, F. Cassano, A. Gopinath, K. Narasimhan, and S. Yao, “Reflexion: Language agents with verbal reinforcement learning,” NeurIPS , vol. 36, pp. 8634–8652, 2023
2023
Earlier work this paper cites.
J. Ruan, Y. Chen, B. Zhang, Z. Xu, T. Bao, H. Mao, Z. Li, X. Zeng, R. Zhao et al. , “Tptu: Task planning and tool usage of large language model-based ai agents,” in NeurIPS , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
C. Packer, V. Fang, S. G. Patil, K. Lin, S. Wooders, and J. E. Gonzalez, “Memgpt: Towards llms as operating systems,” CoRR , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
J. Long, “Large language model guided tree-of-thought,” arXiv preprint arXiv:2305.08291 , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
H. Sun, Y. Zhuang, L. Kong, B. Dai, and C. Zhang, “Adaplanner: Adaptive planning from feedback with language models,” NeurIPS , vol. 36, pp. 58 202–58 245, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
R. Yang, L. Song, Y. Li, S. Zhao, Y. Ge, X. Li, and Y. Shan, “Gpt4tools: Teaching large language model to use tools via self-instruction,” NeurIPS , vol. 36, pp. 71 995–72 007, 2023
2023
Earlier work this paper cites.
D. A. Boiko, R. MacKnight, B. Kline, and G. Gomes, “Autonomous chemical research with large language models,” Nature , vol. 624, no. 7992, pp. 570–578, 2023
2023
Earlier work this paper cites.
H. Jiang, Q. Wu, C.-Y. Lin, Y. Yang, and L. Qiu, “Llmlingua: Compressing prompts for accelerated inference of large language models,” in EMNLP , 2023, pp. 13 358–13 376
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Y. Du, S. Li, A. Torralba, J. B. Tenenbaum, and I. Mordatch, “Improving factuality and reasoning in language models through multiagent debate,” in ICML , 2023
2023
Earlier work this paper cites.
Q. Zhong, L. Ding, J. Liu, B. Du, and D. Tao, “Self-evolution learning for discriminative language model pretraining,” in ACL Findings , 2023, pp. 4130–4145
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
A. Madaan, N. Tandon, P. Gupta, S. Hallinan, L. Gao, S. Wiegreffe, U. Alon, N. Dziri, S. Prabhumoye, Y. Yang et al. , “Self-refine: Iterative refinement with self-feedback,” NeurIPS , vol. 36, pp. 46 534–46 594, 2023
2023
Earlier work this paper cites.
Y. Weng, M. Zhu, F. Xia, B. Li, S. He, S. Liu, B. Sun, K. Liu, and J. Zhao, “Large language models are better reasoners with self-verification,” in EMNLP Findings , 2023, pp. 2550–2575
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
L. Ying, T. Zhi-Xuan, V. Mansinghka, and J. B. Tenenbaum, “Inferring the goals of communicating agents from actions and instructions,” in Proceedings of the AAAI Symposium Series , vol. 2, no. 1, 2023, pp. 26–33
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
X. Deng, Y. Gu, B. Zheng, S. Chen, S. Stevens, B. Wang, H. Sun, and Y. Su, “Mind2web: Towards a generalist agent for the web,” NeurIPS , vol. 36, pp. 28 091–28 114, 2023
2023
Earlier work this paper cites.
Q. Huang, J. Vora, P. Liang, and J. Leskovec, “Benchmarking large language models as ai research agents,” in NeurIPS Workshop , 2023
2023
Earlier work this paper cites.
Y. Qin, Z. Cai, D. Jin, L. Yan, S. Liang, K. Zhu, Y. Lin, X. Han, N. Ding, H. Wang, R. Xie, F. Qi, Z. Liu, M. Sun, and J. Zhou, “WebCPM: Interactive web search for Chinese long-form question answering,” in ACL , A. Rogers, J. Boyd-Graber, and N. Okazaki, Eds. Toronto, Canada: Association for Computational Linguistics, Jul. 2023, pp. 8968–8988
2023
Earlier work this paper cites.
K. Zhang, H. Zhang, G. Li, J. Li, Z. Li, and Z. Jin, “Toolcoder: Teach code generation models to use api search tools,” 2023
2023
Earlier work this paper cites.
T. Schick, J. Dwivedi-Yu, R. Dessì, R. Raileanu, M. Lomeli, E. Hambro, L. Zettlemoyer, N. Cancedda, and T. Scialom, “Toolformer: Language models can teach themselves to use tools,” Advances in Neural Information Processing Systems , vol. 36, pp. 68 539–68 551, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Y. Song, W. Xiong, D. Zhu, W. Wu, H. Qian, M. Song, H. Huang, C. Li, K. Wang, R. Yao, Y. Tian, and S. Li, “Restgpt: Connecting large language models with real-world restful apis,” 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
C. Qian, C. Han, Y. Fung, Y. Qin, Z. Liu, and H. Ji, “CREATOR: Tool creation for disentangling abstract and concrete reasoning of large language models,” in EMNLP Findings , H. Bouamor, J. Pino, and K. Bali, Eds., Singapore, Dec. 2023, pp. 6922–6939
2023
Earlier work this paper cites.
“LangChain,” 1 2023. [Online]. Available: https://github.com/langchain-ai/langchain
2023
Earlier work this paper cites.
“Dify,” 5 2023. [Online]. Available: https://github.com/langgenius/dify
2023
Earlier work this paper cites.
“Ollama,” 7 2023. [Online]. Available: https://github.com/ollama/ollama
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
E. T. Red, “Malicious chatgpt agents: How gpts can quietly grab your data (demo),” Embrace The Red , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
M. Kang, S. Lee, J. Baek, K. Kawaguchi, and S. J. Hwang, “Knowledge-augmented reasoning distillation for small language models in knowledge-intensive tasks,” NeurIPS , vol. 36, pp. 48 573–48 602, 2023
2023
Earlier work this paper cites.
S. Kim, S. Yun, H. Lee, M. Gubri, S. Yoon, and S. J. Oh, “Propile: Probing privacy leakage in large language models,” NeurIPS , vol. 36, pp. 20 750–20 762, 2023
2023
Earlier work this paper cites.
A. Naseh, K. Krishna, M. Iyyer, and A. Houmansadr, “Stealing the decoding algorithms of language models,” in ACM SIGSAC , 2023, pp. 1835–1849
2023
Earlier work this paper cites.
J. Kirchenbauer, J. Geiping, Y. Wen, J. Katz, I. Miers, and T. Goldstein, “A watermark for large language models,” in ICML . PMLR, 2023, pp. 17 061–17 084
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
S. Moore, R. Tong, A. Singh, Z. Liu, X. Hu, Y. Lu, J. Liang, C. Cao, H. Khosravi, P. Denny et al. , “Empowering Education with LLMs - The Next-Gen Interface and Content Generation,” in International Conference on Artificial Intelligence in Education . Springer, 2023, pp. 32–37
2023
Earlier work this paper cites.
P. Henderson, X. Li, D. Jurafsky, T. Hashimoto, M. A. Lemley, and P. Liang, “Foundation Models and Fair Use,” JMLR , vol. 24, no. 400, pp. 1–79, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
A. S. Luccioni, S. Viguier, and A.-L. Ligozat, “Estimating the Carbon Footprint of BLOOM, a 176B Parameter Language Model,” JMLR , vol. 24, no. 253, pp. 1–15, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
X. Ye, M. Xiao, Z. Ning, W. Dai, W. Cui, Y. Du, and Y. Zhou, “Needed: Introducing hierarchical transformer to eye diseases diagnosis,” in Proceedings of the 2023 SIAM International Conference on Data Mining (SDM) . SIAM, 2023, pp. 667–675
2023
Earlier work this paper cites.
X. Feng, Y. Luo, Z. Wang, H. Tang, M. Yang, K. Shao, D. Mguni, Y. Du, and J. Wang, “Chessgpt: Bridging policy learning and language modeling,” in NeurIPS , 2023, pp. 7216–7262
2023
Earlier work this paper cites.
T. Carta, C. Romac, T. Wolf, S. Lamprier, O. Sigaud, and P.-Y. Oudeyer, “Grounding large language models in interactive environments with online reinforcement learning,” in ICML , 2023, pp. 3676–3713
2023
Earlier work this paper cites.
A. Zhu, L. Martin, A. Head, and C. Callison-Burch, “Calypso: Llms as dungeon master’s assistants,” in AAAI , 2023, pp. 380–390
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Y. Sun, Z. Li, K. Fang, C. H. Lee, and A. Asadipour, “Language as reality: a co-creative storytelling game experience in 1001 nights using generative ai,” in AAAI , 2023, pp. 425–434
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
G. V. Aher, R. I. Arriaga, and A. T. Kalai, “Using large language models to simulate multiple humans and replicate human subject studies,” in ICML , 2023, pp. 337–371
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
X. Wang, J. Wei, D. Schuurmans, Q. V. Le, E. H. Chi, S. Narang, A. Chowdhery, and D. Zhou, “Self-consistency improves chain of thought reasoning in language models,” in ICLR , 2023
2023
Earlier work this paper cites.
S. Zhou, U. Alon, S. Agarwal, and G. Neubig, “Codebertscore: Evaluating code generation with pretrained models of code,” in EMNLP , 2023, pp. 13 921–13 937
2023
Earlier work this paper cites.
Z. Wang, S. Zhou, D. Fried, and G. Neubig, “Execution-based evaluation for open-domain code generation,” in EMNLP , 2023, pp. 1271–1290
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Wu, Z. Jiang, A. Khan, Y. Fu, L. Ruis, E. Grefenstette, and T. Rocktäschel, “Chatarena: Multi-agent language game environments for large language models,” 2023
2023
Cited alongside, same era.
2024
Cited alongside, same era.
Y. Wang, D. Xue, S. Zhang, and S. Qian, “Badagent: Inserting and activating backdoor attacks in llm agents,” in ACL , 2024, pp. 9811–9827
2024
Later among the works it cites.
2024
Later among the works it cites.
W. Hua, X. Yang, M. Jin, Z. Li, W. Cheng, R. Tang, and Y. Zhang, “Trustagent: Towards safe and trustworthy llm-based agents through agent constitution,” in EMNLP Findings , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Fei, Y. Yao, Z. Zhang, F. Liu, A. Zhang, and T.-S. Chua, “From multimodal llm to human-level ai: Modality, instruction, reasoning, efficiency and beyond,” in COLING , 2024, pp. 1–8
2024
Cited alongside, same era.
C. Wang, W. Luo, Q. Chen, H. Mai, J. Guo, S. Dong, Z. Li, L. Ma, S. Gao et al. , “Tool-lmm: A large multi-modal model for tool agent learning,” arXiv e-prints , pp. arXiv–2401, 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
M. Xu, H. Du, D. Niyato, J. Kang, Z. Xiong, S. Mao, Z. Han, A. Jamalipour, D. I. Kim, X. Shen et al. , “Unleashing the power of edge-cloud generative ai in mobile networks: A survey of aigc services,” IEEE Communications Surveys & Tutorials , vol. 26, no. 2, pp. 1127–1170, 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Z. Deng, Y. Guo, C. Han, W. Ma, J. Xiong, S. Wen, and Y. Xiang, “Ai agents under threat: A survey of key security challenges and future pathways,” ACM Computing Surveys , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Xiang, Y. Zeng, M. Kang, C. Xu, J. Zhang, Z. Yuan, Z. Chen, C. Xie, F. Jiang, M. Pan et al. , “Clas 2024: The competition for llm and agent safety,” in NeurIPS Workshop , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
B. Chen, G. Li, X. Lin, Z. Wang, and J. Li, “Blockagents: Towards byzantine-robust llm-based multi-agent coordination via blockchain,” in ACM Turing Award Celebration Conference , 2024, pp. 187–192
2024
Later among the works it cites.
2024
Later among the works it cites.
L. Wang, J. Wang, J. Wan, L. Long, Z. Yang, and Z. Qin, “Property existence inference against generative models,” in USENIX , 2024, pp. 2423–2440
2024
Later among the works it cites.
Z. Li, C. Wang, P. Ma, C. Liu, S. Wang, D. Wu, C. Gao, and Y. Liu, “On extracting specialized code abilities from large language models: A feasibility study,” in ICSE , 2024, pp. 1–13
2024
Later among the works it cites.
Y. Lin, Z. Gao, H. Du, D. Niyato, J. Kang, Z. Xiong, and Z. Zheng, “Blockchain-based efficient and trustworthy aigc services in metaverse,” IEEE Transactions on Services Computing , 2024
2024
Later among the works it cites.
X. Shen, Y. Qu, M. Backes, and Y. Zhang, “Prompt stealing attacks against { \{ Text-to-Image } \} generation models,” in USENIX , 2024, pp. 5823–5840
2024
Later among the works it cites.
2024
Later among the works it cites.
B. Hui, H. Yuan, N. Gong, P. Burlina, and Y. Cao, “Pleak: Prompt leaking attacks against large language model applications,” in ACM SIGSAC , 2024, pp. 3600–3614
2024
Later among the works it cites.
P. Tadas and S. Agarmore, “Redefining Work in the Age of AI: Challenges and Pathways to Opportunities,” in SPICES . IEEE, 2024, pp. 1–5
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Deng, Y. Duan, X. Jin, H. Chang, Y. Tian, H. Liu, H. P. Zou, Y. Jin, Y. Xiao, Y. Wang et al. , “Deconstructing The Ethics of Large Language Models from Long-standing Issues to New-emerging Dilemmas: A Survey,” arXiv e-prints , pp. arXiv–2406, 2024
2024
Later among the works it cites.
I. Shumailov, Z. Shumaylov, Y. Zhao, N. Papernot, R. Anderson, and Y. Gal, “AI models collapse when trained on recursively generated data,” Nature , vol. 631, no. 8022, pp. 755–759, 2024
2024
Later among the works it cites.
Y. Xiao, Y. Jin, Y. Bai, Y. Wu, X. Yang, X. Luo, W. Yu, X. Zhao, Y. Liu, Q. Gu et al. , “Large language models can be contextual privacy protection learners,” in EMNLP , 2024, pp. 14 179–14 201
2024
Later among the works it cites.
J. Zhou, “Awesome ai agents for scientific discovery,” https://github.com/zhoujieli/Awesome-LLM-Agents-Scientific-Discovery , 2024
2024
Later among the works it cites.
Y. Jin, Q. Zhao, Y. Wang, H. Chen, K. Zhu, Y. Xiao, and J. Wang, “Agentreview: Exploring peer review dynamics with llm agents,” in EMNLP , 2024, pp. 1208–1226
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Chen, C. Hu, M. Wu, Q. Long, X. Wang, Y. Zhou, and M. Xiao, “Genesum: Large language model-based gene summary extraction,” in 2024 IEEE International Conference on Bioinformatics and Biomedicine (BIBM) . IEEE, 2024, pp. 1438–1443
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
R. W. Lee, K. H. Lee, J. S. Yun, M. S. Kim, and H. S. Choi, “Comparative analysis of m4cxr, an llm-based chest x-ray report generation model, and chatgpt in radiological interpretation,” Journal of Clinical Medicine , vol. 13, no. 23, p. 7057, 2024
2024
Later among the works it cites.
N. Li, C. Gao, M. Li, Y. Li, and Q. Liao, “Econagent: large language model-empowered agents for simulating macroeconomic activities,” ACL , pp. 15 523–15 536, 2024
2024
Later among the works it cites.
Q. Zhao, J. Wang, Y. Zhang, Y. Jin, K. Zhu, H. Chen, and X. Xie, “Competeai: Understanding the competition dynamics in large language model-based agents,” in ICML , 2024, pp. 61 092–61 107
2024
Later among the works it cites.
Z. Ma, Y. Mei, and Z. Su, “Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support,” in AMIA Annual Symposium Proceedings , vol. 2023, 2024, p. 1105
2024
Later among the works it cites.
J. Zhang, X. Xu, N. Zhang, R. Liu, B. Hooi, and S. Deng, “Exploring collaboration mechanisms for llm agents: A social psychology view,” in ACL , 2024, pp. 14 544–14 607
2024
Later among the works it cites.
R. Liu, R. Yang, C. Jia, G. Zhang, D. Zhou, A. M. Dai, D. Yang, and S. Vosoughi, “Training socially aligned language models on simulated social interactions,” in ICLR , 2024
2024
Later among the works it cites.
Y. Dong, X. Jiang, Z. Jin, and G. Li, “Self-collaboration code generation via chatgpt,” ACM Transactions on Software Engineering and Methodology , vol. 33, no. 7, pp. 1–38, 2024
2024
Later among the works it cites.
C. Qian, X. Cong, C. Yang, W. Chen, Y. Su, J. Xu, Z. Liu, and M. Sun, “Chatdev: Communicative agents for software development,” in ACL , 2024, pp. 15 174–15 186
2024
Later among the works it cites.
A. Zhang, Y. Chen, L. Sheng, X. Wang, and T.-S. Chua, “On generative agents in recommendation,” in SIGIR , 2024, pp. 1807–1817
2024
Later among the works it cites.
J. Zhang, Y. Hou, R. Xie, W. Sun, J. McAuley, W. X. Zhao, L. Lin, and J.-R. Wen, “Agentcf: Collaborative learning with autonomous language agents for recommender systems,” in WWW , 2024, pp. 3679–3689
2024
Later among the works it cites.
Z. Wang, Y. Yu, W. Zheng, W. Ma, and M. Zhang, “Macrec: A multi-agent collaboration framework for recommendation,” in SIGIR , 2024, pp. 2760–2764
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Jin, M. Choi, G. Verma, J. Wang, and S. Kumar, “Mm-soc: Benchmarking multimodal large language models in social media platforms,” in ACL Findings , 2024
2024
Later among the works it cites.
Z. Yao, Z. Tang, J. Lou, P. Shen, and W. Jia, “Velo: A vector database-assisted cloud-edge collaborative llm qos optimization framework,” in ICWS . IEEE, 2024, pp. 865–876
2024
Later among the works it cites.
X. Cheng, X. Wang, X. Zhang, T. Ge, S.-Q. Chen, F. Wei, H. Zhang, and D. Zhao, “xrag: Extreme context compression for retrieval-augmented generation with one token,” in NeurIPS , 2024
2024
Later among the works it cites.
Y. Jin, M. Chandra, G. Verma, Y. Hu, M. De Choudhury, and S. Kumar, “Better to ask in english: Cross-lingual evaluation of large language models for healthcare queries,” in WWW , 2024, pp. 2627–2638
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
G. Agrawal, T. Kumarage, Z. Alghamdi, and H. Liu, “Can knowledge graphs reduce hallucinations in llms?: A survey,” in NAACL , 2024, pp. 3947–3960
2024
Later among the works it cites.
T. Gao, H. Yen, J. Yu, and D. Chen, “Enabling large language models to generate text with citations,” in EMNLP , 2024
2024
Later among the works it cites.
K. Zhu, J. Chen, J. Wang, N. Z. Gong, D. Yang, and X. Xie, “Dyval: Dynamic evaluation of large language models for reasoning tasks,” in ICLR , 2024
2024
Later among the works it cites.
K. Zhu, J. Wang, Q. Zhao, R. Xu, and X. Xie, “Dynamic evaluation of large language models by meta probing agents,” in ICML . PMLR, 2024, pp. 62 599–62 617
2024
Later among the works it cites.
J. Yao, X. Yi, Y. Gong, X. Wang, and X. Xie, “Value fulcra: Mapping large language models to the multidimensional spectrum of basic human value,” in NAACL , 2024, pp. 8754–8777
2024
Later among the works it cites.
Z. Xi, W. Chen, X. Guo, W. He, Y. Ding, B. Hong, M. Zhang, J. Wang, S. Jin, E. Zhou et al. , “The rise and potential of large language model based agents: A survey,” Science China Information Sciences , vol. 68, no. 2, p. 121101, 2025
2025
Closest in time.
G. Qu, Q. Chen, W. Wei, Z. Lin, X. Chen, and K. Huang, “Mobile edge intelligence for large language models: A contemporary survey,” IEEE Communications Surveys & Tutorials , 2025
2025
Closest in time.
2025
Closest in time.
J. Zhang, J. Xiang, Z. Yu, F. Teng, X.-H. Chen, J. Chen, M. Zhuge, X. Cheng, S. Hong, J. Wang, B. Liu, Y. Luo, and C. Wu, “AFlow: Automating agentic workflow generation,” in ICLR , 2025
2025
Closest in time.
L. Wang, J. Zhang, H. Yang, Z.-Y. Chen, J. Tang, Z. Zhang, X. Chen, Y. Lin, H. Sun, R. Song et al. , “User behavior simulation with large language model-based agents,” ACM Transactions on Information Systems , vol. 43, no. 2, pp. 1–37, 2025
2025
Closest in time.
W. Wu, Y. Jing, Y. Wang, W. Hu, and D. Tao, “Graph-augmented reasoning: Evolving step-by-step knowledge graph retrieval for llm reasoning,” 2025
2025
Closest in time.
2025
Closest in time.
J.-W. Choi, H. Kim, H. Ong, Y. Yoon, M. Jang, J. Kim et al. , “Reactree: Hierarchical task planning with dynamic tree expansion using llm agent nodes,” 2025
2025
Closest in time.
M. Jafaripour, S. Golestan, S. Miwa, Y. Mitsuka, and O. Zaiane, “Adaptive iterative feedback prompting for obstacle-aware path planning via llms,” in AAAI Workshop , 2025
2025
Closest in time.
S. Wu, S. Zhao, Q. Huang, K. Huang, M. Yasunaga, K. Cao, V. Ioannidis, K. Subbian, J. Leskovec, and J. Y. Zou, “Avatar: Optimizing llm agents for tool usage via contrastive reasoning,” NeurIPS , vol. 37, pp. 25 981–26 010, 2025
2025
Closest in time.
T. Akiba, M. Shing, Y. Tang, Q. Sun, and D. Ha, “Evolutionary optimization of model merging recipes,” Nature Machine Intelligence , pp. 1–10, 2025
2025
Closest in time.
2025
Closest in time.
M. Li, S. Zhao, Q. Wang, K. Wang, Y. Zhou, S. Srivastava, C. Gokmen, T. Lee, E. L. Li, R. Zhang et al. , “Embodied agent interface: Benchmarking llms for embodied decision making,” NeurIPS , vol. 37, pp. 100 428–100 534, 2025
2025
Closest in time.
2025
Closest in time.
T. Xie, D. Zhang, J. Chen, X. Li, S. Zhao, R. Cao, J. H. Toh, Z. Cheng, D. Shin, F. Lei et al. , “Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments,” NeurIPS , vol. 37, pp. 52 040–52 094, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
J. Gehring, K. Zheng, J. Copet, V. Mella, Q. Carbonneaux, T. Cohen, and G. Synnaeve, “Rlef: Grounding code llms in execution feedback with reinforcement learning,” 2025
2025
Closest in time.
“MCP Agent,” 2 2025. [Online]. Available: https://github.com/lastmile-ai/mcp-agent
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
W. Yang, X. Bi, Y. Lin, S. Chen, J. Zhou, and X. Sun, “Watch out for your agents! investigating backdoor threats to llm-based agents,” NeurIPS , vol. 37, pp. 100 938–100 964, 2025
2025
Closest in time.
T. Tong, F. Wang, Z. Zhao, and M. Chen, “Badjudge: Backdoor vulnerabilities of llm-as-a-judge,” in ICLR , 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
E. Debenedetti, J. Zhang, M. Balunovic, L. Beurer-Kellner, M. Fischer, and F. Tramèr, “Agentdojo: A dynamic environment to evaluate prompt injection attacks and defenses for llm agents,” NeurIPS , vol. 37, pp. 82 895–82 920, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Z. Chen, Z. Xiang, C. Xiao, D. Song, and B. Li, “Agentpoison: Red-teaming llm agents via poisoning memory or knowledge bases,” NeurIPS , vol. 37, pp. 130 185–130 213, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
D. A. Alber, Z. Yang, A. Alyakin, E. Yang, S. Rai, A. A. Valliani, J. Zhang, G. R. Rosenbaum, A. K. Amend-Thomas, D. B. Kurland et al. , “Medical large language models are vulnerable to data-poisoning attacks,” Nature Medicine , pp. 1–9, 2025
2025
Closest in time.
AAAI, “Aaai 2025 presidential panel: Future of ai research,” 2025. [Online]. Available: https://aaai.org/wp-content/uploads/2025/03/AAAI-2025-PresPanel-Report-FINAL.pdf
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.