Fetching the paper…
Reading the bibliography…
Recent advances in large language models (LLMs) and vision-language models (VLMs) have enabled powerful autonomous agents capable of complex reasoning and multi-modal tool use.
A lattice model of secure information flow
Dorothy E. Denning · 1976
Earlier work this paper cites.
Point set voting for partial point cloud analysis
Junming Zhang, Weijia Chen, Yuping Wang, Ram Vasudevan, and Matthew Johnson-Roberson · 2021
Earlier work this paper cites.
Langchain – building applications with llms
Harrison Chase · 2022
Earlier work this paper cites.
Chenliang Li, Hehong Chen, Ming Yan, Weizhou Shen, Haiyang Xu, Zhikai Wu, Zhicheng Zhang, Wenmeng Zhou, Yingda Chen, Chen Cheng, Hongzhu Shi, Ji Zhang, Fei Huang, and Jingren Zhou · 2023
Earlier work this paper cites.
Toolformer: Language models can teach themselves to use tools, 2023
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom · 2023
Earlier work this paper cites.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Yongliang Shen, Kaitao Song, Xu Tan, et al · 2023
Earlier work this paper cites.
Eqdrive: Efficient equivariant motion forecasting with multi-modality for autonomous driving
Yuping Wang and Jier Chen · 2023
Earlier work this paper cites.
Equivariant map and agent geometry for autonomous driving motion prediction
Yuping Wang and Jier Chen · 2023
Cited alongside, same era.
Mm-react: Prompting chatgpt for multimodal reasoning and action, 2023
Zhengyuan Yang, Linjie Li, Jianfeng Wang, Kevin Lin, Ehsan Azarnasab, Faisal Ahmed, Zicheng Liu, Ce Liu, Michael Zeng, and Lijuan Wang · 2023
Cited alongside, same era.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, et al · 2023
Cited alongside, same era.
Exploring large language model based intelligent agents: Definitions, methods, and prospects, 2024a
Yuheng Cheng, Ceyao Zhang, Zhengwen Zhang, Xiangrui Meng, Sirui Hong, Wenhao Li, Zihao Wang, Zekai Wang, Feng Yin, Junhua Zhao, and Xiuqiang He · 2024
Cited alongside, same era.
Visualwebarena: Evaluating multimodal agents on realistic visual web tasks, 2024
A new era in llm security: Exploring security concerns in real-world llm-based systems, 2024
Fangzhou Wu, Ning Zhang, Somesh Jha, Patrick McDaniel, and Chaowei Xiao · 2024
Later among the works it cites.
Autotrust: Benchmarking trustworthiness in large vision-language models for autonomous driving
Shuo Xing, Hongyuan Hua, et al · 2024
Later among the works it cites.
Agentharm: A benchmark for measuring harmfulness of llm agents, 2025
Maksym Andriushchenko, Alexandra Souly, Mateusz Dziemian, Derek Duenas, Maxwell Lin, Justin Wang, Dan Hendrycks, Andy Zou, Zico Kolter, Matt Fredrikson, Eric Winsor, Jerome Wynne, Yarin Gal, and Xander Davies · 2025
Closest in time.
Hacking auto-gpt and escaping its docker container
Lukas Euler · 2025
Closest in time.
Dissecting adversarial robustness of multimodal lm agents, 2025
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jing Yu Koh, Robert Lo, Lawrence Jang, Vikram Duvvur, Ming Chong Lim, Po-Yu Huang, Graham Neubig, Shuyan Zhou, Ruslan Salakhutdinov, and Daniel Fried · 2024
Cited alongside, same era.
Permissive information-flow analysis for large language models, 2024
Shoaib Ahmed Siddiqui, Radhika Gaonkar, Boris Köpf, David Krueger, Andrew Paverd, Ahmed Salem, Shruti Tople, Lukas Wutschitz, Menglin Xia, and Santiago Zanella-Béguelin · 2024
Cited alongside, same era.
Exploring large language model based intelligent agents: Definitions, methods, and prospects, 2024b
Yuheng Cheng, Ceyao Zhang, Zhengwen Zhang, Xiangrui Meng, Sirui Hong, Wenhao Li, Zihao Wang, Zekai Wang, Feng Yin, Junhua Zhao, and Xiuqiang He
Cited in the paper.
Autodan-turbo: A lifelong agent for strategy self-exploration to jailbreak llms, 2024a
Xiaogeng Liu, Peiran Li, Edward Suh, Yevgeniy Vorobeychik, Zhuoqing Mao, Somesh Jha, Patrick McDaniel, Huan Sun, Bo Li, and Chaowei Xiao
Cited in the paper.
Toward the unification of generative and discriminative visual foundation model: A survey
Xu Liu, Tong Zhou, Chong Wang, Yuping Wang, Yuanxin Wang, Qinjingwen Cao, Weizhi Du, Yonghuan Yang, Junjun He, Yu Qiao, et al
Cited in the paper.
Uniocc: A unified benchmark for occupancy forecasting and prediction in autonomous driving
Yuping Wang, Xiangyu Huang, Xiaokang Sun, Mingxuan Yan, Shuo Xing, Zhengzhong Tu, and Jiachen Li
Cited in the paper.
Cmp: Cooperative motion prediction with multi-agent communication
Zehao Wang, Yuping Wang, Zhuoyuan Wu, Hengbo Ma, Zhaowei Li, Hang Qiu, and Jiachen Li
Cited in the paper.
Openemma: Open-source multimodal model for end-to-end autonomous driving
Shuo Xing, Chengyuan Qian, Yuping Wang, Hongyuan Hua, Kexin Tian, Yang Zhou, and Zhengzhong Tu
Cited in the paper.
Chen Henry Wu, Rishi Shah, Jing Yu Koh, Ruslan Salakhutdinov, Daniel Fried, and Aditi Raghunathan · 2025
Closest in time.
Xinkai Zou, Yan Liu, Xiongbo Shi, and Chen Yang · 2025
Closest in time.