Fetching the paper…
Reading the bibliography…
Web agents for online shopping have shown great promise in automating user interactions across e-commerce platforms.
Planning and acting in partially observable stochastic domains
Kaelbling, L. P., Littman, M. L., and Cassandra, A. R · 1998
Earlier work this paper cites.
A taxonomy of queries for e-commerce search
Sondhi, P., Sharma, M., Kolari, P., and Zhai, C · 2018
Earlier work this paper cites.
Query reformulation in e-commerce search
Hirsch, S., Guy, I., Nus, A., Dagan, A., and Kurland, O · 2020
Earlier work this paper cites.
Challenges and research opportunities in ecommerce search and recommendations
Tsagkias, M., King, T. H., Kallumadi, S., Murdock, V., and de Rijke, M · 2020
Earlier work this paper cites.
Towards personalized and semantic retrieval: An end-to-end solution for e-commerce search via embedding learning
Zhang, H., Wang, S., Zhang, K., Tang, Z., Jiang, Y., Xiao, Y., Yan, W., and Yang, W · 2020
Earlier work this paper cites.
WebGPT: Browser-assisted question-answering with human feedback
Nakano, R., Hilton, J., Balaji, S., Wu, J., Ouyang, L., Kim, C., Hesse, C., Jain, S., Kosaraju, V., Saunders, W., Jiang, X., Cobbe, K., Eloundou, T., Krueger, G., Button, K., Knight, M., Chess, B., and Schulman, J · 2021
Earlier work this paper cites.
Improving legal judgment prediction through reinforced criminal element extraction
Lyu, Y., Wang, Z., Ren, Z., Ren, P., Chen, Z., Liu, X., Li, Y., Li, H., and Song, H · 2022
Earlier work this paper cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Yao, S., Chen, H., Yang, J., and Narasimhan, K · 2022
Earlier work this paper cites.
Mind2web: Towards a generalist agent for the web
Deng, X., Gu, Y., Zheng, B., Chen, S., Stevens, S., Wang, B., Sun, H., and Su, Y · 2023
Earlier work this paper cites.
Language models can solve computer tasks
Kim, G., Baldi, P., and McAleer, S · 2023
Earlier work this paper cites.
Multi-defendant legal judgment prediction via hierarchical reasoning
Lyu, Y., Hao, J., Wang, Z., Zhao, K., Gao, S., Ren, P., Chen, Z., Wang, F., and Ren, Z · 2023
Earlier work this paper cites.
Feature-level debiased natural language understanding
Lyu, Y., Li, P., Yang, Y., de Rijke, M., Ren, P., Zhao, Y., Yin, D., and Ren, Z · 2023
Earlier work this paper cites.
From pixels to UI actions: Learning to follow instructions via graphical user interfaces
Shaw, P., Joshi, M., Cohan, J., Berant, J., Pasupat, P., Hu, H., Khandelwal, U., Lee, K., and Toutanova, K · 2023
Earlier work this paper cites.
WizardLM: Empowering large language models to follow complex instructions
Xu, C., Sun, Q., Zheng, K., Geng, X., Zhao, P., Feng, J., Tao, C., and Jiang, D · 2023
Earlier work this paper cites.
Variational reasoning over incomplete knowledge graphs for conversational recommendation
Zhang, X., Xin, X., Li, D., Liu, W., Ren, P., Chen, Z., Ma, J., and Ren, Z · 2023
Earlier work this paper cites.
Agent-e: From autonomous web navigation to foundational design principles in agentic systems
Abuelsaad, T., Akkil, D., Dey, P., Jagmohan, A., Vempaty, A., and Kokku, R · 2024
Earlier work this paper cites.
Large language models empowered personalized web agents
Cai, H., Li, Y., Wang, W., Zhu, F., Shen, X., Li, W., and Chua, T · 2024
Earlier work this paper cites.
Identifying high consideration e-commerce search queries
Chen, Z., Choi, J. I., Fetahu, B., and Malmasi, S · 2024
Cited alongside, same era.
Multimodal web navigation with instruction-finetuned foundation models
Furuta, H., Lee, K., Nachum, O., Matsuo, Y., Faust, A., Gu, S. S., and Gur, I · 2024
Cited alongside, same era.
Confucius: Iterative tool learning from introspection feedback by easy-to-difficult curriculum
Gao, S., Shi, Z., Zhu, M., Fang, B., Xin, X., Ren, P., Chen, Z., Ma, J., and Ren, Z · 2024
Cited alongside, same era.
Navigating the digital world as humans do: Universal visual grounding for gui agents
Gou, B., Wang, R., Zheng, B., Xie, Y., Chang, C., Shu, Y., Sun, H., and Su, Y · 2024
Cited alongside, same era.
A real-world webagent with planning, long context understanding, and program synthesis
Gur, I., Furuta, H., Huang, A. V., Safdari, M., Matsuo, Y., Eck, D., and Faust, A · 2024
Cited alongside, same era.
GUI agents with foundation models: A comprehensive survey
Wang, S., Liu, W., Chen, J., Gan, W., Zeng, X., Yu, S., Hao, X., Shao, K., Wang, Y., and Tang, R · 2024
Later among the works it cites.
Assistantbench: Can web agents solve realistic and time-consuming tasks?
Yoran, O., Amouyal, S. J., Malaviya, C., Bogin, B., Press, O., and Berant, J · 2024
Later among the works it cites.
Towards empathetic conversational recommender systems
Zhang, X., Xie, R., Lyu, Y., Xin, X., Ren, P., Liang, M., Zhang, B., Kang, Z., de Rijke, M., and Ren, Z · 2024
Later among the works it cites.
Gpt-4v(ision) is a generalist web agent, if grounded
Zheng, B., Gou, B., Kil, J., Sun, H., and Su, Y · 2024
Later among the works it cites.
WebArena: A realistic web environment for building autonomous agents
Zhou, S., Xu, F. F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Ou, T., Bisk, Y., Fried, D., Alon, U., and Neubig, G · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
VideoWebArena: Evaluating long context multimodal agents with video understanding web tasks
Jang, L., Li, Y., Ding, C., Lin, J., Liang, P. P., Zhao, D., Bonatti, R., and Koishida, K · 2024
Cited alongside, same era.
VisualWebArena: Evaluating multimodal agents on realistic visual web tasks
Koh, J. Y., Lo, R., Jang, L., Duvvur, V., Lim, M. C., Huang, P., Neubig, G., Zhou, S., Salakhutdinov, R., and Fried, D · 2024
Cited alongside, same era.
AutoWebGLM: A large language model-based web navigating agent
Lai, H., Liu, X., Iong, I. L., Yao, S., Chen, Y., Shen, P., Yu, H., Zhang, H., Zhang, X., Dong, Y., and Tang, J · 2024
Cited alongside, same era.
Weblinx: Real-world website navigation with multi-turn dialogue
Lù, X. H., Kasner, Z., and Reddy, S · 2024
Cited alongside, same era.
Knowtuning: Knowledge-aware fine-tuning for large language models
Lyu, Y., Yan, L., Wang, S., Shi, H., Yin, D., Ren, P., Chen, Z., de Rijke, M., and Ren, Z · 2024
Cited alongside, same era.
GAIA: a benchmark for general AI assistants
Mialon, G., Fourrier, C., Wolf, T., LeCun, Y., and Scialom, T · 2024
Cited alongside, same era.
Browser use = state of the art web agent, 2024
Müller, M. and Žunič, G · 2024
Cited alongside, same era.
Try deep research and Gemini 2.0 flash experimental
Gemini · 2025
Closest in time.
Agent-centric information access
Kanoulas, E., Eustratiadis, P., Li, Y., Lyu, Y., Pal, V., Poerwawinata, G., Qiao, J., and Wang, Z · 2025
Closest in time.
InfiGUI-R1: Advancing multimodal gui agents from reactive actors to deliberative reasoners
Liu, Y., Li, P., Xie, C., Hu, X., Han, X., Zhang, S., Yang, H., and Wu, F · 2025
Closest in time.
Ning, L., Liang, Z., Jiang, Z., Qu, H., Ding, Y., Fan, W., Wei, X.-y., Lin, S., Liu, H., Yu, P. S., et al · 2025
Closest in time.
Introducing deep research
OpenAI · 2025
Closest in time.
Tool learning in the wild: Empowering language models as automatic tool agents
Shi, Z., Gao, S., Yan, L., Feng, Y., Chen, X., Chen, Z., Yin, D., Verberne, S., and Ren, Z · 2025
Closest in time.
BEARCUBS: A benchmark for computer-using web agents
Song, Y., Thai, K., Pham, C. M., Chang, Y., Nadaf, M., and Iyyer, M · 2025
Closest in time.
A cooperative multi-agent framework for zero-shot named entity recognition
Wang, Z., Zhao, Z., Lyu, Y., Chen, Z., de Rijke, M., and Ren, Z · 2025
Closest in time.
An illusion of progress? assessing the current state of web agents
Xue, T., Qi, W., Shi, T., Song, C. H., Gou, B., Song, D., Sun, H., and Su, Y · 2025
Closest in time.
Improving sequential recommenders through counterfactual augmentation of system exposure
Zhao, Z., Ren, Z., Yang, J., Yan, Z., Wang, Z., Yang, L., Ren, P., Chen, Z., de Rijke, M., and Xin, X · 2025
Closest in time.
SkillWeaver: Web agents can self-improve by discovering and honing skills
Zheng, B., Fatemi, M. Y., Jin, X., Wang, Z. Z., Gandhi, A., Song, Y., Gu, Y., Srinivasa, J., Liu, G., Neubig, G., et al · 2025
Closest in time.