Fetching the paper…
Reading the bibliography…
We introduce NNetNav, a method for unsupervised interaction with websites that generates synthetic demonstrations for training browser agents.
Understanding natural language
Winograd, T · 1972
Earlier work this paper cites.
Plow: a collaborative task learning agent
Allen, J., Chambers, N., Ferguson, G., Galescu, L., Jung, H., Swift, M., and Taysom, W · 2007
Earlier work this paper cites.
Reinforcement learning for mapping instructions to actions
Branavan, S., Chen, H., Zettlemoyer, L., and Barzilay, R · 2009
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
Chen, D. and Mooney, R · 2011
Earlier work this paper cites.
Guiding reinforcement learning exploration using natural language
Harrison, B., Ehsan, U., and Riedl, M. O · 2017
Earlier work this paper cites.
Mapping instructions and visual observations to actions with reinforcement learning
Misra, D., Langford, J., and Artzi, Y · 2017
Earlier work this paper cites.
World of bits: An open-domain platform for web-based agents
Shi, T., Karpathy, A., Fan, L., Hernandez, J., and Liang, P · 2017
Earlier work this paper cites.
Reinforcement learning on web interfaces using workflow-guided exploration
Liu, E. Z., Guu, K., Pasupat, P., Shi, T., and Liang, P · 2018
Earlier work this paper cites.
The curious case of neural text degeneration
Holtzman, A., Buys, J., Du, L., Forbes, M., and Choi, Y · 2019
Earlier work this paper cites.
Zero: Memory optimizations toward training trillion parameter models
Rajbhandari, S., Rasley, J., Ruwase, O., and He, Y · 2020
Earlier work this paper cites.
A data-driven approach for learning to control computers
Humphreys, P. C., Raposo, D., Pohlen, T., Thornton, G., Chhaparia, R., Muldal, A., Abramson, J., Georgiev, P., Santoro, A., and Lillicrap, T · 2022
Earlier work this paper cites.
Improving intrinsic exploration with language abstractions
Mu, J., Zhong, V., Raileanu, R., Jiang, M., Goodman, N., Rocktäschel, T., and Grefenstette, E · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., ichter, b., Xia, F., Chi, E., Le, Q. V., and Zhou, D · 2022
Earlier work this paper cites.
ReAct: Synergizing reasoning and acting in language models
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K. R., and Cao, Y · 2022
Cited alongside, same era.
Out of one, many: Using language models to simulate human samples
Argyle, L. P., Busby, E. C., Fulda, N., Gubler, J. R., Rytting, C., and Wingate, D · 2023
Cited alongside, same era.
Guiding pretraining in reinforcement learning with large language models
Du, Y., Watkins, O., Wang, Z., Colas, C., Darrell, T., Abbeel, P., Gupta, A., and Andreas, J · 2023
Cited alongside, same era.
Multimodal web navigation with instruction-finetuned foundation models
Furuta, H., Nachum, O., Lee, K.-H., Matsuo, Y., Gu, S. S., and Gur, I · 2023
Cited alongside, same era.
Language models can solve computer tasks
Kim, G., Baldi, P., and McAleer, S · 2023
Dubey, A., Jauhri, A., Pandey, A., Kadian, A., Al-Dahle, A., Letman, A., Mathur, A., Schelten, A., Yang, A., Fan, A., et al · 2024
Closest in time.
Webvoyager: Building an end-to-end web agent with large multimodal models
He, H., Yao, W., Ma, K., Yu, W., Dai, Y., Zhang, H., Lan, Z., and Yu, D · 2024
Closest in time.
Autowebglm: Bootstrap and reinforce a large language model-based web navigating agent, 2024
Lai, H., Liu, X., Iong, I. L., Yao, S., Chen, Y., Shen, P., Yu, H., Zhang, H., Zhang, X., Dong, Y., and Tang, J · 2024
Closest in time.
Weblinx: Real-world website navigation with multi-turn dialogue
Lù, X. H., Kasner, Z., and Reddy, S · 2024
Closest in time.
BAGEL: Bootstrapping agents by guiding exploration with language
Murty, S., Manning, C. D., Shaw, P., Joshi, M., and Lee, K · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Efficient memory management for large language model serving with pagedattention
Kwon, W., Li, Z., Zhuang, S., Sheng, Y., Zheng, L., Yu, C. H., Gonzalez, J. E., Zhang, H., and Stoica, I · 2023
Cited alongside, same era.
Role play with large language models
Shanahan, M., McDonell, K., and Reynolds, L · 2023
Cited alongside, same era.
Reflexion: an autonomous agent with dynamic memory and self-reflection
Shinn, N., Labash, B., and Gopinath, A · 2023
Cited alongside, same era.
Heap: Hierarchical policies for web actions using llms
Sodhi, P., Branavan, S., and McDonald, R · 2023
Cited alongside, same era.
Adaplanner: Adaptive planning from feedback with language models
Sun, H., Zhuang, Y., Kong, L., Dai, B., and Zhang, C · 2023
Cited alongside, same era.
How far can camels go? exploring the state of instruction tuning on open resources, 2023
Wang, Y., Ivison, H., Dasigi, P., Hessel, J., Khot, T., Chandu, K. R., Wadden, D., MacMillan, K., Smith, N. A., Beltagy, I., and Hajishirzi, H · 2023
Cited alongside, same era.
Webarena: A realistic web environment for building autonomous agents
Zhou, S., Xu, F. F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Bisk, Y., Fried, D., Alon, U., et al · 2023
Cited alongside, same era.
Closest in time.
Synatra: Turning indirect knowledge into direct demonstrations for digital agents at scale
Ou, T., Xu, F. F., Madaan, A., Liu, J., Lo, R., Sridhar, A., Sengupta, S., Roth, D., Neubig, G., and Zhou, S · 2024
Closest in time.
Large language models can self-improve at web agent tasks
Patel, A., Hofmarcher, M., Leoveanu-Condrei, C., Dinu, M.-C., Callison-Burch, C., and Hochreiter, S · 2024
Closest in time.
Scribeagent: Towards specialized web agents using production-scale workflow data
Shen, J., Jain, A., Xiao, Z., Amlekar, I., Hadji, M., Podolny, A., and Talwalkar, A · 2024
Closest in time.
Wang, Z. Z., Mao, J., Fried, D., and Neubig, G · 2024
Closest in time.
Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments
Xie, T., Zhang, D., Chen, J., Li, X., Zhao, S., Cao, R., Hua, T. J., Cheng, Z., Shin, D., Lei, F., et al · 2024
Closest in time.
Agenttrek: Agent trajectory synthesis via guiding replay with web tutorials
Xu, Y., Lu, D., Shen, Z., Wang, J., Wang, Z., Mao, Y., Xiong, C., and Yu, T · 2024
Closest in time.
Proposer-agent-evaluator (pae): Autonomous skill discovery for foundation model internet agents
Zhou, Y., Yang, Q., Lin, K., Bai, M., Zhou, X., Wang, Y.-X., Levine, S., and Li, E · 2024
Closest in time.