Fetching the paper…
Reading the bibliography…
Recently, large language models (LLMs) have demonstrated exceptional capabilities in serving as the foundation for AI assistants.
zxcvbn: { \{ Low-Budget } \} Password Strength Estimation. In 25th USENIX Security Symposium (USENIX Security 16) . 157–173
Daniel Lowe Wheeler. 2016 · 2016
Earlier work this paper cites.
Rico: A mobile app dataset for building data-driven design applications. In Proceedings of the 30th Annual ACM Symposium on User Interface Software and Technology . 845–854
Biplab Deka, Zifeng Huang, Chad Franzen, Joshua Hibschman, Daniel Afergan, Yang Li, Jeffrey Nichols, and Ranjitha Kumar. 2017 · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Humanoid: A deep learning-based approach to automated black-box android app testing. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 1070–1073
Yuanchun Li, Ziyue Yang, Yao Guo, and Xiangqun Chen. 2019 · 2019
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Adam Roberts, Colin Raffel, Katherine Lee, Michael Matena, Noam Shazeer, Peter J Liu, Sharan Narang, Wei Li, and Yanqi Zhou. 2019 · 2019
Earlier work this paper cites.
Flin: A flexible natural language interface for web navigation
Sahisnu Mazumder and Oriana Riva. 2020 · 2020
Earlier work this paper cites.
Testing apps with real-world inputs. In Proceedings of the IEEE/ACM 1st International Conference on Automation of Software Test . 1–10
Tanapuch Wanwarang, Nataniel P Borges Jr, Leon Bettscheider, and Andreas Zeller. 2020 · 2020
Earlier work this paper cites.
LaMDA: our breakthrough conversation technology
Eli Collins. 2021 · 2021
Earlier work this paper cites.
NLP-assisted web element identification toward script-free testing. In 2021 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 639–643
Hiroyuki Kirinuki, Shinsuke Matsumoto, Yoshiki Higo, and Shinji Kusumoto. 2021 · 2021
Earlier work this paper cites.
Glider: A reinforcement learning approach to extract UI scripts from websites. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval . 1420–1430
Yuanchun Li and Oriana Riva. 2021 · 2021
Earlier work this paper cites.
Cutting down on prompts and parameters: Simple few-shot learning with language models
Robert L Logan IV, Ivana Balažević, Eric Wallace, Fabio Petroni, Sameer Singh, and Sebastian Riedel. 2021 · 2021
Earlier work this paper cites.
Prompt programming for large language models: Beyond the few-shot paradigm. In Extended Abstracts of the 2021 CHI Conference on Human Factors in Computing Systems . 1–7
Laria Reynolds and Kyle McDonell. 2021 · 2021
Earlier work this paper cites.
Fast and reliable end-to-end testing for modern web apps | Playwright
2022 · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al · 2022
Cited alongside, same era.
CrawLabel: computing natural-language labels for UI test cases. In Proceedings of the 3rd ACM/IEEE International Conference on Automation of Software Test . 103–114
Yu Liu, Rahulkrishna Yandrapally, Anup K Kalia, Saurabh Sinha, Rachel Tzoref-Brill, and Ali Mesbah. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
A real-world webagent with planning, long context understanding, and program synthesis
Izzeddin Gur, Hiroki Furuta, Austin Huang, Mustafa Safdari, Yutaka Matsuo, Douglas Eck, and Aleksandra Faust. 2023 · 2023
Later among the works it cites.
Cogagent: A visual language model for gui agents
Wenyi Hong, Weihan Wang, Qingsong Lv, Jiazheng Xu, Wenmeng Yu, Junhui Ji, Yan Wang, Zihan Wang, Yuxiao Dong, Ming Ding, et al · 2023
Later among the works it cites.
Faria Huq, Jeffrey P Bigham, and Nikolas Martelaro. 2023 · 2023
Later among the works it cites.
Prompt Injection attack against LLM-integrated Applications
Yi Liu, Gelei Deng, Yuekang Li, Kailong Wang, Tianwei Zhang, Yepang Liu, Haoyu Wang, Yan Zheng, and Yang Liu. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Common Crawl - Open Repository of Web Crawl Data
2023 · 2023
Cited alongside, same era.
Introducing ChatGPT and Whisper APIs
2023 · 2023
Cited alongside, same era.
WebPilot - Copilot for All
2023 · 2023
Cited alongside, same era.
A More Accessible Web with Natural Language Interface. In Proceedings of the 20th International Web for All Conference . 153–155
Xiang Deng. 2023 · 2023
Cited alongside, same era.
Mind2Web: Towards a Generalist Agent for the Web
Xiang Deng, Yu Gu, Boyuan Zheng, Shijie Chen, Samuel Stevens, Boshi Wang, Huan Sun, and Yu Su. 2023 · 2023
Cited alongside, same era.
Multimodal Web Navigation with Instruction-Finetuned Foundation Models
Hiroki Furuta, Ofir Nachum, Kuang-Huei Lee, Yutaka Matsuo, Shixiang Shane Gu, and Izzeddin Gur. 2023 · 2023
Cited alongside, same era.
Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection. In Proceedings of the 16th ACM Workshop on Artificial Intelligence and Security . 79–90
Kai Greshake, Sahar Abdelnabi, Shailesh Mishra, Christoph Endres, Thorsten Holz, and Mario Fritz. 2023 · 2023
Cited alongside, same era.
Paloma Sodhi, SRK Branavan, and Ryan McDonald. 2023 · 2023
Later among the works it cites.
Large language models as optimizers
Chengrun Yang, Xuezhi Wang, Yifeng Lu, Hanxiao Liu, Quoc V Le, Denny Zhou, and Xinyun Chen. 2023a · 2023
Later among the works it cites.
Set-of-mark prompting unleashes extraordinary visual grounding in gpt-4v
Jianwei Yang, Hao Zhang, Feng Li, Xueyan Zou, Chunyuan Li, and Jianfeng Gao. 2023b · 2023
Later among the works it cites.
Responsible Task Automation: Empowering Large Language Models as Responsible Task Automators
Zhizheng Zhang, Xiaoyi Zhang, Wenxuan Xie, and Yan Lu. 2023 · 2023
Later among the works it cites.
Universal and transferable adversarial attacks on aligned language models
Andy Zou, Zifan Wang, J Zico Kolter, and Matt Fredrikson. 2023 · 2023
Later among the works it cites.
GPT-4V-Act
2024 · 2024
Closest in time.
Exposing Limitations of Language Model Agents in Sequential-Task Compositions on the Web. In ICLR 2024 Workshop on Large Language Model (LLM) Agents
Hiroki Furuta, Yutaka Matsuo, Aleksandra Faust, and Izzeddin Gur. [n. d.] · 2024
Closest in time.
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Hongliang He, Wenlin Yao, Kaixin Ma, Wenhao Yu, Yong Dai, Hongming Zhang, Zhenzhong Lan, and Dong Yu. 2024 · 2024
Closest in time.