Fetching the paper…
Reading the bibliography…
Despite advances in automated testing, manual testing remains prevalent due to the high maintenance demands associated with test script fragility-scripts often break with minor changes in application structure.
The right answer for the wrong question: consequences of type III error for public health research
S Schwartz and K M Carpenter. 1999 · 1999
Earlier work this paper cites.
Automating Web navigation with the WebVCR
Vinod Anupam, Juliana Freire, Bharat Kumar, and Daniel Lieuwen. 2000 · 2000
Earlier work this paper cites.
Experimentation in Software Engineering . The Kluwer International Series in Software Engineering, Vol. 6
Claes Wohlin, Per Runeson, Martin Höst, Magnus C. Ohlsson, Björn Regnell, and Anders Wesslén. 2000 · 2000
Earlier work this paper cites.
WaRR: A tool for high-fidelity web application record and replay. In 2011 IEEE/IFIP 41st International Conference on Dependable Systems & Networks (DSN) . IEEE, Hong Kong, 403–410
Silviu Andrica and George Candea. 2011 · 2011
Earlier work this paper cites.
Reducing Web Test Cases Aging by Means of Robust XPath Locators. In 2014 IEEE International Symposium on Software Reliability Engineering Workshops . IEEE, Naples, Italy, 449–454
Maurizio Leotta, Andrea Stocco, Filippo Ricca, and Paolo Tonella. 2014b · 2014
Earlier work this paper cites.
Using Multi-Locators to Increase the Robustness of Web Test Cases. In 2015 IEEE 8th International Conference on Software Testing, Verification and Validation (ICST) . IEEE, Graz, Austria, 1–10
Maurizio Leotta, Andrea Stocco, Filippo Ricca, and Paolo Tonella. 2015 · 2015
Earlier work this paper cites.
How to Build a Benchmark. In Proceedings of the 6th ACM/SPEC International Conference on Performance Engineering . ACM, Austin Texas, USA, 333–336
Jóakim V. Kistowski, Jeremy A. Arnold, Karl Huppler, Klaus-Dieter Lange, John L. Henning, and Paul Cao. 2015 · 2015
Earlier work this paper cites.
WATERFALL: an incremental approach for repairing record-replay tests of web applications. In Proceedings of the 2016 24th ACM SIGSOFT International Symposium on Foundations of Software Engineering (FSE 2016) . Association for Computing Machinery, Seattle, WA, USA, 751–762
Mouna Hammoudi, Gregg Rothermel, and Andrea Stocco. 2016a · 2016
Earlier work this paper cites.
Why do Record/Replay Tests of Web Applications Break?. In 2016 IEEE International Conference on Software Testing, Verification and Validation (ICST) . IEEE, Chicago, IL, USA, 180–190
Mouna Hammoudi, Gregg Rothermel, and Paolo Tonella. 2016b · 2016
Earlier work this paper cites.
Card-sorting: From text to themes
T. Zimmermann. 2016 · 2016
Earlier work this paper cites.
Chapter Three - Three Open Problems in the Context of E2E Web Testing and a Vision: NEONATE
Filippo Ricca, Maurizio Leotta, and Andrea Stocco. 2019 · 2018
Earlier work this paper cites.
Visual web test repair. In Proceedings of the 2018 26th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering (ESEC/FSE 2018) . Association for Computing Machinery, Lake Buena Vista, FL, USA, 503–514
Andrea Stocco, Rahulkrishna Yandrapally, and Ali Mesbah. 2018 · 2018
Earlier work this paper cites.
Comparing the effort and effectiveness of automated and manual tests. In 2019 14th Iberian Conference on Information Systems and Technologies (CISTI) . IEEE, Coimbra, Portugal, 1–6
Ignacio Dobles, Alexandra Martínez, and Christian Quesada-López. 2019 · 2019
Earlier work this paper cites.
WebRR: self-replay enhanced robust record/replay for web application testing. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering (ESEC/FSE 2020) . Association for Computing Machinery, Virtual Event, USA, 1498–1508
Zhenyue Long, Guoquan Wu, Xiaojiang Chen, Wei Chen, and Jun Wei. 2020 · 2020
Earlier work this paper cites.
Erratum: Leveraging Flexible Tree Matching to repair broken locators in web automation scripts
Sacha Brisset, Romain Rouvoy, Lionel Seinturier, and Renaud Pawlak. 2022 · 2021
Earlier work this paper cites.
Khyathi Raghavi Chandu, Yonatan Bisk, and Alan W. Black. 2021 · 2021
Cited alongside, same era.
An Improving Approach for DOM-Based Web Test Suite Repair. In Web Engineering , Marco Brambilla, Richard Chbeir, Flavius Frasincar, and Ioana Manolescu (Eds.). Springer International Publishing, Biarritz, France, 372–387
Wei Chen, Hanyang Cao, and Xavier Blanc. 2021 · 2021
Cited alongside, same era.
Generating and selecting resilient and maintainable locators for Web automated testing
Vu Nguyen, Thanh To, and Gia-Han Diep. 2021 · 2021
Cited alongside, same era.
WebVLN: Vision-and-Language Navigation on Websites
Qi Chen, Dileepa Pitawela, Chongyang Zhao, Gengze Zhou, Hsiang-Ting Chen, and Qi Wu. 2023 · 2023
Cited alongside, same era.
AgentQuest: A Modular Benchmark Framework to Measure Progress and Improve LLM Agents
Luca Gioacchini, Giuseppe Siracusano, Davide Sanvito, Kiril Gashteovski, David Friede, Roberto Bifulco, and Carolin Lawrence. 2024 · 2024
Later among the works it cites.
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Hongliang He, Wenlin Yao, Kaixin Ma, Wenhao Yu, Yong Dai, Hongming Zhang, Zhenzhong Lan, and Dong Yu. 2024 · 2024
Later among the works it cites.
OpenWebAgent: An Open Toolkit to Enable Web Agents on Large Language Models. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 3: System Demonstrations) , Yixin Cao, Yang Feng, and Deyi Xiong (Eds.). Association for Computational Linguistics, Bangkok, Thailand, 72–81
Iat Long Iong, Xiao Liu, Yuxuan Chen, Hanyu Lai, Shuntian Yao, Pengbo Shen, Hao Yu, Yuxiao Dong, and Jie Tang. 2024 · 2024
Later among the works it cites.
AutoWebGLM: A Large Language Model-based Web Navigating Agent. In Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD ’24) . Association for Computing Machinery, New York, NY, USA, 5295–5306
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xiang Deng, Yu Gu, Boyuan Zheng, Shijie Chen, Samuel Stevens, Boshi Wang, Huan Sun, and Yu Su. 2023 · 2023
Cited alongside, same era.
Towards Autonomous Testing Agents via Conversational Large Language Models
Robert Feldt, Sungmin Kang, Juyeon Yoon, and Shin Yoo. 2023 · 2023
Cited alongside, same era.
Don’t Generate, Discriminate: A Proposal for Grounding Language Models to Real-World Environments
Yu Gu, Xiang Deng, and Yu Su. 2023 · 2023
Cited alongside, same era.
A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
Izzeddin Gur, Hiroki Furuta, Austin Huang, Mustafa Safdari, Yutaka Matsuo, Douglas Eck, and Aleksandra Faust. 2023 · 2023
Cited alongside, same era.
Challenges of End-to-End Testing with Selenium WebDriver and How to Face Them: A Survey. In 2023 IEEE Conference on Software Testing, Verification and Validation (ICST) . IEEE, Dublin, Ireland, 339–350
Maurizio Leotta, Boni García, Filippo Ricca, and Jim Whitehead. 2023 · 2023
Cited alongside, same era.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Cited alongside, same era.
Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V
Jianwei Yang, Hao Zhang, Feng Li, Xueyan Zou, Chunyuan Li, and Jianfeng Gao. 2023 · 2023
Cited alongside, same era.
LLM for Test Script Generation and Migration: Challenges, Capabilities, and Opportunities. In 2023 IEEE 23rd International Conference on Software Quality, Reliability, and Security (QRS) . IEEE, Chiang Mai, Thailand, 206–217
Shengcheng Yu, Chunrong Fang, Yuchen Ling, Chentian Wu, and Zhenyu Chen. 2023 · 2023
Cited alongside, same era.
Hanyu Lai, Xiao Liu, Iat Long Iong, Shuntian Yao, Yuxuan Chen, Pengbo Shen, Hao Yu, Hanchen Zhang, Xiaohan Zhang, Yuxiao Dong, and Jie Tang. 2024 · 2024
Later among the works it cites.
Make LLM a Testing Expert: Bringing Human-like Interaction to Mobile GUI Testing via Functionality-aware Decisions. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering (ICSE ’24) . Association for Computing Machinery, Lisbon, Portugal
Zhe Liu, Chunyang Chen, Junjie Wang, Mengzhuo Chen, Boyu Wu, Xing Che, Dandan Wang, and Qing Wang. 2024 · 2024
Later among the works it cites.
Building Better AI Agents: A Provocation on the Utilisation of Persona in LLM-based Conversational Agents. In ACM Conversational User Interfaces 2024 . ACM, Luxembourg Luxembourg, 1–6
Guangzhi Sun, Xiao Zhan, and Jose Such. 2024 · 2024
Later among the works it cites.
Software Testing With Large Language Models: Survey, Landscape, and Vision
Junjie Wang, Yuchao Huang, Chunyang Chen, Zhe Liu, Song Wang, and Qing Wang. 2024a · 2024
Later among the works it cites.
A Survey on Large Language Model based Autonomous Agents
Lei Wang, Chen Ma, Xueyang Feng, Zeyu Zhang, Hao Yang, Jingsen Zhang, Zhiyuan Chen, Jiakai Tang, Xu Chen, Yankai Lin, Wayne Xin Zhao, Zhewei Wei, and Ji-Rong Wen. 2024b · 2024
Later among the works it cites.
TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks
Frank F. Xu, Yufan Song, Boxuan Li, Yuxuan Tang, Kritanjali Jain, Mengxue Bao, Zora Z. Wang, Xuhui Zhou, Zhitong Guo, Murong Cao, Mingyang Yang, Hao Yang Lu, Amaad Martin, Zhe Su, Leander Maben, Raj Mehta, Wayne Chi, Lawrence Jang, Yiqing Xie, Shuyan Zhou, and Graham Neubig. 2024 · 2024
Later among the works it cites.
On the Evaluation of Large Language Models in Unit Test Generation. In Proceedings of the 39th IEEE/ACM International Conference on Automated Software Engineering (ASE ’24) . Association for Computing Machinery, Sacramento, California, USA, 1607–1619
Lin Yang, Chen Yang, Shutao Gao, Weijing Wang, Bo Wang, Qihao Zhu, Xiao Chu, Jianyi Zhou, Guangtai Liang, Qianxiang Wang, and Junjie Chen. 2024 · 2024
Later among the works it cites.
GPT-4V(ision) is a Generalist Web Agent, if Grounded. In Proceedings of the 41st International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 235) , Ruslan Salakhutdinov, Zico Kolter, Katherine Heller, Adrian Weller, Nuria Oliver, Jonathan Scarlett, and Felix Berkenkamp (Eds.). PMLR, Vienna, Austria, 61349–61385
Boyuan Zheng, Boyu Gou, Jihyung Kil, Huan Sun, and Yu Su. 2024 · 2024
Later among the works it cites.
WebArena: A Realistic Web Environment for Building Autonomous Agents
Shuyan Zhou, Frank F. Xu, Hao Zhu, Xuhui Zhou, Robert Lo, Abishek Sridhar, Xianyi Cheng, Tianyue Ou, Yonatan Bisk, Daniel Fried, Uri Alon, and Graham Neubig. 2024 · 2024
Later among the works it cites.
Agent-as-a-Judge: Evaluate Agents with Agents
Mingchen Zhuge, Changsheng Zhao, Dylan Ashley, Wenyi Wang, Dmitrii Khizbullin, Yunyang Xiong, Zechun Liu, Ernie Chang, Raghuraman Krishnamoorthi, Yuandong Tian, Yangyang Shi, Vikas Chandra, and Jürgen Schmidhuber. 2024 · 2024
Later among the works it cites.
UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Yujia Qin, Yining Ye, Junjie Fang, Haoming Wang, Shihao Liang, Shizuo Tian, Junda Zhang, Jiahao Li, Yunxin Li, Shijue Huang, Wanjun Zhong, Kuanye Li, Jiale Yang, Yu Miao, Woyu Lin, Longxiang Liu, Xu Jiang, Qianli Ma, Jingyu Li, Xiaojun Xiao, Kai Cai, Chuang Li, Yaowei Zheng, Chaolin Jin, Chen Li, Xiao Zhou, Minchao Wang, Haoli Chen, Zhaojian Li, Haihua Yang, Haifeng Liu, Feng Lin, Tao Peng, Xin Liu, and Guang Shi. 2025 · 2025
Closest in time.