Fetching the paper…
Reading the bibliography…
Large Foundation Models (LFMs) have unlocked new possibilities in human-computer interaction, particularly with the rise of mobile Graphical User Interface (GUI) Agents capable of interacting with mobile GUIs.
The art of Prolog (2nd ed.): advanced programming techniques
Leon Sterling and Ehud Shapiro · 1994
Earlier work this paper cites.
Foundations of Databases: The Logical Level
Serge Abiteboul, Richard Hull, and Victor Vianu · 1995
Earlier work this paper cites.
Introduction to Formal Hardware Verification: Methods and Tools for Designing Correct Circuits and Systems
Thomas Kropf · 1999
Earlier work this paper cites.
A static analyzer for large safety-critical software
Bruno Blanchet, Patrick Cousot, Radhia Cousot, Jérôme Feret, Laurent Mauborgne, Antoine Miné, David Monniaux, and Xavier Rival · 2007
Earlier work this paper cites.
Principles of Model Checking
Christel Baier and Joost-Pieter Katoen · 2008
Earlier work this paper cites.
Z3: An Efficient SMT Solver
Leonardo de Moura and Nikolaj Bjørner · 2008
Earlier work this paper cites.
The art of software testing
Glenford J Myers, Corey Sandler, and Tom Badgett · 2011
Earlier work this paper cites.
Systems and Software Verification: Model-Checking Techniques and Tools
B. Berard, P. McKenzie, M. Bidoit, A. Finkel, F. Laroussinie, A. Petit, L. Petrucci, and P. Schnoebelen · 2014
Earlier work this paper cites.
The seahorn verification framework
Arie Gurfinkel, Temesghen Kahsai, Anvesh Komuravelli, and Jorge A. Navas · 2015
Earlier work this paper cites.
Horn Clause Solvers for Program Verification , pages 24–51
Nikolaj Bjørner, Arie Gurfinkel, Ken McMillan, and Andrey Rybalchenko · 2015
Earlier work this paper cites.
Modelplex: Verified runtime validation of verified cyber-physical system models
Stefan Mitsch and André Platzer · 2016
Earlier work this paper cites.
Jayhorn: A framework for verifying java programs
Temesghen Kahsai, Philipp Rümmer, Huascar Sanchez, and Martin Schäf · 2016
Earlier work this paper cites.
Model checking, 2nd Edition
Edmund M. Clarke, Orna Grumberg, Daniel Kroening, Doron A. Peled, and Helmut Veith · 2018
Earlier work this paper cites.
Veriphy: verified controller executables from verified cyber-physical system models
Rose Bohrer, Yong Kiam Tan, Stefan Mitsch, Magnus O. Myreen, and André Platzer · 2018
Earlier work this paper cites.
Introduction to runtime verification
Ezio Bartocci, Yliès Falcone, Adrian Francalanza, and Giles Reger · 2018
Earlier work this paper cites.
Assessing the factual accuracy of generated text
Ben Goodrich, Vrindavan Rao, Peter J. Liu, and Mina Saleh · 2019
Earlier work this paper cites.
Software verification and validation of safe autonomous cars: A systematic literature review
Nijat Rajabli, Francesco Flammini, Roberto Nardone, and Valeria Vittorini · 2020
Earlier work this paper cites.
Rusthorn: Chc-based verification for rust programs
Yusuke Matsushita, Takeshi Tsukada, and Naoki Kobayashi · 2021
Earlier work this paper cites.
Entity-level factual consistency of abstractive text summarization, 2021
Feng Nan, Ramesh Nallapati, Zhiling Wang, Cicero Nogueira dos Santos, Huan Zhu, Dong Zhang, Kathy McKeown, and Bing Xiang · 2021
Earlier work this paper cites.
Retrieval augmentation reduces hallucination in conversation, 2021
Kurt Shuster, Spencer Poff, M. Chen, Douwe Kiela, and Jason Weston · 2021
Cited alongside, same era.
Autoformalization with large language models
Yuhuai Wu, Albert Qiaochu Jiang, Wenda Li, Markus N. Rabe, Charles Staats, Mateja Jamnik, and Christian Szegedy · 2022
Cited alongside, same era.
Creating safe agi that benefits all of humanity, 2023
openai · 2023
Cited alongside, same era.
Talk to claude, 2023
anthropic · 2023
Cited alongside, same era.
Introducing llama 2, 2023
meta · 2023
Cited alongside, same era.
Empowering llm to use smartphone for intelligent task automation
Hao Wen, Yuanchun Li, Guohong Liu, Shanhui Zhao, Tao Yu, Toby Jia-Jun Li, Shiqi Jiang, Yunhao Liu, Yaqin Zhang, and Yunxin Liu · 2023
Cited alongside, same era.
Gpt-4v(ision) is a generalist web agent, if grounded
Boyuan Zheng, Boyu Gou, Jihyung Kil, Huan Sun, and Yu Su · 2024
Later among the works it cites.
Androidworld: A dynamic benchmarking environment for autonomous agents, 2024
Christopher Rawles, Sarah Clinckemaillie, Yifan Chang, Jonathan Waltz, Gabrielle Lau, Marybeth Fair, Alice Li, William Bishop, Wei Li, Folawiyo Campbell-Ajala, Daniel Toyama, Robert Berry, Divya Tyamagundlu, Timothy Lillicrap, and Oriana Riva · 2024
Later among the works it cites.
Cogagent: A visual language model for gui agents, 2024
Wenyi Hong, Weihan Wang, Qingsong Lv, Jiazheng Xu, Wenmeng Yu, Junhui Ji, Yan Wang, Zihan Wang, Yuxuan Zhang, Juanzi Li, Bin Xu, Yuxiao Dong, Ming Ding, and Jie Tang · 2024
Later among the works it cites.
Understanding the weakness of large language model agents within a complex android environment, 2024
Mingzhe Xing, Rongkai Zhang, Hui Xue, Qi Chen, Fan Yang, and Zhen Xiao · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Appagent: Multimodal agents as smartphone users, 2023
Chi Zhang, Zhao Yang, Jiaxuan Liu, Yucheng Han, Xin Chen, Zebiao Huang, Bin Fu, and Gang Yu · 2023
Cited alongside, same era.
Grounding complex natural language commands for temporal tasks in unseen environments
Jason Xinyu Liu, Ziyi Yang, Ifrah Idrees, Sam Liang, Benjamin Schornstein, Stefanie Tellex, and Ankit Shah · 2023
Cited alongside, same era.
Chain-of-verification reduces hallucination in large language models, 2023
Shehzaad Dhuliawala, Mojtaba Komeili, Jing Xu, Roberta Raileanu, Xian Li, Asli Celikyilmaz, and Jason Weston · 2023
Cited alongside, same era.
Haoqiang Kang, Juntong Ni, and Huaxiu Yao · 2023
Cited alongside, same era.
Self-refine: Iterative refinement with self-feedback, 2023
Aman Madaan, Niket Tandon, Prakhar Gupta, et al · 2023
Cited alongside, same era.
FactScore: Fine-grained atomic evaluation of factual precision in long form text generation, 2023
Sewon Min, Kalpesh Krishna, Ximing Lyu, Mike Lewis, Wen-tau Yih, Pang Wei Koh, Mohit Iyyer, Luke Zettlemoyer, and Hannaneh Hajishirzi · 2023
Cited alongside, same era.
Llamatouch: A faithful and scalable testbed for mobile ui task automation
Li Zhang, Shihe Wang, Xianqing Jia, Zhihan Zheng, Yunhe Yan, Longxi Gao, Yuanchun Li, and Mengwei Xu · 2024
Later among the works it cites.
Plug in the safety chip: Enforcing constraints for llm-driven robot agents
Ziyi Yang, Shreyas Sundara Raman, Ankit Shah, and Stefanie Tellex · 2024
Later among the works it cites.
SELP: generating safe and efficient task plans for robot agents with large language models
Yi Wu, Zikang Xiong, Yiran Hu, Shreyash S. Iyengar, Nan Jiang, Aniket Bera, Lin Tan, and Suresh Jagannathan · 2024
Later among the works it cites.
Large multimodal agents: A survey, 2024
Junlin Xie, Zhihong Chen, Ruifei Zhang, Xiang Wan, and Guanbin Li · 2024
Later among the works it cites.
Mengjia Niu, Hao Li, Jie Shi, Hamed Haddadi, and Fan Mo · 2024
Later among the works it cites.
Introducing operator, 2025
openai · 2025
Closest in time.
Computer use (beta), 2025
anthropic · 2025
Closest in time.
Ike Obi, Vishnunandan LN Venkatesh, Weizheng Wang, Ruiqi Wang, Dayoon Suh, Temitope I Amosa, Wonse Jo, and Byung-Cheol Min · 2025
Closest in time.
Llm-powered gui agents in phone automation: Surveying progress and prospects
William Liu, Liang Liu, Yaxuan Guo, Han Xiao, Weifeng Lin, Yuxiang Chai, Shuai Ren, Xiaoyu Liang, Linghao Li, Wenhao Wang, et al · 2025
Closest in time.
Large language model-brained gui agents: A survey, 2025
Chaoyun Zhang, Shilin He, Jiaxu Qian, Bowen Li, Liqun Li, Si Qin, Yu Kang, Minghua Ma, Guyue Liu, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang, and Qi Zhang · 2025
Closest in time.
Gui agents with foundation models: A comprehensive survey, 2025
Shuai Wang, Weiwen Liu, Jingxuan Chen, Yuqi Zhou, Weinan Gan, Xingshan Zeng, Yuhan Che, Shuai Yu, Xinlong Hao, Kun Shao, Bin Wang, Chuhan Wu, Yasheng Wang, Ruiming Tang, and Jianye Hao · 2025
Closest in time.
State, 2025
Meta · 2025
Closest in time.
State and jetpack compose, 2025
Google · 2025
Closest in time.
State, 2025
Apple · 2025
Closest in time.