Fetching the paper…
Reading the bibliography…
Mobile phone agents can assist people in automating daily tasks on their phones, which have emerged as a pivotal research spotlight.
K. Lieberherr, I. Holland, and A. Riel, “Object-oriented programming: An objective sense of style,” ACM Sigplan Notices , vol. 23, no. 11, pp. 323–334, 1988
1988
Earlier work this paper cites.
S. Robertson, H. Zaragoza et al. , “The probabilistic relevance framework: Bm25 and beyond,” Foundations and Trends® in Information Retrieval , vol. 3, no. 4, pp. 333–389, 2009
2009
Earlier work this paper cites.
Y. Li, J. He, X. Zhou, Y. Zhang, and J. Baldridge, “Mapping natural language instructions to mobile ui action sequences,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , 2020, pp. 8198–8210
2020
Earlier work this paper cites.
Y. Li, G. Li, L. He, J. Zheng, H. Li, and Z. Guan, “Widget captioning: Generating natural language description for mobile user interface elements,” in Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2020, pp. 5495–5510
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
B. Wang, G. Li, X. Zhou, Z. Chen, T. Grossman, and Y. Li, “Screen2words: Automatic mobile ui summarization with multimodal learning,” in The 34th Annual ACM Symposium on User Interface Software and Technology , 2021, pp. 498–510
2021
Earlier work this paper cites.
T. J.-J. Li, L. Popowski, T. Mitchell, and B. A. Myers, “Screen2vec: Semantic embedding of gui screens and gui components,” in Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems , 2021, pp. 1–15
2021
Earlier work this paper cites.
X. Zhang, L. De Greef, A. Swearngin, S. White, K. Murray, L. Yu, Q. Shan, J. Nichols, J. Wu, C. Fleizach et al. , “Screen recognition: Creating accessibility metadata for mobile applications from pixels,” in Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems , 2021, pp. 1–15
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Chen, A. Swearngin, J. Wu, T. Barik, J. Nichols, and X. Zhang, “Towards complete icon labeling in mobile applications,” in Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems , 2022, pp. 1–14
2022
Earlier work this paper cites.
A. Burns, D. Arsan, S. Agrawal, R. Kumar, K. Saenko, and B. A. Plummer, “A dataset for interactive vision-language navigation with unknown command feasibility,” in European Conference on Computer Vision . Springer, 2022, pp. 312–328
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
L. Sun, X. Chen, L. Chen, T. Dai, Z. Zhu, and K. Yu, “Meta-gui: Towards multi-modal conversational agents on mobile gui,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 6699–6712
2022
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
G. Li and Y. Li, “Spotlight: Mobile ui understanding using vision-language models with a focus,” in The Eleventh International Conference on Learning Representations , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
B. Wang, G. Li, and Y. Li, “Enabling conversational interaction with mobile ui using large language models,” in Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems , 2023, pp. 1–17
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
A. Team, “Autogpt,” 2023. [Online]. Available: https://github.com/Significant-Gravitas/AutoGPT
2023
Cited alongside, same era.
2023
Cited alongside, same era.
G. Li, H. Hammoud, H. Itani, D. Khizbullin, and B. Ghanem, “Camel: Communicative agents for" mind" exploration of large language model society,” Advances in Neural Information Processing Systems , vol. 36, pp. 51 991–52 008, 2023
H. Wen, Y. Li, G. Liu, S. Zhao, T. Yu, T. J.-J. Li, S. Jiang, Y. Liu, Y. Zhang, and Y. Liu, “Autodroid: Llm-powered task automation in android,” in Proceedings of the 30th Annual International Conference on Mobile Computing and Networking , 2024, pp. 543–557
2024
Later among the works it cites.
Y. Ge, W. Hua, K. Mei, J. Tan, S. Xu, Z. Li, Y. Zhang et al. , “Openagi: When llm meets domain experts,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Later among the works it cites.
Y. Shen, K. Song, X. Tan, D. Li, W. Lu, and Y. Zhuang, “Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Later among the works it cites.
T. Guo, X. Chen, Y. Wang, R. Chang, S. Pei, N. Chawla, O. Wiest, and X. Zhang, “Large language model based multi-agents: A survey of progress and challenges.” in 33rd International Joint Conference on Artificial Intelligence (IJCAI 2024) . IJCAI; Cornell arxiv, 2024
2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
J. S. Park, J. O’Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein, “Generative agents: Interactive simulacra of human behavior,” in Proceedings of the 36th annual acm symposium on user interface software and technology , 2023, pp. 1–22
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Later among the works it cites.
S. Hong, M. Zhuge, J. Chen, X. Zheng, Y. Cheng, J. Wang, C. Zhang, Z. Wang, S. K. S. Yau, Z. Lin et al. , “Metagpt: Meta programming for a multi-agent collaborative framework,” in The Twelfth International Conference on Learning Representations , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
A. Zhang, Y. Chen, L. Sheng, X. Wang, and T.-S. Chua, “On generative agents in recommendation,” in Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval , 2024, pp. 1807–1817
2024
Later among the works it cites.
H. Sun, Y. Liu, C. Wu, H. Yan, C. Tai, X. Gao, S. Shang, and R. Yan, “Harnessing multi-role capabilities of large language models for open-domain question answering,” in Proceedings of the ACM on Web Conference 2024 , 2024, pp. 4372–4382
2024
Later among the works it cites.
H. Sun, W. Xu, W. Liu, J. Luan, B. Wang, S. Shang, J.-R. Wen, and R. Yan, “Determlr: Augmenting llm-based logical reasoning from indeterminacy to determinacy,” in Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2024, pp. 9828–9862
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
J.-t. Huang, E. J. Li, M. H. Lam, T. Liang, W. Wang, Y. Yuan, W. Jiao, X. Wang, Z. Tu, and M. R. Lyu, “How far are we on the decision-making of llms? evaluating llms’ gaming ability in multi-agent environments,” CoRR , 2024
2024
Later among the works it cites.
Q. Wu, W. Xu, W. Liu, T. Tan, L. Liujianfeng, A. Li, J. Luan, B. Wang, and S. Shang, “Mobilevlm: A vision-language model for better intra-and inter-ui understanding,” in Findings of the Association for Computational Linguistics: EMNLP 2024 , 2024, pp. 10 231–10 251
2024
Later among the works it cites.
M. Xing, R. Zhang, H. Xue, Q. Chen, F. Yang, and Z. Xiao, “Understanding the weakness of large language model agents within a complex android environment,” in Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining , 2024, pp. 6061–6072
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Liu, H. Sun, W. Guo, X. Xiao, C. Mao, Z. Yu, and R. Yan, “Bidev: Bilateral defusing verification for complex claim fact-checking,” in Proceedings of the AAAI Conference on Artificial Intelligence , 2025
2025
Closest in time.