Fetching the paper…
Reading the bibliography…
Reproducing game bugs, particularly crash bugs in continuously evolving games like Minecraft, is a notoriously manual, time-consuming, and challenging process to automate; insights from a key decision maker from Minecraft we interviewed confirm this, highlighting that a substantial portion of crash reports necessitate manual scenario reconstruction.
Zimmermann, T., Premraj, R., Bettenburg, N., Just, S., Schroter, A., and Weiss, C · 2010
Earlier work this paper cites.
Interrater reliability: the kappa statistic
McHugh, M. L · 2012
Earlier work this paper cites.
Prismarinejs/mineflayer: Create minecraft bots with a powerful, stable, and high level javascript api., 2013
PrismarineJS · 2013
Earlier work this paper cites.
How is video game development different from software development in open source?
Pascarella, L., Palomba, F., Di Penta, M., and Bacchelli, A · 2018
Earlier work this paper cites.
Wuji: Automatic online combat game testing using evolutionary deep reinforcement learning
Zheng, Y., Xie, X., Su, T., Ma, L., Hao, J., Meng, Z., Liu, Y., Shen, R., Chen, Y., and Fan, C · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D · 2020
Earlier work this paper cites.
The significance of bug report elements
Soltani, M., Hermans, F., and Bäck, T · 2020
Earlier work this paper cites.
A survey of video game testing
Politowski, C., Petrillo, F., and Gueheneuc, Y.-G · 2021
Earlier work this paper cites.
We’ll fix it in post: What do bug fixes in video game update notes tell us?
Truelove, A., Santana de Almeida, E., and Ahmed, I · 2021
Cited alongside, same era.
Game updates enhance players’ engagement: a case of dota2
Zhong, X., and Xu, J · 2021
Cited alongside, same era.
Large language models are few-shot testers: Exploring llm-based general bug reproduction
Kang, S., Yoon, J., and Yoo, S · 2023
Cited alongside, same era.
Self-refine: iterative refinement with self-feedback
Madaan, A., Tandon, N., Gupta, P., Hallinan, S., Gao, L., Wiegreffe, S., Alon, U., Dziri, N., Prabhumoye, S., Yang, Y., Gupta, S., Majumder, B. P., Hermann, K., Welleck, S., Yazdanbakhsh, A., and Clark, P · 2023
Cited alongside, same era.
The rise and potential of large language model based agents: A survey, 2023
Xi, Z., Chen, W., Guo, X., He, W., Ding, Y., Hong, B., Zhang, M., Wang, J., Jin, S., Zhou, E., Fan, X., Wang, X., Xiong, L., Zhou, Y., Wang, W., Jiang, C., Zou, Y., Liu, X., Yin, Z., Dou, S., Weng, R., Cheng, W., Zhang, Q., Qin, W., Zheng, Y., Qiu, X., Huang, X., and Gui, T · 2023
Cited alongside, same era.
A survey on large language models for code generation, 2024
Jiang, J., Wang, F., Shen, J., Kim, S., and Kim, S · 2024
Later among the works it cites.
Vision-driven automated mobile gui testing via multimodal large language model
Liu, Z., Li, C., Chen, C., Wang, J., Wu, B., Wang, Y., Hu, J., and Wang, Q · 2024
Later among the works it cites.
Omniparser for pure vision based gui agent, 2024
Lu, Y., Yang, J., Shen, Y., and Awadallah, A · 2024
Later among the works it cites.
GPT-4o System Card, 2024
OpenAI · 2024
Later among the works it cites.
Tool learning with foundation models, 2024
Qin, Y., Hu, S., Lin, Y., Chen, W., Ding, N., Cui, G., Zeng, Z., Huang, Y., Xiao, C., Han, C., Fung, Y. R., Su, Y., Wang, H., Qian, C., Tian, R., Zhu, K., Liang, S., Shen, X., Xu, B., Zhang, Z., Ye, Y., Li, B., Tang, Z., Yi, J., Zhu, Y., Dai, Z., Yan, L., Cong, X., Lu, Y., Zhao, W., Huang, Y., Yan, J., Han, X., Sun, X., Li, D., Phang, J., Yang, C., Wu, T., Ji, H., Liu, Z., and Sun, M · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
ReAct: Synergizing reasoning and acting in language models
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K., and Cao, Y · 2023
Cited alongside, same era.
Chatbr: Automated assessment and improvement of bug report quality using chatgpt
Bo, L., Ji, W., Sun, X., Zhang, T., Wu, X., and Wei, Y · 2024
Cited alongside, same era.
Prompting is all you need: Automated android bug replay with large language models
Feng, S., and Chen, C · 2024
Cited alongside, same era.
Voyager: An open-ended embodied agent with large language models
Wang, G., Xie, Y., Jiang, Y., Mandlekar, A., Xiao, C., Zhu, Y., Fan, L., and Anandkumar, A
Cited in the paper.
The effect of sampling temperature on problem solving in large language models, 2024
Renze, M., and Guven, E · 2024
Later among the works it cites.
Feedback-driven automated whole bug report reproduction for android apps
Wang, D., Zhao, Y., Feng, S., Zhang, Z., Halfond, W. G. J., Chen, C., Sun, X., Shi, J., and Yu, T · 2024
Later among the works it cites.
Improving retrieval augmented language model with self-reasoning, 2024
Xia, Y., Zhou, J., Shi, Z., Chen, J., and Huang, H · 2024
Later among the works it cites.