Fetching the paper…
Reading the bibliography…
Test-driven development (TDD) is the practice of writing tests first and coding later, and the proponents of TDD expound its numerous benefits.
K. Beck, Test driven development: By example . Addison-Wesley Professional, 2002
2002
Earlier work this paper cites.
T. F. Bissyandé, D. Lo, L. Jiang, L. Réveillere, J. Klein, and Y. Le Traon, “Got issues? who cares about it? a large scale investigation of issue trackers from github,” in 2013 IEEE 24th international symposium on software reliability engineering (ISSRE) . IEEE, 2013, pp. 188–197
2013
Earlier work this paper cites.
R. Just, D. Jalali, and M. D. Ernst, “Defects4j: A database of existing faults to enable controlled testing studies for java programs,” in Proceedings of the 2014 international symposium on software testing and analysis , 2014, pp. 437–440
2014
Earlier work this paper cites.
M. Martinez, T. Durieux, R. Sommerard, J. Xuan, and M. Monperrus, “Automatic repair of real bugs in java: A large-scale experiment on the defects4j dataset,” Empirical Software Engineering , vol. 22, pp. 1936–1964, 2017
2017
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
X. Tan, M. Zhou, and Z. Sun, “A first look at good first issues on github,” in ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2020, pp. 398–409
2020
Earlier work this paper cites.
J. Zhou, S. Wang, C.-P. Bezemer, Y. Zou, and A. E. Hassan, “Studying the association between bountysource bounties and the issue-addressing likelihood of github issue reports,” IEEE Transactions on Software Engineering , vol. 47, no. 12, pp. 2919–2933, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
V. J. Hellendoorn, J. Tsay, M. Mukherjee, and M. Hirzel, “Towards automating code review at scale,” in Proceedings of the 29th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering , ser. ESEC/FSE 2021. New York, NY, USA: Association for Computing Machinery, 2021, p. 1479–1482. [Online]. Available: https://doi.org/10.1145/3468264.3473134
2021
Earlier work this paper cites.
M. Golzadeh, A. Decan, D. Legay, and T. Mens, “A ground-truth dataset and classification model for detecting bots in github issue and pr comments,” Journal of Systems and Software , vol. 175, p. 110911, 2021
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
T. Ahmed and P. Devanbu, “Few-shot training LLMs for project-specific code-summarization,” in Proceedings of the 37th IEEE/ACM International Conference on Automated Software Engineering , 2022, pp. 1–5
2022
Earlier work this paper cites.
Q.-C. Bui, R. Scandariato, and N. E. D. Ferreyra, “Vul4j: A dataset of reproducible java vulnerabilities geared towards the study of program repair techniques,” in Proceedings of the 19th International Conference on Mining Software Repositories , 2022, pp. 464–468
2022
Earlier work this paper cites.
J. Wu, H. He, W. Xiao, K. Gao, and M. Zhou, “Demystifying software release note issues on github,” in Proceedings of the 30th IEEE/ACM International Conference on Program Comprehension , 2022, pp. 602–613
2022
Earlier work this paper cites.
S. Kang, J. Yoon, and S. Yoo, “Large language models are few-shot testers: Exploring LLM-based general bug reproduction,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 2312–2323
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
N. Nashid, M. Sintaha, and A. Mesbah, “Retrieval-based prompt selection for code-related few-shot learning,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 2450–2462
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Schäfer, S. Nadi, A. Eghbali, and F. Tip, “An empirical evaluation of using large language models for automated unit test generation,” IEEE Transactions on Software Engineering , 2023
2023
Cited alongside, same era.
C. E. Jimenez, J. Yang, A. Wettig, S. Yao, K. Pei, O. Press, and K. R. Narasimhan, “Swe-bench: Can language models resolve real-world github issues?” in The Twelfth International Conference on Learning Representations , 2024
2024
Cited alongside, same era.
N. Chowdhury, J. Aung, C. J. Shern, O. Jaffe, D. Sherburn, G. Starace, E. Mays, R. Dias, M. Aljubeh, M. Glaese, C. E. Jimenez, J. Yang, K. Liu, and A. Madry, “Introducing SWE-bench Verified,” Aug. 2024. [Online]. Available: https://openai.com/index/introducing-swe-bench-verified/
2024
Cited alongside, same era.
L. Plein, W. C. Ouédraogo, J. Klein, and T. F. Bissyandé, “Automatic generation of test cases based on bug reports: a feasibility study with large language models,” in International Conference on Software Engineering: Companion Proceedings (ICSE-Companion) , May 2024, pp. 360–361. [Online]. Available: https://doi.org/10.1145/3639478.3643119
2024
Cited alongside, same era.
N. Mündler, M. N. Müller, J. He, and M. Vechev, “SWT-bench: Testing and validating real-world bug-fixes with code agents,” in Advances in Neural Information Processing Systems (NeurIPS) , Dec. 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
T. Ahmed, K. S. Pai, P. Devanbu, and E. Barr, “Automatic semantic augmentation of language model prompts (for code summarization),” in Proceedings of the IEEE/ACM 46th International Conference on Software Engineering , 2024, pp. 1–13
2024
Cited alongside, same era.
2024
Closest in time.
N. Wadhwa, A. Sonwane, D. Arora, A. Mehrotra, S. Utpala, R. B. Bairi, A. Kanade, and N. Natarajan, “Masai: Modular architecture for software-engineering ai agents,” in NeurIPS 2024 Workshop on Open-World Agents
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Y. Zhang, H. Ruan, Z. Fan, and A. Roychoudhury, “AutoCodeRover: Autonomous program improvement,” in International Symposium on Software Testing and Analysis (ISSTA) , 2024, pp. 1592–1604. [Online]. Available: https://doi.org/10.1145/3650212.3680384
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
M. Smytzek, M. Eberlein, B. Serce, L. Grunske, and A. Zeller, “Tests4py: A benchmark for system testing,” in Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering , 2024, pp. 557–561
2024
Closest in time.
2024
Closest in time.
J. Wang, Y. Huang, C. Chen, Z. Liu, S. Wang, and Q. Wang, “Software testing with large language models: Survey, landscape, and vision,” IEEE Transactions on Software Engineering , 2024
2024
Closest in time.
G. Ryan, S. Jain, M. Shang, S. Wang, X. Ma, M. K. Ramanathan, and B. Ray, “Code-aware prompting: A study of coverage guided test generation in regression setting using LLM,” FSE , 2024
2024
Closest in time.