Fetching the paper…
Reading the bibliography…
Automated unit test generators, particularly search-based software testing tools like EvoSuite, are capable of generating tests with high coverage.
K. L. Beck, Test-Driven Development - By Example , ser. The Addison-Wesley signature series. Addison-Wesley, 2003
2003
Earlier work this paper cites.
A. Zeller, Why Programs Fail: A Guide to Systematic Debugging . Morgan Kaufmann Publishers Inc., 2005
2005
Earlier work this paper cites.
C. Pacheco and M. D. Ernst, “Randoop: Feedback-directed random testing for java,” in Conf. on Object-Oriented Programming Systems and Applications (OOPSLA-Companion) . ACM, 2007, pp. 815–816
2007
Earlier work this paper cites.
S. Ren, D. Guo, S. Lu, L. Zhou et al. , “Codebleu: a method for automatic evaluation of code synthesis,” arXiv , 2020. [Online]. Available: https://doi.org/10.48550/arXiv.2009.10297
2009
Earlier work this paper cites.
J. Graylin, J. E. Hale, R. K. Smith, H. David et al. , “Cyclomatic complexity and lines of code: empirical evidence of a stable linear relationship,” Journal of Software Engineering and Applications , vol. 2, no. 03, p. 137, 2009
2009
Earlier work this paper cites.
S. Ali, L. C. Briand, H. Hemmati, and R. K. Panesar-Walawege, “A systematic review of the application and empirical investigation of search-based test case generation,” IEEE Trans. Software Eng. , vol. 36, no. 6, pp. 742–762, 2010
2010
Earlier work this paper cites.
L. Baresi and M. Miraz, “Testful: automatic unit-test generation for java classes,” in 32nd IEEE/ACM International Conference on Software Engineering (ICSE) . ACM, 2010, pp. 281–284
2010
Earlier work this paper cites.
G. Fraser and A. Arcuri, “EvoSuite: Automatic test suite generation for object-oriented software,” in Proc. Joint Meeting Symp. Foundations of Software Engineering and the European Softw. Eng. Conf. (ESEC/FSE) . ACM, 2011, pp. 416–419
2011
Earlier work this paper cites.
G. M. Sullivan and R. Feinn, “Using effect size—or why the p value is not enough,” Journal of graduate medical education , vol. 4, no. 3, pp. 279–282, 2012
2012
Earlier work this paper cites.
——, “Whole test suite generation,” IEEE Transactions on Software Engineering , vol. 39, no. 2, pp. 276–291, 2013
2013
Earlier work this paper cites.
G. Fraser and A. Arcuri, “EvoSuite: On the challenges of test case generation in the real world,” in International Conference on Software Testing, Verification and Validation (ICST) . IEEE, 2013, pp. 362–369
2013
Earlier work this paper cites.
A. Arcuri and G. Fraser, “Parameter tuning or default values? an empirical investigation in search-based software engineering,” Empirical Software Engineering , vol. 18, pp. 594–623, 2013
2013
Earlier work this paper cites.
S. Afshan, P. McMinn, and M. Stevenson, “Evolving readable string test inputs using a natural language model to reduce human oracle cost,” in International Conference on Software Testing, Verification and Validation (ICST) . IEEE, 2013, pp. 352–361
2013
Earlier work this paper cites.
A. J. Ko, B. Dosono, and N. Duriseti, “Thirty years of software problems in the news,” in Proc. Int’l Workshop on Cooperative and Human Aspects of Software Engineering (CHASE) . ACM, 2014, pp. 32–39
2014
Earlier work this paper cites.
G. Fraser and A. Arcuri, “A large scale evaluation of automated unit test generation using evosuite,” ACM Transactions on Software Engineering and Methodology (TOSEM) , vol. 24, no. 2, p. 8, 2014
2014
Earlier work this paper cites.
M. Beller, G. Gousios, A. Panichella, and A. Zaidman, “When, how, and why developers (do not) test in their IDEs,” in Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering (ESEC/FSE) . ACM, 2015, pp. 179–190
2015
Earlier work this paper cites.
M. Beller, G. Gousios, and A. Zaidman, “How (much) do developers test?” in 37th IEEE/ACM International Conference on Software Engineering (ICSE) . IEEE Computer Society, 2015, pp. 559–562
2015
Earlier work this paper cites.
G. Fraser, M. Staats, P. McMinn, A. Arcuri, and F. Padberg, “Does automated unit test generation really help software testers? A controlled empirical study,” ACM Trans. Softw. Eng. Methodol. , vol. 24, no. 4, pp. 23:1–23:49, 2015
2015
Earlier work this paper cites.
G. Fraser and A. Arcuri, “Achieving scalable mutation-based generation of whole test suites,” Empirical Software Engineering , vol. 20, no. 3, pp. 783–812, 2015
2015
Earlier work this paper cites.
S. Shamshiri, R. Just, J. M. Rojas, G. Fraser et al. , “Do automatically generated unit tests find real faults? an empirical study of effectiveness and challenges,” in International Conference on Automated Software Engineering (ASE) . IEEE, 2015, pp. 201––211
2015
Earlier work this paper cites.
S. Lo and S. Andrews, “To transform or not to transform: using generalized linear mixed models to analyse reaction time data,” Frontiers in Psychology , vol. 6, 2015
2015
Earlier work this paper cites.
F. Palomba, A. Panichella, A. Zaidman, R. Oliveto, and A. De Lucia, “Automatic test case generation: What if test code quality matters?” in Proceedings of the 25th International Symposium on Software Testing and Analysis (ISSTA) . ACM, 2016, pp. 130–141
2016
Earlier work this paper cites.
F. Palomba, D. Di Nucci, A. Panichella, R. Oliveto, and A. De Lucia, “On the diffusion of test smells in automatically generated test code: An empirical study,” in 2016 IEEE/ACM 9th International Workshop on Search-Based Software Testing (SBST) , 2016, pp. 5–14
2016
Earlier work this paper cites.
B. Zhang, E. Hill, and J. Clause, “Towards automatically generating descriptive names for unit tests,” in Proc. Int’l Conf. on Automated Software Engineering (ASE) . ACM, 2016, pp. 625–636
2016
Earlier work this paper cites.
S. Panichella, A. Panichella, M. Beller, A. Zaidman, and H. C. Gall, “The impact of test case summaries on bug fixing performance: An empirical investigation,” in Proc. Int’l Conference on Software Engineering (ICSE) , 2016, pp. 547–558
2016
Earlier work this paper cites.
S. Vegas, C. Apa, and N. Juristo, “Crossover designs in software engineering experiments: Benefits and perils,” IEEE Transactions on Software Engineering , vol. 42, no. 2, pp. 120–135, 2016
2016
Cited alongside, same era.
M. M. Almasi, H. Hemmati, G. Fraser, A. Arcuri, and J. Benefelds, “An industrial evaluation of unit test generation: Finding real faults in a financial application,” in Proc. Int’l Conf. on Software Engineering: Software Engineering in Practice Track (ICSE-SEIP) . IEEE, 2017
2017
Cited alongside, same era.
E. Daka, J. M. Rojas, and G. Fraser, “Generating unit tests with descriptive names or: Would you name your children thing1 and thing2?” in Proceedings of the International Symposium on Software Testing and Analysis (ISSTA) . ACM, 2017, pp. 57–67
2017
Cited alongside, same era.
A. Panichella, F. M. Kifetew, and P. Tonella, “Automated test case generation as a many-objective optimisation problem with dynamic selection of the targets,” IEEE Trans. Software Eng. , vol. 44, pp. 122–158, 2018
OpenAI:, J. Achiam, S. Adler, S. Agarwal et al. , “Gpt-4 technical report,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2303.08774
2023
Later among the works it cites.
A. Al-Kaswan, T. Ahmed, M. Izadi, A. A. Sawant et al. , “Extending source code pre-trained language models to summarise decompiled binaries,” in 2023 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 2023, pp. 260–271
2023
Later among the works it cites.
B. Rozière, J. Gehring, F. Gloeckle, S. Sootla et al. , “Code llama: Open foundation models for code,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2308.12950
2023
Later among the works it cites.
R. Li, L. B. Allal, Y. Zi, N. Muennighoff et al. , “Starcoder: may the source be with you!” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2305.06161
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
A. Arcuri, “An experience report on applying software testing academic results in industry: we need usable automated test generation,” Empirical Software Engineering , vol. 23, no. 4, pp. 1959–1981, 2018
2018
Cited alongside, same era.
G. Grano, S. Scalabrino, H. C. Gall, and R. Oliveto, “An empirical investigation on the readability of manual and generated test cases,” in International Conference on Program Comprehension (ICPC) . IEEE, 2018, pp. 348–351
2018
Cited alongside, same era.
R. H. B. Christensen, “Cumulative link models for ordinal regression with the r package ordinal,” Submitted in J. Stat. Software , vol. 35, 2018
2018
Cited alongside, same era.
J. Campos, Y. Ge, N. Albunian, G. Fraser et al. , “An empirical evaluation of evolutionary algorithms for unit test suite generation,” Information and Software Technology , vol. 104, pp. 207–235, 2018
2018
Cited alongside, same era.
M. Beller, G. Gousios, A. Panichella, S. Proksch et al. , “Developer testing in the IDE: patterns, beliefs, and behavior,” IEEE Trans. Software Eng. , vol. 45, no. 3, pp. 261–284, 2019
2019
Cited alongside, same era.
G. Grano, F. Palomba, D. Di Nucci, A. De Lucia, and H. C. Gall, “Scented since the beginning: On the diffuseness of test smells in automatically generated test code,” Journal of Systems and Software , vol. 156, pp. 312–327, 2019
2019
Cited alongside, same era.
V. Khorikov, Unit Testing Principles, Practices, and Patterns . Manning, 2019
2019
Cited alongside, same era.
D. Roy, Z. Zhang, M. Ma, V. Arnaoudova et al. , “Deeptc-enhancer: Improving the readability of automatically generated tests,” in Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering , 2020, pp. 287–298
2020
Cited alongside, same era.
N. Rao, K. Jain, U. Alon, C. Le Goues, and V. J. Hellendoorn, “Cat-lm training language models on aligned code and tests,” in 38th IEEE/ACM International Conference on Automated Software Engineering (ASE) , 2023, pp. 409–420
2023
Later among the works it cites.
C. Lemieux, J. P. Inala, S. K. Lahiri, and S. Sen, “Codamosa: Escaping coverage plateaus in test generation with pre-trained large language models,” in IEEE/ACM 45th International Conference on Software Engineering (ICSE) , 2023, pp. 919–931
2023
Later among the works it cites.
S. Ouyang, J. M. Zhang, M. Harman, and M. Wang, “LLM is like a box of chocolates: the non-determinism of ChatGPT in code generation,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2308.02828
2023
Later among the works it cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma et al. , “Chain-of-thought prompting elicits reasoning in large language models,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2201.11903
2023
Later among the works it cites.
J. Li, G. Li, Y. Li, and Z. Jin, “Structured chain-of-thought prompting for code generation,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2305.06599
2023
Later among the works it cites.
G. Marvin, N. Hellen, D. Jjingo, and J. Nakatumba-Nabende, “Prompt engineering in large language models,” in International Conference on Data Intelligence and Cognitive Informatics . Springer, 2023, pp. 387–402
2023
Later among the works it cites.
J. Zamfirescu-Pereira, R. Y. Wong, B. Hartmann, and Q. Yang, “Why Johnny can’t prompt: how non-AI experts try (and fail) to design LLM prompts,” in Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems , 2023, pp. 1–21
2023
Later among the works it cites.
A. Fan, B. Gokkaya, M. Harman, M. Lyubarskiy et al. , “Large language models for software engineering: Survey and open problems,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/arXiv.2310.03533
2023
Later among the works it cites.
A. Deljouyi and A. Zaidman, “Generating understandable unit tests through end-to-end test scenario carving,” in Proceedings of the 23rd IEEE International Working Conference on Source Code Analysis and Manipulation (SCAM) . IEEE, 2023, pp. 107–118
2023
Later among the works it cites.
A. M. Dakhel, A. Nikanjam, V. Majdinasab, F. Khomh, and M. C. Desmarais, “Effective test generation using pre-trained large language models and mutation testing,” 2023
2023
Later among the works it cites.
B. Steenhoek, M. Tufano, N. Sundaresan, and A. Svyatkovskiy, “Reinforcement learning from automatic feedback for high-quality unit test generation,” arXiv , 2023. [Online]. Available: https://doi.org/10.48550/ArXiv.2310.02368
2023
Later among the works it cites.
A. Khatami and A. Zaidman, “State-of-the-practice in quality assurance in Java-based open source software development,” Software: Practice and Experience , vol. 54, no. 8, pp. 1408–1446, 2024
2024
Closest in time.
C. E. Brandt, A. Khatami, M. Wessel, and A. Zaidman, “Shaken, not stirred: How developers like their amplified tests,” IEEE Trans. Software Eng. , vol. 50, no. 5, pp. 1264–1280, 2024
2024
Closest in time.
M. Schäfer, S. Nadi, A. Eghbali, and F. Tip, “An empirical evaluation of using large language models for automated unit test generation,” IEEE Transactions on Software Engineering , vol. 50, no. 1, pp. 85–105, 2024
2024
Closest in time.
K. El Haji, C. Brandt, and A. Zaidman, “Using github copilot for test generation in python: An empirical study,” in Proceedings of the International Conference on Automation of Software Test (AST) . ACM, 2024
2024
Closest in time.
M. L. Siddiq, J. C. S. Santos, R. H. Tanvir, N. Ulfat et al. , “Using large language models to generate junit tests: An empirical study,” in International Conference on Evaluation and Assessment in Software Engineering (EASE) . ACM, 2024
2024
Closest in time.
“Replication package of UTGen,” 2024. [Online]. Available: https://doi.org/10.5281/zenodo.13329464
2024
Closest in time.
M. Izadi, J. Katzy, T. Van Dam, M. Otten et al. , “Language models for code completion: A practical evaluation,” in Proceedings of the IEEE/ACM 46th International Conference on Software Engineering , 2024, pp. 1–13
2024
Closest in time.
B. Baudry, K. Etemadi, S. Fang, Y. Gamage et al. , “Generative ai to generate test data generators,” arXiv , 2024. [Online]. Available: https://doi.org/10.48550/arXiv.2401.17626
2024
Closest in time.
N. Alshahwan, J. Chheda, A. Finegenova, B. Gokkaya et al. , “Automated unit test improvement using large language models at Meta,” arXiv , 2024. [Online]. Available: https://doi.org/10.48550/ArXiv.2402.09171
2024
Closest in time.