Fetching the paper…
Reading the bibliography…
Unit testing is essential for software reliability, yet manual test creation is time-consuming and often neglected.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al. (2020) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
Beck K (2000) Extreme programming explained: embrace change. addison-wesley professional
2000
Earlier work this paper cites.
Sen K, Marinov D, Agha G (2005) Cute: A concolic unit testing engine for c. ACM SIGSOFT Software Engineering Notes 30(5):263–272
2005
Earlier work this paper cites.
Pacheco C, Ernst MD (2007) Randoop: feedback-directed random testing for java. In: Companion to the 22nd ACM SIGPLAN conference on Object-oriented programming systems and applications companion, pp 815–816
2007
Earlier work this paper cites.
Buse RP, Weimer WR (2009) Learning a metric for code readability. IEEE Transactions on software engineering 36(4):546–558
2009
Earlier work this paper cites.
Pinto GH, Vergilio SR (2010) A multi-objective genetic algorithm to test data generation. In: 2010 22nd IEEE International Conference on Tools with Artificial Intelligence, IEEE, vol 1, pp 129–134
2010
Earlier work this paper cites.
Fraser G, Arcuri A (2011) Evosuite: automatic test suite generation for object-oriented software. In: Proceedings of the 19th ACM SIGSOFT symposium and the 13th European conference on Foundations of software engineering, pp 416–419
2011
Earlier work this paper cites.
Posnett D, Hindle A, Devanbu P (2011) A simpler model of software readability. In: Proceedings of the 8th working conference on mining software repositories, pp 73–82
2011
Earlier work this paper cites.
Dorn J (2012) A general software readability model. MCS Thesis available from (http://www cs virginia edu/weimer/students/dorn-mcs-paper pdf) 5:11–14
2012
Earlier work this paper cites.
Arcuri A, Fraser G (2013) Parameter tuning or default values? an empirical investigation in search-based software engineering. Empirical Software Engineering 18:594–623
2013
Earlier work this paper cites.
Fraser G, Arcuri A (2013) Evosuite: On the challenges of test case generation in the real world. In: 2013 IEEE sixth international conference on software testing, verification and validation, IEEE, pp 362–369
2013
Earlier work this paper cites.
Daka E, Fraser G (2014) A survey on unit testing practices and problems. In: 2014 IEEE 25th International Symposium on Software Reliability Engineering, IEEE, pp 201–211
2014
Earlier work this paper cites.
Daka E, Campos J, Fraser G, Dorn J, Weimer W (2015) Modeling readability to improve unit tests. In: Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering, pp 107–118
2015
Earlier work this paper cites.
Fraser G, Staats M, McMinn P, Arcuri A, Padberg F (2015) Does automated unit test generation really help software testers? a controlled empirical study. ACM Transactions on Software Engineering and Methodology (TOSEM) 24(4):1–49
2015
Earlier work this paper cites.
Rojas JM, Fraser G, Arcuri A (2015) Automated unit test generation during software development: A controlled experiment and think-aloud observations. In: Proceedings of the 2015 international symposium on software testing and analysis, pp 338–349
2015
Earlier work this paper cites.
Shamshiri S, Just R, Rojas JM, Fraser G, McMinn P, Arcuri A (2015) Do automatically generated unit tests find real faults? an empirical study of effectiveness and challenges (t). In: 2015 30th IEEE/ACM International Conference on Automated Software Engineering (ASE), IEEE, pp 201–211
2015
Earlier work this paper cites.
Palomba F, Di Nucci D, Panichella A, Oliveto R, De Lucia A (2016) On the diffusion of test smells in automatically generated test code: An empirical study. In: Proceedings of the 9th international workshop on search-based software testing, pp 5–14
2016
Earlier work this paper cites.
Almasi MM, Hemmati H, Fraser G, Arcuri A, Benefelds J (2017) An industrial evaluation of unit test generation: Finding real faults in a financial application. In: 2017 IEEE/ACM 39th International Conference on Software Engineering: Software Engineering in Practice Track (ICSE-SEIP), IEEE, pp 263–272
2017
Earlier work this paper cites.
Arcuri A, Fraser G, Just R (2017) Private api access and functional mocking in automated unit test generation. In: 2017 IEEE international conference on software testing, verification and validation (ICST), IEEE, pp 126–137
2017
Earlier work this paper cites.
Panichella A, Kifetew FM, Tonella P (2017) Automated test case generation as a many-objective optimisation problem with dynamic selection of the targets. IEEE Transactions on Software Engineering 44(2):122–158
2017
Earlier work this paper cites.
Grano G, Scalabrino S, Gall HC, Oliveto R (2018) An empirical investigation on the readability of manual and generated test cases. In: Proceedings of the 26th Conference on Program Comprehension, pp 348–351
2018
Earlier work this paper cites.
Scalabrino S, Linares-Vásquez M, Oliveto R, Poshyvanyk D (2018) A comprehensive model for code readability. Journal of Software: Evolution and Process 30(6):e1958
2018
Earlier work this paper cites.
Shamshiri S, Rojas JM, Galeotti JP, Walkinshaw N, Fraser G (2018) How do automatically generated unit tests influence software maintenance? In: 2018 IEEE 11th international conference on software testing, verification and validation (ICST), IEEE, pp 250–261
2018
Earlier work this paper cites.
Peruma A, Almalki KS, Newman CD, Mkaouer MW, Ouni A, Palomba F (2019) On the distribution of test smells in open source android applications: An exploratory study.(2019). Citado na p 13
2019
Earlier work this paper cites.
Zeller A, Gopinath R, Böhme M, Fraser G, Holler C (2019) The fuzzing book
2019
Cited alongside, same era.
Panichella A, Panichella S, Fraser G, Sawant AA, Hellendoorn VJ (2020) Revisiting test smells in automatically generated tests: limitations, pitfalls, and opportunities. In: 2020 IEEE international conference on software maintenance and evolution (ICSME), IEEE, pp 523–533
2020
Cited alongside, same era.
Peruma A, Almalki K, Newman CD, Mkaouer MW, Ouni A, Palomba F (2020) Tsdetect: An open source test smells detection tool. In: Proceedings of the 28th ACM joint meeting on european software engineering conference and symposium on the foundations of software engineering, pp 1650–1654
2020
Cited alongside, same era.
Dantas CEC, Maia MA (2021) Readability and understandability scores for snippet assessment: An exploratory study. arXiv preprint arXiv:210809181
2021
Cited alongside, same era.
Amatriain X (2024) Prompt design and engineering: Introduction and advanced methods. arXiv preprint arXiv:240114423
2024
Closest in time.
Chen Y, Hu Z, Zhi C, Han J, Deng S, Yin J (2024) Chatunitest: A framework for llm-based test generation. In: Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering, pp 572–576
2024
Closest in time.
Dakhel AM, Nikanjam A, Majdinasab V, Khomh F, Desmarais MC (2024) Effective test generation using pre-trained large language models and mutation testing. Information and Software Technology 171:107468
2024
Closest in time.
Deljouyi A, Koohestani R, Izadi M, Zaidman A (2024) Leveraging large language models for enhancing the understandability of generated unit tests. arXiv preprint arXiv:240811710
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Abdullin AM, Itsykson VM (2022) Kex: A platform for analysis of jvm programs. Information and Control Systems (1 (116)):30–43
2022
Cited alongside, same era.
Kojima T, Gu SS, Reid M, Matsuo Y, Iwasawa Y (2022) Large language models are zero-shot reasoners. Advances in neural information processing systems 35:22199–22213
2022
Cited alongside, same era.
Mi Q, Hao Y, Ou L, Ma W (2022) Towards using visual, semantic and structural features to improve code readability classification. Journal of Systems and Software 193:111454
2022
Cited alongside, same era.
Oliveira D, Bruno R, Madeiral F, Masuhara H, Castor F (2022) A systematic literature review on the impact of formatting elements on program understandability. Available at SSRN 4182156
2022
Cited alongside, same era.
Panichella A, Panichella S, Fraser G, Sawant AA, Hellendoorn VJ (2022) Test smells 20 years later: detectability, validity, and reliability. Empirical Software Engineering 27(7):170
2022
Cited alongside, same era.
Si C, Gan Z, Yang Z, Wang S, Wang J, Boyd-Graber J, Wang L (2022) Prompting gpt-3 to be reliable. arXiv preprint arXiv:221009150
2022
Cited alongside, same era.
Wang X, Wei J, Schuurmans D, Le Q, Chi E, Narang S, Chowdhery A, Zhou D (2022) Self-consistency improves chain of thought reasoning in language models. arXiv preprint arXiv:220311171
2022
Cited alongside, same era.
Wei J, Wang X, Schuurmans D, Bosma M, Xia F, Chi E, Le QV, Zhou D, et al. (2022) Chain-of-thought prompting elicits reasoning in large language models. Advances in neural information processing systems 35:24824–24837
2022
Cited alongside, same era.
2024
Closest in time.
Hossain SB, Dwyer M (2024) Togll: Correct and strong test oracle generation with llms. arXiv preprint arXiv:240503786
2024
Closest in time.
Jiang AQ, Sablayrolles A, Roux A, Mensch A, Savary B, Bamford C, Chaplot DS, Casas Ddl, Hanna EB, Bressand F, et al. (2024) Mixtral of experts. arXiv preprint arXiv:240104088
2024
Closest in time.
Kumar A, Haiduc S, Das PP, Chakrabarti PP (2024) Llms as evaluators: A novel approach to evaluate bug report summarization. arXiv preprint arXiv:240900630
2024
Closest in time.
Macedo M, Tian Y, Cogo FR, Adams B (2024) Exploring the impact of the output format on the evaluation of large language models for code translation. arXiv preprint arXiv:240317214
2024
Closest in time.
OpenAI (2023) Gpt-3.5-turbo. URL: https://platformopenaicom/docs/models/gpt-3-5-turbo/[accessed 2024-05-21]
2024
Closest in time.
Sahoo P, Singh AK, Saha S, Jain V, Mondal S, Chadha A (2024) A systematic survey of prompt engineering in large language models: Techniques and applications. arXiv preprint arXiv:240207927
2024
Closest in time.
Sergeyuk A, Lvova O, Titov S, Serova A, Bagirov F, Kirillova E, Bryksin T (2024) Reassessing java code readability models with a human-centered approach. In: Proceedings of the 32nd IEEE/ACM International Conference on Program Comprehension, pp 225–235
2024
Closest in time.
Sun W, Miao Y, Li Y, Zhang H, Fang C, Liu Y, Deng G, Liu Y, Chen Z (2024) Source code summarization in the era of large language models. arXiv preprint arXiv:240707959
2024
Closest in time.
Tang Y, Liu Z, Zhou Z, Luo X (2024) Chatgpt vs sbst: A comparative assessment of unit test suite generation. IEEE Transactions on Software Engineering
2024
Closest in time.
Winkler D, Urbanke P, Ramler R (2024) Investigating the readability of test code. Empirical Software Engineering 29(2):53
2024
Closest in time.
Yang L, Yang C, Gao S, Wang W, Wang B, Zhu Q, Chu X, Zhou J, Liang G, Wang Q, et al. (2024) On the evaluation of large language models in unit test generation. In: Proceedings of the 39th IEEE/ACM International Conference on Automated Software Engineering, pp 1607–1619
2024
Closest in time.
Yuan Z, Liu M, Ding S, Wang K, Chen Y, Peng X, Lou Y (2024) Evaluating and improving chatgpt for unit test generation. Proceedings of the ACM on Software Engineering 1(FSE):1703–1726
2024
Closest in time.
Biagiola M, Ghislotti G, Tonella P (2025) Improving the readability of automatically generated tests using large language models. In: 2025 IEEE Conference on Software Testing, Verification and Validation (ICST), IEEE, pp 162–173
2025
Closest in time.
Deljouyi A, Koohestani R, Izadi M, Zaidman A (2025) Leveraging large language models for enhancing the understandability of generated unit tests. in 2025 ieee/acm 47th international conference on software engineering (icse). Los Alamitos, CA, USA pp 392–404
2025
Closest in time.
Li J, Li G, Li Y, Jin Z (2025) Structured chain-of-thought prompting for code generation. ACM Transactions on Software Engineering and Methodology 34(2):1–23
2025
Closest in time.
Molina F, Gorla A, d’Amorim M (2025) Test oracle automation in the era of llms. ACM Transactions on Software Engineering and Methodology 34(5):1–24
2025
Closest in time.
Ouédraogo WC, Plein L, Kabore K, Habib A, Klein J, Lo D, Bissyandé TF (2025) Enriching automatic test case generation by extracting relevant test inputs from bug reports. Empirical Software Engineering 30(3):85
2025
Closest in time.
Zhang Y, Lu Q, Liu K, Dou W, Zhu J, Qian L, Zhang C, Lin Z, Wei J (2025) Citywalk: Enhancing llm-based c++ unit test generation via project-dependency awareness and language-specific knowledge. arXiv preprint arXiv:250116155
2025
Closest in time.