Fetching the paper…
Reading the bibliography…
Motivation.
M. McCloskey and N. J. Cohen, “Catastrophic interference in connectionist networks: The sequential learning problem,” in Psychology of learning and motivation . Elsevier, 1989, vol. 24, pp. 109–165
1989
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “BLEU: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting of the Association for Computational Linguistics , 2002, pp. 311–318
2002
Earlier work this paper cites.
C. K. Roy and J. R. Cordy, “A survey on software clone detection research,” Queen’s School of computing TR , vol. 541, no. 115, pp. 64–68, 2007
2007
Earlier work this paper cites.
L. Jiang, G. Misherghi, Z. Su, and S. Glondu, “Deckard: Scalable and accurate tree-based detection of code clones,” in 29th International Conference on Software Engineering (ICSE’07) . IEEE, 2007, pp. 96–105
2007
Earlier work this paper cites.
J. Svajlenko, J. F. Islam, I. Keivanloo, C. K. Roy, and M. M. Mia, “Towards a big data curated benchmark of inter-project code clones,” in 2014 IEEE International Conference on Software Maintenance and Evolution . IEEE, 2014, pp. 476–480
2014
Earlier work this paper cites.
R. Just, D. Jalali, and M. D. Ernst, “Defects4J: A database of existing faults to enable controlled testing studies for java programs,” in Proceedings of the 2014 international symposium on software testing and analysis , 2014, pp. 437–440
2014
Earlier work this paper cites.
H. Sajnani, V. Saini, J. Svajlenko, C. K. Roy, and C. V. Lopes, “SourcererCC: Scaling code clone detection to big-code,” in Proceedings of the 38th International Conference on Software Engineering , 2016, pp. 1157–1168
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
C. V. Lopes, P. Maj, P. Martins, V. Saini, D. Yang, J. Zitny, H. Sajnani, and J. Vitek, “Déjàvu: a map of code duplicates on github,” Proceedings of the ACM on Programming Languages , vol. 1, no. OOPSLA, pp. 1–28, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
D. Yang, P. Martins, V. Saini, and C. Lopes, “Stack Overflow in GitHub: any snippets there?” in 2017 IEEE/ACM 14th International Conference on Mining Software Repositories (MSR) . IEEE, 2017, pp. 280–290
2017
Earlier work this paper cites.
J. Li, P. He, J. Zhu, and M. R. Lyu, “Software defect prediction via convolutional neural network,” in 2017 IEEE international conference on software quality, reliability and security (QRS) . IEEE, 2017, pp. 318–328
2017
Earlier work this paper cites.
X. Gu, H. Zhang, and S. Kim, “Deep code search,” in Proceedings of the 40th International Conference on Software Engineering , 2018, pp. 933–944
2018
Earlier work this paper cites.
X. Hu, G. Li, X. Xia, D. Lo, S. Lu, and Z. Jin, “Summarizing source code with transferred api knowledge,” 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
B. Wei, G. Li, X. Xia, Z. Fu, and Z. Jin, “Code generation as a dual task of code summarization,” Advances in neural information processing systems , vol. 32, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. LeClair, S. Jiang, and C. McMillan, “A neural model for generating natural language summaries of program subroutines,” in 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 2019, pp. 795–806
2019
Earlier work this paper cites.
M. Gharehyazie, B. Ray, M. Keshani, M. S. Zavosht, A. Heydarnoori, and V. Filkov, “Cross-project code clones in GitHub,” Empirical Software Engineering , vol. 24, pp. 1538–1573, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Allamanis, “The adverse effects of code duplication in machine learning models of code,” in Proceedings of the 2019 ACM SIGPLAN International Symposium on New Ideas, New Paradigms, and Reflections on Programming and Software , 2019, pp. 143–153
2019
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2022
Later among the works it cites.
T. Ahmed and P. Devanbu, “Multilingual training for software engineering,” in Proceedings of the 44th International Conference on Software Engineering , 2022, pp. 1443–1455
2022
Later among the works it cites.
2022
Later among the works it cites.
L. Di Grazia and M. Pradel, “Code search: A survey of techniques for finding code,” ACM Computing Surveys , vol. 55, no. 11, pp. 1–31, 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
2021
Cited alongside, same era.
X. Zhou, D. Han, and D. Lo, “Assessing generalizability of CodeBERT,” in 2021 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 2021, pp. 425–436
2021
Cited alongside, same era.
Y. Golubev and T. Bryksin, “On the nature of code cloning in open-source java projects,” in 2021 IEEE 15th International Workshop on Software Clones (IWSC) . IEEE, 2021, pp. 22–28
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2023
Later among the works it cites.
B. Min, H. Ross, E. Sulem, A. P. B. Veyseh, T. H. Nguyen, O. Sainz, E. Agirre, I. Heintz, and D. Roth, “Recent advances in natural language processing via large pre-trained language models: A survey,” ACM Computing Surveys , vol. 56, no. 2, pp. 1–40, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Zakeri-Nasrabadi, S. Parsa, M. Ramezani, C. Roy, and M. Ekhtiarzadeh, “A systematic literature review on source code similarity measurement and clone detection: Techniques, applications, and challenges,” Journal of Systems and Software , p. 111796, 2023
2023
Later among the works it cites.
A. Karmakar, M. Allamanis, and R. Robbes, “JEMMA: An extensible java dataset for ml4code applications,” Empirical Software Engineering , vol. 28, no. 2, p. 54, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
M. Schäfer, S. Nadi, A. Eghbali, and F. Tip, “An empirical evaluation of using large language models for automated unit test generation,” IEEE Transactions on Software Engineering , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
R. Tufano, L. Pascarella, and G. Bavota, “Automating code-related tasks through transformers: The impact of pre-training,” in 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2023, pp. 2425–2437
2023
Later among the works it cites.
2023
Later among the works it cites.
T. Sharma, M. Kechagia, S. Georgiou, R. Tiwari, I. Vats, H. Moazen, and F. Sarro, “A survey on machine learning techniques applied to source code,” Journal of Systems and Software , vol. 209, p. 111934, 2024
2024
Closest in time.
S. Iyer, I. Konstas, A. Cheung, and L. Zettlemoyer, “Summarizing source code using a neural attention model,” in 54th Annual Meeting of the Association for Computational Linguistics 2016 . Association for Computational Linguistics, 2016, pp. 2073–2083
2083
Closest in time.
M. Allamanis, H. Peng, and C. Sutton, “A convolutional attention network for extreme summarization of source code,” in International conference on machine learning . PMLR, 2016, pp. 2091–2100
2091
Closest in time.