Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated remarkable potential in code generation.
R. A. Fisher, “On the interpretation of χ \chi 2 from contingency tables, and the calculation of p,” Journal of the royal statistical society , vol. 85, no. 1, pp. 87–94, 1922
1922
Earlier work this paper cites.
J. L. Fleiss, “Measuring nominal scale agreement among many raters.” Psychological bulletin , vol. 76, no. 5, p. 378, 1971
1971
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting of the Association for Computational Linguistics , 2002, pp. 311–318
2002
Earlier work this paper cites.
C.-Y. Lin, “Rouge: A package for automatic evaluation of summaries,” in Text summarization branches out , 2004, pp. 74–81
2004
Earlier work this paper cites.
S. Banerjee and A. Lavie, “Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,” in Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and/or summarization , 2005, pp. 65–72
2005
Earlier work this paper cites.
I. Sutskever, G. E. Hinton, and G. W. Taylor, “The recurrent temporal restricted boltzmann machine,” Advances in neural information processing systems , vol. 21, 2008
2008
Earlier work this paper cites.
T. Cohn, P. Blunsom, and S. Goldwater, “Inducing tree-substitution grammars,” The Journal of Machine Learning Research , vol. 11, pp. 3053–3096, 2010
2010
Earlier work this paper cites.
S. Gulwani, “Dimensions in program synthesis,” in Proceedings of the 12th international ACM SIGPLAN symposium on Principles and practice of declarative programming , 2010, pp. 13–24
2010
Earlier work this paper cites.
T. T. Nguyen, A. T. Nguyen, H. A. Nguyen, and T. N. Nguyen, “A statistical semantic language model for source code,” in Proceedings of the 2013 9th Joint Meeting on Foundations of Software Engineering , 2013, pp. 532–542
2013
Earlier work this paper cites.
M. Allamanis and C. Sutton, “Mining idioms from source code,” in Proceedings of the 22nd acm sigsoft international symposium on foundations of software engineering , 2014, pp. 472–483
2014
Earlier work this paper cites.
V. Raychev, M. Vechev, and E. Yahav, “Code completion with statistical language models,” in Proceedings of the 35th ACM SIGPLAN conference on programming language design and implementation , 2014, pp. 419–428
2014
Earlier work this paper cites.
Z. Liu, Y. Dou, J. Jiang, and J. Xu, “Automatic code generation of convolutional neural networks in fpga implementation,” in 2016 International conference on field-programmable technology (FPT) . IEEE, 2016, pp. 61–68
2016
Earlier work this paper cites.
A. Eriguchi, K. Hashimoto, and Y. Tsuruoka, “Tree-to-sequence attentional neural machine translation,” in Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2016, pp. 823–833
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
S. Jiang, A. Armaly, and C. McMillan, “Automatically generating commit messages from diffs using neural machine translation,” in 2017 32nd IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 2017, pp. 135–146
2017
Earlier work this paper cites.
P. Yin and G. Neubig, “A syntactic neural model for general-purpose code generation,” in Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2017, pp. 440–450
2017
Earlier work this paper cites.
Y. Wan, Z. Zhao, M. Yang, G. Xu, H. Ying, J. Wu, and P. S. Yu, “Improving automatic source code summarization via deep reinforcement learning,” in Proceedings of the 33rd ACM/IEEE international conference on automated software engineering , 2018, pp. 397–407
2018
Earlier work this paper cites.
B. Zhang and R. Sennrich, “Root mean square layer normalization,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
Z. Sun, Q. Zhu, L. Mou, Y. Xiong, G. Li, and L. Zhang, “A grammar-based structural cnn decoder for code generation,” in Proceedings of the AAAI conference on artificial intelligence , vol. 33, no. 01, 2019, pp. 7055–7062
2019
Earlier work this paper cites.
A. Svyatkovskiy, S. K. Deng, S. Fu, and N. Sundaresan, “Intellicode compose: Code generation using transformer,” in Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2020, pp. 1433–1443
2020
Earlier work this paper cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang et al. , “Codebert: A pre-trained model for programming and natural languages,” in Findings of the Association for Computational Linguistics: EMNLP 2020 , 2020, pp. 1536–1547
2020
Earlier work this paper cites.
D. Guo, S. Ren, S. Lu, Z. Feng, D. Tang, L. Shujie, L. Zhou, N. Duan, A. Svyatkovskiy, S. Fu et al. , “Graphcodebert: Pre-training code representations with data flow,” in International Conference on Learning Representations , 2020
2020
Earlier work this paper cites.
B. Wei, Y. Li, G. Li, X. Xia, and Z. Jin, “Retrieve and refine: exemplar-based neural comment generation,” in 2020 35th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 2020, pp. 349–360
2020
Earlier work this paper cites.
L. Massarelli, F. Petroni, A. Piktus, M. Ott, T. Rocktäschel, V. Plachouras, F. Silvestri, and S. Riedel, “How decoding strategies affect the verifiability of generated text,” in Findings of the Association for Computational Linguistics: EMNLP 2020 , 2020, pp. 223–235
2020
Earlier work this paper cites.
Z. Gao, X. Xia, J. Grundy, D. Lo, and Y.-F. Li, “Generating question titles for stack overflow from mined code snippets,” ACM Transactions on Software Engineering and Methodology (TOSEM) , vol. 29, no. 4, pp. 1–37, 2020
2020
Earlier work this paper cites.
E. J. Hu, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, W. Chen et al. , “Lora: Low-rank adaptation of large language models,” in International Conference on Learning Representations , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
B. Lester, R. Al-Rfou, and N. Constant, “The power of scale for parameter-efficient prompt tuning,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , 2021, pp. 3045–3059
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Lu, D. Guo, S. Ren, J. Huang, A. Svyatkovskiy, A. Blanco, C. Clement, D. Drain, D. Jiang, D. Tang et al. , “Codexglue: A machine learning benchmark dataset for code understanding and generation,” in Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1) , 2021
2021
Earlier work this paper cites.
W. Ahmad, S. Chakraborty, B. Ray, and K.-W. Chang, “Unified pre-training for program understanding and generation,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , 2021, pp. 2655–2668
2021
Earlier work this paper cites.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi, “Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , 2021, pp. 8696–8708
2021
Earlier work this paper cites.
X. L. Li and P. Liang, “Prefix-tuning: Optimizing continuous prompts for generation,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , 2021, pp. 4582–4597
2021
Earlier work this paper cites.
A. Mastropaolo, S. Scalabrino, N. Cooper, D. N. Palacio, D. Poshyvanyk, R. Oliveto, and G. Bavota, “Studying the usage of text-to-text transfer transformer to support code-related tasks,” in 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 2021, pp. 336–347
2021
Earlier work this paper cites.
M. Shah, R. Shenoy, and R. Shankarmani, “Natural language to python source code using transformers,” in 2021 International Conference on Intelligent Technologies (CONIT) . IEEE, 2021, pp. 1–4
2021
Earlier work this paper cites.
P. Liguori, E. Al-Hossami, V. Orbinato, R. Natella, S. Shaikh, D. Cotroneo, and B. Cukic, “Evil: exploiting software via natural language,” in 2021 IEEE 32nd International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 2021, pp. 321–332
2021
Earlier work this paper cites.
W. Du and S. Ding, “A survey on multi-agent deep reinforcement learning: from the perspective of challenges and applications,” Artificial Intelligence Review , vol. 54, pp. 3215–3238, 2021
2021
Earlier work this paper cites.
Y. Li, D. Choi, J. Chung, N. Kushman, J. Schrittwieser, R. Leblond, T. Eccles, J. Keeling, F. Gimeno, A. Dal Lago et al. , “Competition-level code generation with alphacode,” Science , vol. 378, no. 6624, pp. 1092–1097, 2022
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-thought prompting elicits reasoning in large language models,” Advances in Neural Information Processing Systems , vol. 35, pp. 24 824–24 837, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa, “Large language models are zero-shot reasoners,” Advances in neural information processing systems , vol. 35, pp. 22 199–22 213, 2022
2022
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” in The Eleventh International Conference on Learning Representations , 2022
2022
Cited alongside, same era.
A. Zeng, X. Liu, Z. Du, Z. Wang, H. Lai, M. Ding, Z. Yang, Y. Xu, W. Zheng, X. Xia et al. , “Glm-130b: An open bilingual pre-trained model,” in The Eleventh International Conference on Learning Representations , 2022
2022
Cited alongside, same era.
Y. Yang, X. Xia, D. Lo, and J. Grundy, “A survey on deep learning for software engineering,” ACM Computing Surveys (CSUR) , vol. 54, no. 10s, pp. 1–73, 2022
2022
Cited alongside, same era.
T. Dao, D. Fu, S. Ermon, A. Rudra, and C. Ré, “Flashattention: Fast and memory-efficient exact attention with io-awareness,” Advances in Neural Information Processing Systems , vol. 35, pp. 16 344–16 359, 2022
2022
Cited alongside, same era.
Y. Gu, X. Han, Z. Liu, and M. Huang, “Ppt: Pre-trained prompt tuning for few-shot learning,” in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2022, pp. 8410–8423
2022
Cited alongside, same era.
G. Yang, K. Liu, X. Chen, Y. Zhou, C. Yu, and H. Lin, “Ccgir: Information retrieval-based code comment generation method for smart contracts,” Knowledge-Based Systems , vol. 237, p. 107858, 2022
2022
Cited alongside, same era.
G. Yang, X. Chen, Y. Zhou, and C. Yu, “Dualsc: Automatic generation and summarization of shellcode via transformer and dual learning,” in 2022 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 2022, pp. 361–372
2022
Cited alongside, same era.
S. Chakraborty, T. Ahmed, Y. Ding, P. T. Devanbu, and B. Ray, “Natgen: generative pre-training by “naturalizing” source code,” in Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering , 2022, pp. 18–30
2022
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Z. Zhang, C. Chen, B. Liu, C. Liao, Z. Gong, H. Yu, J. Li, and R. Wang, “Unifying the perspectives of nlp and software engineering: A survey on language models for code,” 2023
2023
Closest in time.
G. Yang, Y. Zhou, X. Chen, X. Zhang, T. Han, and T. Chen, “Exploitgen: Template-augmented exploit code generation based on codebert,” Journal of Systems and Software , vol. 197, p. 111577, 2023
2023
Closest in time.
2023
Closest in time.
G. Yang, Y. Zhou, X. Zhang, X. Chen, T. Han, and T. Chen, “Assessing and improving syntactic adversarial robustness of pre-trained models for code translation,” 2023
2023
Closest in time.
I. Team, “Internlm: A multilingual language model with progressively enhanced capabilities,” https://github.com/InternLM/InternLM , 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Q. Zheng, X. Xia, X. Zou, Y. Dong, S. Wang, Y. Xue, L. Shen, Z. Wang, A. Wang, Y. Li et al. , “Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x,” in Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining , 2023, pp. 5673–5684
2023
Closest in time.
K. Liu, X. Chen, C. Chen, X. Xie, and Z. Cui, “Automated question title reformulation by mining modification logs from stack overflow,” IEEE Transactions on Software Engineering , 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
H. Face, “Text generation strategies,” 2023, https://huggingface.co/docs/transformers/generation_strategies#decoding-strategies
2023
Closest in time.
X. Liu, Y. Zheng, Z. Du, M. Ding, Y. Qian, Z. Yang, and J. Tang, “Gpt understands, too,” AI Open , 2023
2023
Closest in time.
“alpaca-lora: Instruct-tune llama on consumer hardware,” https://github.com/tloen/alpaca-lora , 2023
2023
Closest in time.
D. Zan, B. Chen, F. Zhang, D. Lu, B. Wu, B. Guan, W. Yongji, and J.-G. Lou, “Large language models meet nl2code: A survey,” in Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2023, pp. 7443–7464
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
N. D. Q. B. Anh T. V. Dau, Jin L.C. Guo, “Bootstrapping code-text pretrained language model to detect inconsistency between code and comment,” EACL 2024 - Demonstration track , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
J. Yang, A. Prabhakar, K. Narasimhan, and S. Yao, “Intercode: Standardizing and benchmarking interactive coding with execution feedback,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
S. Iyer, I. Konstas, A. Cheung, and L. Zettlemoyer, “Summarizing source code using a neural attention model,” in 54th Annual Meeting of the Association for Computational Linguistics 2016 . Association for Computational Linguistics, 2016, pp. 2073–2083
2083
Closest in time.