Fetching the paper…
Reading the bibliography…
Recent years have seen the successful application of large pre-trained models to code representation learning, resulting in substantial improvements on many code-related downstream tasks.
CodeSearchNet Challenge: Evaluating the State of Semantic Code Search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2020 · 1909
Earlier work this paper cites.
Manual and Automatic Evaluation of Summaries. In Proceedings of the ACL-02 Workshop on Automatic Summarization - Volume 4 . 45–51
Chin-Yew Lin and Eduard Hovy. 2002 · 2002
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation. In Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics . 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
SCELMo: Source Code Embeddings from Language Models
Rafael Michael Karampatsis and Charles Sutton. 2020 · 2004
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments. In Proceedings of the ACL Workshop on Intrinsic and Extrinsic Evaluation Measures for Machine Translation and/or Summarization . 65–72
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Exploring Software Naturalness through Neural Language Models
Luca Buratti, Saurabh Pujar, Mihaela Bornea, Scott McCarley, Yunhui Zheng, Gaetano Rossiello, Alessandro Morari, Jim Laredo, Veronika Thost, Yufan Zhuang, and Giacomo Domeniconi. 2020 · 2006
Earlier work this paper cites.
A Systematic Review of Software Development Cost Estimation Studies
Magne Jorgensen and Martin Shepperd. 2007 · 2007
Earlier work this paper cites.
Towards Automatically Generating Summary Comments for Java Methods. In 25th IEEE/ACM International Conference on Automated Software Engineering, ASE 2010 . ACM, 43–52
Giriprasad Sridhara, Emily Hill, Divya Muppaneni, Lori Pollock, and K Vijay-Shanker. 2010 · 2010
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) . 1724–1734
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Convolutional Neural Networks for Sentence Classification. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing, EMNLP 2014 . ACL, 1746–1751
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Incorporating Copying Mechanism in Sequence-to-Sequence Learning. In Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 1631–1640
Jiatao Gu, Zhengdong Lu, Hang Li, and Victor OK Li. 2016 · 2016
Earlier work this paper cites.
Neural Machine Translation of Rare Words with Subword Units. In Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 1715–1725
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Earlier work this paper cites.
Towards String-To-Tree Neural Machine Translation. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) . 132–140
Roee Aharoni and Yoav Goldberg. 2017 · 2017
Earlier work this paper cites.
OpenNMT: Open-Source Toolkit for Neural Machine Translation. In Proceedings of ACL 2017, System Demonstrations . 67–72
Guillaume Klein, Yoon Kim, Yuntian Deng, Jean Senellart, and Alexander M Rush. 2017 · 2017
Earlier work this paper cites.
A Parallel Corpus of Python Functions and Documentation Strings for Automated Code Documentation and Code Generation. In Proceedings of the Eighth International Joint Conference on Natural Language Processing (Volume 2: Short Papers) . 314–319
Antonio Valerio Miceli-Barone and Rico Sennrich. 2017 · 2017
Earlier work this paper cites.
Get To The Point: Summarization with Pointer-Generator Networks. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 1073–1083
Abigail See, Peter J Liu, and Christopher D Manning. 2017 · 2017
Earlier work this paper cites.
Attention is All You Need. In Advances in Neural Information Processing Systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
code2seq: Generating Sequences from Structured Representations of Code. In International Conference on Learning Representations
Uri Alon, Shaked Brody, Omer Levy, and Eran Yahav. 2018 · 2018
Earlier work this paper cites.
Tree-to-tree Neural Networks for Program Translation. In Advances in Neural Information Processing Systems , Vol. 31. Curran Associates, Inc
Xinyun Chen, Chang Liu, and Dawn Song. 2018 · 2018
Earlier work this paper cites.
Deep Code Search. In 2018 IEEE/ACM 40th International Conference on Software Engineering (ICSE) . IEEE, 933–944
Xiaodong Gu, Hongyu Zhang, and Sunghun Kim. 2018 · 2018
Earlier work this paper cites.
Deep Code Comment Generation. In 2018 IEEE/ACM 26th International Conference on Program Comprehension (ICPC) . IEEE, 200–20010
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2018a · 2018
Earlier work this paper cites.
Summarizing Source Code with Transferred API Knowledge. In Proceedings of the 27th International Joint Conference on Artificial Intelligence, IJCAI 2018 . 2269–2275
Xing Hu, Ge Li, Xin Xia, David Lo, Shuai Lu, and Zhi Jin. 2018b · 2018
Cited alongside, same era.
Deep Contextualized Word Representations. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) . 2227–2237
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Self-Attention with Relative Position Representations. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 2 (Short Papers) . 464–468
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Cited alongside, same era.
Improving Automatic Source Code Summarization via Deep Reinforcement Learning. In Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering . 397–407
CodeBERT: A Pre-Trained Model for Programming and Natural Languages. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: Findings . 1536–1547
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Later among the works it cites.
Deep Code Comment Generation with Hybrid Lexical and Syntactical Information
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2020 · 2020
Later among the works it cites.
Control Flow Graph Embedding Based on Multi-Instance Decomposition for Bug Localization. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 34. 4223–4230
Xuan Huo, Ming Li, and Zhi-Hua Zhou. 2020 · 2020
Later among the works it cites.
Learning and Evaluating Contextual Embedding of Source Code. In Proceedings of the 37th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 119) . PMLR, 5110–5121
Aditya Kanade, Petros Maniatis, Gogul Balakrishnan, and Kensen Shi. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yao Wan, Zhou Zhao, Min Yang, Guandong Xu, Haochao Ying, Jian Wu, and Philip S Yu. 2018 · 2018
Cited alongside, same era.
The Adverse Effects of Code Duplication in Machine Learning Models of Code. In Proceedings of the 2019 ACM SIGPLAN International Symposium on New Ideas, New Paradigms, and Reflections on Programming and Software . 143–153
Miltiadis Allamanis. 2019 · 2019
Cited alongside, same era.
code2vec: Learning Distributed Representations of Code
Uri Alon, Meital Zilberstein, Omer Levy, and Eran Yahav. 2019 · 2019
Cited alongside, same era.
ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators. In International Conference on Learning Representations
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D Manning. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
A Neural Model for Generating Natural Language Summaries of Program Subroutines. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 795–806
Alexander LeClair, Siyuan Jiang, and Collin McMillan. 2019 · 2019
Cited alongside, same era.
Recommendations for Datasets for Source Code Summarization. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . 3931–3937
Alexander LeClair and Collin McMillan. 2019 · 2019
Cited alongside, same era.
Decoupled Weight Decay Regularization. In International Conference on Learning Representations
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
Improved Code Summarization via a Graph Neural Network. In Proceedings of the 28th International Conference on Program Comprehension . 184–195
Alexander LeClair, Sakib Haque, Lingfei Wu, and Collin McMillan. 2020 · 2020
Later among the works it cites.
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 7871–7880
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Multi-task Learning based Pre-trained Language Model for Code Completion. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 473–485
Fang Liu, Ge Li, Yunfei Zhao, and Zhi Jin. 2020 · 2020
Later among the works it cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Later among the works it cites.
Intellicode Compose: Code Generation Using Transformer. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1433–1443
Alexey Svyatkovskiy, Shao Kun Deng, Shengyu Fu, and Neel Sundaresan. 2020 · 2020
Later among the works it cites.
Retrieval-based Neural Source Code Summarization. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering . 1385–1397
Jian Zhang, Xu Wang, Hongyu Zhang, Hailong Sun, and Xudong Liu. 2020 · 2020
Later among the works it cites.
Holistic Combination of Structural and Textual Code Information for Context based API Recommendation
Chi Chen, Xin Peng, Zhenchang Xing, Jun Sun, Xin Wang, Yifan Zhao, and Wenyun Zhao. 2021 · 2021
Later among the works it cites.
ACCL: Architecting Highly Scalable Distributed Training Systems with Highly-Efficient Collective Communication Library
Jianbo Dong, Shaochuang Wang, Fei Feng, Zheng Cao, Heng Pan, Lingbo Tang, Pengcheng Li, Hao Li, Qianyuan Ran, Yiqun Guo, et al · 2021
Later among the works it cites.
GraphCodeBERT: Pre-training Code Representations with Data Flow. In International Conference on Learning Representations, ICLR 2021
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, LIU Shujie, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu, et al · 2021
Later among the works it cites.
TreeBERT: A tree-based pre-trained model for programming language. In Proceedings of the Thirty-Seventh Conference on Uncertainty in Artificial Intelligence , Vol. 161. PMLR, 54–63
Xue Jiang, Zhuoran Zheng, Chen Lyu, Liang Li, and Lei Lyu. 2021 · 2021
Later among the works it cites.
Improving Code Summarization with Block-wise Abstract Syntax Tree Splitting. In 29th IEEE/ACM International Conference on Program Comprehension, ICPC 2021
Chen Lin, Zhichao Ouyang, Junqing Zhuang, Jianqiang Chen, Hui Li, and Rongxin Wu. 2021 · 2021
Later among the works it cites.
Studying the usage of text-to-text transfer transformer to support code-related tasks. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 336–347
Antonio Mastropaolo, Simone Scalabrino, Nathan Cooper, David Nader Palacio, Denys Poshyvanyk, Rocco Oliveto, and Gabriele Bavota. 2021 · 2021
Later among the works it cites.
Copy That! Editing Sequences by Copying Spans. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 13622–13630
Sheena Panthaplackel, Miltiadis Allamanis, and Marc Brockschmidt. 2021 · 2021
Later among the works it cites.
Code Completion by Modeling Flattened Abstract Syntax Trees as Graphs. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 14015–14023
Yanlin Wang and Hui Li. 2021 · 2021
Later among the works it cites.
Code Summarization with Structure-induced Transformer. In Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 . Association for Computational Linguistics, 1078–1090
Hongqiu Wu, Hai Zhao, and Min Zhang. 2021 · 2021
Later among the works it cites.
Exploiting Method Names to Improve Code Summarization: A Deliberation Multi-Task Learning Approach. In 29th IEEE/ACM International Conference on Program Comprehension, ICPC 2021 . IEEE, 138–148
Rui Xie, Wei Ye, Jinan Sun, and Shikun Zhang. 2021 · 2021
Later among the works it cites.
A Multi-Modal Transformer-based Code Summarization Approach for Smart Contracts. In 29th IEEE/ACM International Conference on Program Comprehension, ICPC 2021 . IEEE
Zhen Yang, Jacky Keung, Xiao Yu, Xiaodong Gu, Zhengyuan Wei, Xiaoxue Ma, and Miao Zhang. 2021 · 2021
Later among the works it cites.