Fetching the paper…
Reading the bibliography…
Code intelligence leverages machine learning techniques to extract knowledge from extensive code corpora, with the aim of developing intelligent tools to improve the quality and productivity of computer programming.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, et al · 1901
Earlier work this paper cites.
Combining graph-based learning with automated data collection for code vulnerability detection
Huanting Wang, Guixin Ye, Zhanyong Tang, Shin Hwei Tan, et al · 1958
Earlier work this paper cites.
CoSQL: A Conversational Text-to-SQL Challenge Towards Cross-Domain Natural Language Interfaces to Databases. In EMNLP . 1962–1979
Tao Yu, Rui Zhang, Heyang Er, Suyi Li, Eric Xue, Bo Pang, Xi Victoria Lin, et al · 1979
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W. J. Zhu · 2002
Earlier work this paper cites.
LLVM: A compilation framework for lifelong program analysis & transformation. In CGO . 75–86
Chris Lattner and Vikram Adve. 2004 · 2004
Earlier work this paper cites.
Faster than C: Static type inference with Starkiller
Michael Salib. 2004 · 2004
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
C. Y. Lin · 2004
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
S. Banerjee and A. Lavie · 2005
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Geoffrey E Hinton, Simon Osindero, and Yee-Whye Teh. 2006 · 2006
Earlier work this paper cites.
Why software is eating the world
Marc Andreessen. 2011 · 2011
Earlier work this paper cites.
Efficient Estimation of Word Representations in Vector Space. In ICLR
Tomás Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Alex Graves, Greg Wayne, and Ivo Danihelka. 2014 · 2014
Earlier work this paper cites.
Structured generative models of natural source code. In ICML . 649–657
Chris Maddison and Daniel Tarlow. 2014 · 2014
Earlier work this paper cites.
Code completion with statistical language models. In ICPC . 419–428
Veselin Raychev, Martin Vechev, and Eran Yahav. 2014 · 2014
Earlier work this paper cites.
Modeling and discovering vulnerabilities with code property graphs. In S&P . 590–604
Fabian Yamaguchi, Nico Golde, Daniel Arp, and Konrad Rieck. 2014 · 2014
Earlier work this paper cites.
How can I use this method?. In ICSE , Vol. 1. 880–890
Laura Moreno, Gabriele Bavota, Massimiliano Di Penta, Rocco Oliveto, and Andrian Marcus. 2015 · 2015
Earlier work this paper cites.
Toward deep learning software repositories. In MSR . 334–345
Martin White, Christopher Vendome, Mario Linares-Vásquez, and Denys Poshyvanyk. 2015 · 2015
Earlier work this paper cites.
Gated feedback recurrent neural networks
Junyoung Chung, Caglar Gulcehre, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Improved semantic parsers for if-then statements. In ACL . 726–736
Islam Beltagy and Chris Quirk. 2016 · 2016
Earlier work this paper cites.
Automated correction for syntax errors in programming assignments using recurrent neural networks
Sahil Bhatia and Rishabh Singh. 2016 · 2016
Earlier work this paper cites.
Language to Logical Form with Neural Attention. In ACL
Li Dong and Mirella Lapata. 2016 · 2016
Earlier work this paper cites.
Deep API learning. In Proceedings of the 2016 24th ACM SIGSOFT International Symposium on Foundations of Software Engineering . 631–642
Xiaodong Gu, Hongyu Zhang, Dongmei Zhang, and Sunghun Kim. 2016 · 2016
Earlier work this paper cites.
Gated Graph Sequence Neural Networks. In ICLR
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard S. Zemel. 2016 · 2016
Earlier work this paper cites.
Latent Predictor Networks for Code Generation. In ACL . 599–609
Wang Ling, Phil Blunsom, Edward Grefenstette, Karl Moritz Hermann, Tomáš Kočiskỳ, Fumin Wang, and Andrew Senior. 2016 · 2016
Earlier work this paper cites.
Latent attention for if-then program synthesis
Chang Liu, Xinyun Chen, Eui Chul Shin, Mingcheng Chen, and Dawn Song. 2016a · 2016
Earlier work this paper cites.
Neural code completion
Chang Liu, Xin Wang, Richard Shin, Joseph E Gonzalez, and Dawn Song. 2016b · 2016
Earlier work this paper cites.
Convolutional neural networks over tree structures for programming language processing. In AAAI , Vol. 30
Lili Mou, Ge Li, Lu Zhang, Tao Wang, and Zhi Jin. 2016 · 2016
Earlier work this paper cites.
Probabilistic model for code with decision trees
Veselin Raychev, Pavol Bielik, and Martin Vechev. 2016 · 2016
Earlier work this paper cites.
SVF: interprocedural static value-flow analysis in LLVM. In Proceedings of the 25th international conference on compiler construction . 265–266
Yulei Sui and Jingling Xue. 2016 · 2016
Earlier work this paper cites.
Automatically learning semantic features for defect prediction. In ICSE . 297–308
Song Wang, Taiyue Liu, and Lin Tan. 2016 · 2016
Earlier work this paper cites.
Deep learning code fragments for code clone detection. In ASE . 87–98
Martin White, Michele Tufano, Christopher Vendome, and Denys Poshyvanyk. 2016 · 2016
Earlier work this paper cites.
Probabilistic model for code with decision trees
Veselin Raychev, Pavol Bielik, and Martin Vechev · 2016
Earlier work this paper cites.
Smartpaste: Learning to adapt source code
Miltiadis Allamanis and Marc Brockschmidt. 2017 · 2017
Earlier work this paper cites.
DeepCoder: Learning to Write Programs. In ICLR
Matej Balog, Alexander L. Gaunt, Marc Brockschmidt, Sebastian Nowozin, and Daniel Tarlow. 2017 · 2017
Earlier work this paper cites.
A Parallel Corpus of Python Functions and Documentation Strings for Automated Code Documentation and Code Generation. In IJCNLP . 314–319
Antonio Valerio Miceli Barone and Rico Sennrich. 2017 · 2017
Earlier work this paper cites.
Towards evaluating the robustness of neural networks. In S&P . 39–57
Nicholas Carlini and David Wagner. 2017 · 2017
Earlier work this paper cites.
Synthesizing benchmarks for predictive modeling. In CGO . 86–99
Chris Cummins, Pavlos Petoumenos, Zheng Wang, and Hugh Leather. 2017 · 2017
Earlier work this paper cites.
Robustfill: Neural program learning under noisy i/o. In ICML . 990–998
Jacob Devlin, Jonathan Uesato, Surya Bhupatiraju, Rishabh Singh, Abdel-rahman Mohamed, and Pushmeet Kohli. 2017 · 2017
Earlier work this paper cites.
DeepAM: Migrate APIs with Multi-Modal Sequence to Sequence Learning. In IJCAI . 3675–3681
Xiaodong Gu, Hongyu Zhang, Dongmei Zhang, and Sunghun Kim. 2017 · 2017
Earlier work this paper cites.
Deepfix: Fixing common c language errors by deep learning. In AAAI
Rahul Gupta, Soham Pal, Aditya Kanade, and Shirish Shevade. 2017 · 2017
Earlier work this paper cites.
An unsupervised approach for discovering relevant tutorial fragments for APIs. In ICSE . 38–48
He Jiang, Jingxuan Zhang, Zhilei Ren, and Tao Zhang. 2017 · 2017
Earlier work this paper cites.
Exploring API embedding for API usages and applications. In ICSE . 438–449
Trong Duc Nguyen, Anh Tuan Nguyen, Hung Dang Phan, and Tien N Nguyen. 2017 · 2017
Earlier work this paper cites.
Abstract Syntax Networks for Code Generation and Semantic Parsing. In ACL . 1139–1149
Maxim Rabinovich, Mitchell Stern, and Dan Klein. 2017 · 2017
Earlier work this paper cites.
Neural Programming by Example. In AAAI . 1539–1545
Chengxun Shu and Hongyu Zhang. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In NeurIPS . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Supervised Deep Features for Software Functional Clone Detection by Exploiting Lexical and Syntactical Information in Source Code.. In IJCAI . 3034–3040
Huihui Wei and Ming Li. 2017 · 2017
Earlier work this paper cites.
A Syntactic Neural Model for General-Purpose Code Generation. In ACL . 440–450
Pengcheng Yin and Graham Neubig. 2017 · 2017
Earlier work this paper cites.
Seq2sql: Generating structured queries from natural language using reinforcement learning
Victor Zhong, Caiming Xiong, and Richard Socher. 2017 · 2017
Earlier work this paper cites.
A parallel corpus of python functions and documentation strings for automated code documentation and code generation
Antonio Valerio Miceli Barone and Rico Sennrich · 2017
Earlier work this paper cites.
A survey of machine learning for big code and naturalness
Miltiadis Allamanis, Earl T Barr, Premkumar Devanbu, and Charles Sutton. 2018a · 2018
Earlier work this paper cites.
code2seq: Generating Sequences from Structured Representations of Code. In ICLR
Uri Alon, Shaked Brody, Omer Levy, and Eran Yahav. 2018 · 2018
Earlier work this paper cites.
Neural Code Comprehension: A Learnable Representation of Code Semantics. In NeurIPS . 3589–3601
Tal Ben-Nun, Alice Shoshana Jakobovits, and Torsten Hoefler. 2018 · 2018
Earlier work this paper cites.
An Encoder-Decoder Framework Translating Natural Language to Database Queries. In IJCAI . 3977–3983
Ruichu Cai, Boyan Xu, Zhenjie Zhang, Xiaoyan Yang, Zijian Li, and Zhihao Liang. 2018 · 2018
Earlier work this paper cites.
Tree-to-tree Neural Networks for Program Translation. In NeurIPS . 2552–2562
Xinyun Chen, Chang Liu, and Dawn Song. 2018 · 2018
Earlier work this paper cites.
Automatic feature learning for predicting vulnerable software components
Hoa Khanh Dam, Truyen Tran, Trang Thi Minh Pham, Shien Wee Ng, John Grundy, and Aditya Ghose. 2018 · 2018
Earlier work this paper cites.
Robust physical-world attacks on deep learning visual classification. In CVPR . 1625–1634
Kevin Eykholt, Ivan Evtimov, Earlence Fernandes, Bo Li, Amir Rahmati, Chaowei Xiao, Atul Prakash, Tadayoshi Kohno, and Dawn Song. 2018 · 2018
Earlier work this paper cites.
AllenNLP: A Deep Semantic Natural Language Processing Platform. In Proceedings of Workshop for NLP Open Source Software (NLP-OSS) . 1–6
Matt Gardner, Joel Grus, Mark Neumann, Oyvind Tafjord, et al · 2018
Earlier work this paper cites.
Deep code search. In ICSE . 933–944
Xiaodong Gu, Hongyu Zhang, and Sunghun Kim. 2018 · 2018
Earlier work this paper cites.
Deep reinforcement learning for programming language correction
Rahul Gupta, Aditya Kanade, and Shirish Shevade. 2018 · 2018
Earlier work this paper cites.
Learning to Repair Software Vulnerabilities with Generative Adversarial Networks. In NeurIPS . 7944–7954
Jacob Harer, Onur Ozdemir, Tomo Lazovich, Christopher P. Reale, Rebecca L. Russell, Louis Y. Kim, and Sang Peter Chin. 2018 · 2018
Earlier work this paper cites.
MaxSMT-based type inference for Python 3. In International Conference on Computer Aided Verification . 12–19
Mostafa Hassan, Caterina Urban, Marco Eilers, and Peter Müller. 2018 · 2018
Earlier work this paper cites.
Learning to generate corrective patches using neural machine translation
Hideaki Hata, Emad Shihab, and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
Retrieval-Based Neural Code Generation. In EMNLP . 925–930
Shirley Anugrah Hayati, Raphael Olivier, Pravalika Avvaru, Pengcheng Yin, Anthony Tomasic, and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
Deep learning type inference. In ESEC/FSE . 152–162
Vincent J Hellendoorn, Christian Bird, Earl T Barr, and Miltiadis Allamanis. 2018 · 2018
Earlier work this paper cites.
Code vectors: Understanding programs through embedded abstracted symbolic traces. In ESEC/FSE . 163–174
Jordan Henkel, Shuvendu K Lahiri, Ben Liblit, and Thomas Reps. 2018 · 2018
Earlier work this paper cites.
Summarizing source code with transferred api knowledge.(2018). In IJCAI , Vol. 19. 2269–2275
Xing Hu, Ge Li, Xin Xia, David Lo, Shuai Lu, and Zhi Jin. 2018b · 2018
Earlier work this paper cites.
Mapping Language to Code in Programmatic Context. In EMNLP . 1643–1652
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
Maximal divergence sequential autoencoder for binary software vulnerability detection. In ICLR
Tue Le, Tuan Nguyen, Trung Le, Dinh Phung, Paul Montague, Olivier De Vel, and Lizhen Qu. 2018 · 2018
Earlier work this paper cites.
Adapting neural text classification for improved software categorization. In ICSME . 461–472
Alexander LeClair, Zachary Eberhart, and Collin McMillan. 2018 · 2018
Earlier work this paper cites.
Deepbugs: A learning approach to name-based bug detection
Michael Pradel and Koushik Sen. 2018 · 2018
Earlier work this paper cites.
Syntax and sensibility: Using language models to detect and correct syntax errors. In SANER . 311–322
Eddie Antonio Santos, Joshua Charles Campbell, Dhvani Patel, Abram Hindle, and José Nelson Amaral. 2018 · 2018
Earlier work this paper cites.
Neural Program Repair by Jointly Learning to Localize and Repair. In ICLR
Marko Vasic, Aditya Kanade, Petros Maniatis, David Bieber, and Rishabh Singh. 2018 · 2018
Earlier work this paper cites.
Improving automatic source code summarization via deep reinforcement learning. In ASE . 397–407
Yao Wan, Zhou Zhao, Min Yang, Guandong Xu, Haochao Ying, Jian Wu, and Philip S Yu. 2018 · 2018
Earlier work this paper cites.
Deepsim: deep learning code functional similarity. In ESEC/FSE . 141–151
Gang Zhao and Jeff Huang. 2018 · 2018
Earlier work this paper cites.
Improving automatic source code summarization via deep reinforcement learning
Yao Wan, Zhou Zhao, Min Yang, Guandong Xu, Haochao Ying, Jian Wu, and Philip S Yu · 2018
Earlier work this paper cites.
Deep code comment generation
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin · 2018
Earlier work this paper cites.
Deep learning type inference
Vincent J Hellendoorn, Christian Bird, Earl T Barr, and Miltiadis Allamanis · 2018
Earlier work this paper cites.
Staqc: A systematically mined question-code dataset from stack overflow
Ziyu Yao, Daniel S Weld, Wei-Peng Chen, and Huan Sun · 2018
Earlier work this paper cites.
StackOverflow
2019 · 2019
Earlier work this paper cites.
code2vec: Learning distributed representations of code
Uri Alon, Meital Zilberstein, Omer Levy, and Eran Yahav. 2019 · 2019
Earlier work this paper cites.
AutoPandas: neural-backed generators for program synthesis
Rohan Bavishi, Caroline Lemieux, Roy Fox, Koushik Sen, and Ion Stoica. 2019 · 2019
Earlier work this paper cites.
Generative Code Modeling with Graphs. In ICLR
Marc Brockschmidt, Miltiadis Allamanis, Alexander L. Gaunt, and Oleksandr Polozov. 2019 · 2019
Earlier work this paper cites.
Learning-based recursive aggregation of abstract syntax trees for code clone detection. In SANER . 95–104
Lutz Büch and Artur Andrzejak. 2019 · 2019
Earlier work this paper cites.
When deep learning met code search. In ESEC/FSE . 964–974
Jose Cambronero, Hongyu Li, Seohyun Kim, Koushik Sen, and Satish Chandra. 2019 · 2019
Earlier work this paper cites.
On evaluating adversarial robustness
Nicholas Carlini, Anish Athalye, Nicolas Papernot, Wieland Brendel, Jonas Rauber, Dimitris Tsipras, Ian Goodfellow, Aleksander Madry, and Alexey Kurakin. 2019 · 2019
Earlier work this paper cites.
Sequencer: Sequence-to-sequence learning for end-to-end program repair
Zimin Chen, Steve James Kommrusch, Michele Tufano, Louis-Noël Pouchet, Denys Poshyvanyk, and Martin Monperrus. 2019 · 2019
Earlier work this paper cites.
Open vocabulary learning on source code with a graph-structured cache. In ICML . 1475–1485
Milan Cvitkovic, Badal Singh, and Animashree Anandkumar. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Structured Neural Summarization. In ICLR
Patrick Fernandes, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Cited alongside, same era.
Neural attribution for semantic bug-localization in student programs
R Gupta, A Kanade, and S Shevade. 2019 · 2019
Cited alongside, same era.
Codesearchnet challenge: Evaluating the state of semantic code search
Hamel Husain, Ho-Hsiang Wu, Tiferet Gazit, Miltiadis Allamanis, and Marc Brockschmidt. 2019 · 2019
Cited alongside, same era.
A manual inspection of Defects4J bugs and its implications for automatic program repair
Jiajun Jiang, Yingfei Xiong, and Xin Xia. 2019 · 2019
Cited alongside, same era.
SySeVR: A framework for using deep learning to detect software vulnerabilities
Zhen Li, Deqing Zou, Shouhuai Xu, Hai Jin, Yawei Zhu, and Zhaoxuan Chen. 2021d · 2021
Later among the works it cites.
Improving Code Summarization with Block-wise Abstract Syntax Tree Splitting. In ICPC . IEEE, 184–195
Chen Lin, Zhichao Ouyang, Junqing Zhuang, Jianqiang Chen, Hui Li, and Rongxin Wu. 2021 · 2021
Later among the works it cites.
Deep Graph Matching and Searching for Semantic Code Retrieval
Xiang Ling, Lingfei Wu, Saizhuo Wang, Gaoning Pan, Tengfei Ma, Fangli Xu, Alex X. Liu, Chunming Wu, and Shouling Ji. 2021a · 2021
Later among the works it cites.
A Practical Black-box Attack on Source Code Authorship Identification Classifiers
Qianjun Liu, Shouling Ji, Changchang Liu, and Chunming Wu. 2021a · 2021
Later among the works it cites.
Combining Graph Neural Networks with Expert Knowledge for Smart Contract Vulnerability Detection
Zhenguang Liu, Peng Qian, Xiaoyang Wang, Yuan Zhuang, Lin Qiu, and Xun Wang. 2021b · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A neural model for generating natural language summaries of program subroutines. In ICSE . 795–806
Alexander LeClair, Siyuan Jiang, and Collin McMillan. 2019 · 2019
Cited alongside, same era.
Improving bug detection via context-based code representation learning and attention-based neural networks
Yi Li, Shaohua Wang, Tien N Nguyen, and Son Van Nguyen. 2019 · 2019
Cited alongside, same era.
Learning to spot and refactor inconsistent method names. In ICSE . 1–12
Kui Liu, Dongsun Kim, Tegawendé F Bissyandé, Taeyoung Kim, Kisub Kim, Anil Koyuncu, Suntae Kim, and Yves Le Traon. 2019 · 2019
Cited alongside, same era.
NL2Type: inferring JavaScript function types from natural language information. In ICSE . 304–315
Rabee Sohail Malik, Jibesh Patra, and Michael Pradel. 2019 · 2019
Cited alongside, same era.
DeepDelta: learning to repair compilation errors. In ESEC/FSE . 925–936
Ali Mesbah, Andrew Rice, Emily Johnston, Nick Glorioso, and Edward Aftandilian. 2019 · 2019
Cited alongside, same era.
Clcdsa: cross language code clone detection using syntactical features and api documentation. In ASE . 1026–1037
Kawser Wazed Nafi, Tonny Shekha Kar, Banani Roy, Chanchal K Roy, and Kevin A Schneider. 2019 · 2019
Cited alongside, same era.
Learning to infer program sketches. In ICML . 4861–4870
Maxwell Nye, Luke Hewitt, Joshua Tenenbaum, and Armando Solar-Lezama. 2019 · 2019
Cited alongside, same era.
CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation. In NeurIPS Datasets and Benchmarks
Shuai Lu, Daya Guo, Shuo Ren, Junjie Huang, Alexey Svyatkovskiy, et al · 2021
Later among the works it cites.
Studying the usage of text-to-text transfer transformer to support code-related tasks. In ICSE . 336–347
Antonio Mastropaolo, Simone Scalabrino, Nathan Cooper, David Nader Palacio, Denys Poshyvanyk, et al · 2021
Later among the works it cites.
Modeling Functional Similarity in Source Code with Graph-Based Siamese Networks
Nikita Mehrotra, Navdha Agarwal, Piyush Gupta, Saket Anand, David Lo, and Rahul Purandare. 2021 · 2021
Later among the works it cites.
Adversarial Attacks to API Recommender Systems: Time to Wake Up and Smell the Coffee?. In ASE . 253–265
Phuong T Nguyen, Claudio Di Sipio, Juri Di Rocco, Massimiliano Di Penta, and Davide Di Ruscio. 2021 · 2021
Later among the works it cites.
Deep Just-In-Time Inconsistency Detection Between Comments and Source Code. In AAAI , Vol. 35. 427–435
Sheena Panthaplackel, Junyi Jessy Li, Milos Gligoric, and Raymond J Mooney. 2021 · 2021
Later among the works it cites.
How could Neural Networks understand Programs?. In ICML , Vol. 139. 8476–8486
Dinglan Peng, Shuxin Zheng, Yatao Li, Guolin Ke, Di He, and Tie-Yan Liu. 2021 · 2021
Later among the works it cites.
PyExplainer: Explaining the Predictions of Just-In-Time Defect Models. In ASE . 407–418
Chanathip Pornprasit, Chakkrit Tantithamthavorn, Jirayus Jiarpakdee, Michael Fu, and Patanamon Thongtanunam. 2021 · 2021
Later among the works it cites.
Understanding neural code intelligence through program simplification. In ESEC/FSE . ACM, 441–452
Md. Rafiqul Islam Rabin, Vincent J. Hellendoorn, and Mohammad Amin Alipour. 2021 · 2021
Later among the works it cites.
You autocomplete me: Poisoning vulnerabilities in neural code completion. In USENIX Security
Roei Schuster, Congzheng Song, Eran Tromer, and Vitaly Shmatikov. 2021 · 2021
Later among the works it cites.
Explanation-Guided Backdoor Poisoning Attacks Against Malware Classifiers. In USENIX Security
Giorgio Severi, Jim Meyer, Scott Coull, and Alina Oprea. 2021 · 2021
Later among the works it cites.
API2Com: On the Improvement of Automatically Generated Code Comments Using API Documentations. In ICPC . IEEE, 411–421
Ramin Shahbazi, Rishab Sharma, and Fatemeh H. Fard. 2021 · 2021
Later among the works it cites.
CAST: Enhancing Code Summarization with Hierarchical Splitting and Reconstruction of Abstract Syntax Trees. In EMNLP . 4053–4062
Ensheng Shi, Yanlin Wang, Lun Du, Hongyu Zhang, Shi Han, et al · 2021
Later among the works it cites.
Fast and memory-efficient neural code completion. In MSR . 329–340
Alexey Svyatkovskiy, Sebastian Lee, Anna Hadjitofi, Maik Riechert, Juliana Vicente Franco, and Miltiadis Allamanis. 2021 · 2021
Later among the works it cites.
SynCoBERT: Syntax-Guided Multi-Modal Contrastive Pre-Training for Code Representation
Xin Wang, Yasheng Wang, Fei Mi, Pingyi Zhou, Yao Wan, Xiao Liu, Li Li, Hao Wu, Jin Liu, and Xin Jiang. 2021b · 2021
Later among the works it cites.
Code completion by modeling flattened abstract syntax trees as graphs. In AAAI , Vol. 35. 14015–14023
Yanlin Wang and Hui Li. 2021 · 2021
Later among the works it cites.
Code Summarization with Structure-induced Transformer. In Findings of ACL . 1078–1090
Hongqiu Wu, Hai Zhao, and Min Zhang. 2021 · 2021
Later among the works it cites.
Exploiting Method Names to Improve Code Summarization: A Deliberation Multi-Task Learning Approach. In ICPC . IEEE, 138–148
Rui Xie, Wei Ye, Jinan Sun, and Shikun Zhang. 2021 · 2021
Later among the works it cites.
A Multi-Modal Transformer-based Code Summarization Approach for Smart Contracts. In ICPC . IEEE, 1–12
Zhen Yang, Jacky Keung, Xiao Yu, Xiaodong Gu, Zhengyuan Wei, Xiaoxue Ma, and Miao Zhang. 2021 · 2021
Later among the works it cites.
Interpretable Program Synthesis. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–16
Tianyi Zhang, Zhiyang Chen, Yuanli Zhu, Priyan Vaithilingam, Xinyu Wang, and Elena L Glassman. 2021a · 2021
Later among the works it cites.
Assessing Generalizability of CodeBERT. In ICSME . IEEE, 425–436
Xin Zhou, DongGyun Han, and David Lo. 2021 · 2021
Later among the works it cites.
A syntax-guided edit decoder for neural program repair. In ESEC/FSE . 341–353
Qihao Zhu, Zeyu Sun, Yuan-an Xiao, Wenjie Zhang, Kang Yuan, Yingfei Xiong, and Lu Zhang. 2021 · 2021
Later among the works it cites.
Interpreting deep learning-based vulnerability detector predictions based on heuristic searching
Deqing Zou, Yawei Zhu, Shouhuai Xu, Zhen Li, Hai Jin, and Hengkai Ye. 2021 · 2021
Later among the works it cites.
Language-Agnostic Representation Learning of Source Code from Structure and Context. In ICLR
Daniel Zügner, Tobias Kirschstein, Michele Catasta, Jure Leskovec, and Stephan Günnemann. 2021 · 2021
Later among the works it cites.
Cosqa: 20, 000+ web queries for code search and question answering
Junjie Huang, Duyu Tang, Linjun Shou, Ming Gong, Ke Xu, Daxin Jiang, Ming Zhou, and Nan Duan · 2021
Later among the works it cites.
Project codenet: A large-scale AI for code dataset for learning a diversity of coding tasks
Ruchir Puri, David S. Kung, Geert Janssen, Wei Zhang, Giacomo Domeniconi, et al · 2021
Later among the works it cites.
ProGraML: A Graph-based Program Representation for Data Flow Analysis and Compiler Optimizations
Chris Cummins, Zacharias Fisches, Tal Ben-Nun, Torsten Hoefler, Michael O’Boyle, and Hugh Leather · 2021
Later among the works it cites.
Codesc: A large code-description parallel dataset
Masum Hasan, Tanveer Muttaqueen, Abdullah Al Ishtiaq, Kazi Sajeed Mehrab, Md. Mahim Anjum Haque, Tahmid Hasan, Wasi Uddin Ahmad, Anindya Iqbal, and Rifat Shahriyar · 2021
Later among the works it cites.
Avatar: A parallel corpus for java-python program translation
Wasi Uddin Ahmad, Md Golam Rahman Tushar, Saikat Chakraborty, and Kai-Wei Chang · 2021
Later among the works it cites.
Pytorrent: A python library corpus for large-scale language models
Mehdi Bahrami, NC Shrikanth, Shade Ruangwan, Lei Liu, Yuji Mizobuchi, Masahiro Fukuyori, Wei-Peng Chen, Kazuki Munakata, and Tim Menzies · 2021
Later among the works it cites.
Codeqa: A question answering dataset for source code comprehension
Chenxiao Liu and Xiaojun Wan · 2021
Later among the works it cites.
Codexglue: A machine learning benchmark dataset for code understanding and generation
Shuai Lu, Daya Guo, Shuo Ren, Junjie Huang, Alexey Svyatkovskiy, et al · 2021
Later among the works it cites.
Multilingual training for Software Engineering. In ICSE
Toufique Ahmed and Premkumar Devanbu. 2022 · 2022
Later among the works it cites.
MVD: Memory-Related Vulnerability Detection Based on Flow-Sensitive Graph Neural Networks. In ICSE . 1456–1468
Sicong Cao, Xiaobing Sun, Lili Bo, Rongxin Wu, Bin Li, and Chuanqi Tao. 2022 · 2022
Later among the works it cites.
ERNIE-Code: Beyond English-Centric Cross-lingual Pretraining for Programming Languages
Yekun Chai, Shuohuan Wang, Chao Pang, Yu Sun, Hao Tian, and Hua Wu. 2022a · 2022
Later among the works it cites.
On the transferability of pre-trained language models for low-resource programming languages. In ICPC . ACM, 401–412
Fuxiang Chen, Fatemeh H. Fard, David Lo, and Timofey Bryksin. 2022 · 2022
Later among the works it cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Later among the works it cites.
Pangu-coder: Program synthesis with function-level language modeling
Fenia Christopoulou, Gerasimos Lampouras, Milan Gritta, Guchun Zhang, Yinpeng Guo, Zhongqi Li, Qi Zhang, Meng Xiao, Bo Shen, Lin Li, et al · 2022
Later among the works it cites.
Zero-shot program representation learning. In ICPC . ACM, 60–70
Nan Cui, Yuze Jiang, Xiaodong Gu, and Beijun Shen. 2022 · 2022
Later among the works it cites.
Fine-grained Co-Attentive Representation Learning for Semantic Code Search. In SANER . 396–407
Zhongyang Deng, Ling Xu, Chao Liu, Meng Yan, Zhou Xu, and Yan Lei. 2022 · 2022
Later among the works it cites.
Towards Learning (Dis)-Similarity of Source Code from Program Contrasts. In ACL . 6300–6312
Yangruibo Ding, Luca Buratti, Saurabh Pujar, Alessandro Morari, Baishakhi Ray, and Saikat Chakraborty. 2022 · 2022
Later among the works it cites.
Incoder: A generative model for code infilling and synthesis
Daniel Fried, Armen Aghajanyan, Jessy Lin, Sida Wang, Eric Wallace, Freda Shi, Ruiqi Zhong, Wen-tau Yih, Luke Zettlemoyer, and Mike Lewis. 2022 · 2022
Later among the works it cites.
VulRepair: a T5-based automated software vulnerability repair. In ESEC/FSE . 935–947
Michael Fu, Chakkrit Tantithamthavorn, Trung Le, Van Nguyen, and Dinh Q. Phung. 2022 · 2022
Later among the works it cites.
M2TS: multi-scale multi-modal approach based on transformer for source code summarization. In ICPC . ACM, 24–35
Yuexiu Gao and Chen Lyu. 2022 · 2022
Later among the works it cites.
Source Code Summarization with Structural Relative Position Guided Transformer. In SANER . 13–24
Zi Gong, Cuiyun Gao, Yasheng Wang, Wenchao Gu, Yun Peng, and Zenglin Xu. 2022 · 2022
Later among the works it cites.
Accelerating Code Search with Deep Hashing and Code Classification. In ACL . 2534–2544
Wenchao Gu, Yanlin Wang, Lun Du, Hongyu Zhang, Shi Han, Dongmei Zhang, and Michael R. Lyu. 2022 · 2022
Later among the works it cites.
Cross-Language Binary-Source Code Matching with Intermediate Representations. In SANER
Yi Gui, Yao Wan, Hongyu Zhang, Huifang Huang, Yulei Sui, Guandong Xu, Zhiyuan Shao, and Hai Jin. 2022 · 2022
Later among the works it cites.
On the effectiveness of pretrained models for API learning. In ICPC . ACM, 309–320
Mohammad Abdul Hadi, Imam Nur Bani Yusuf, Ferdian Thung, Kien Gia Luong, Lingxiao Jiang, Fatemeh H. Fard, and David Lo. 2022 · 2022
Later among the works it cites.
TreeCen: Building Tree Graph for Scalable Semantic Code Clone Detection. In ASE . ACM, 109:1–109:12
Yutao Hu, Deqing Zou, Junru Peng, Yueming Wu, Junjie Shan, and Hai Jin. 2022 · 2022
Later among the works it cites.
Prompt-tuned Code Language Model as a Neural Knowledge Base for Type Inference in Statically-Typed Partial Code. In ASE . ACM, 79:1–79:13
Qing Huang, Zhiqiang Yuan, Zhenchang Xing, Xiwei Xu, Liming Zhu, and Qinghua Lu. 2022 · 2022
Later among the works it cites.
Automatically Generating Code Comment Using Heterogeneous Graph Neural Networks. In SANER . 1078–1088
Dun Jin, Peiyu Liu, and Zhenfang Zhu. 2022 · 2022
Later among the works it cites.
Competition-Level Code Generation with AlphaCode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, et al · 2022
Later among the works it cites.
AST-Probe: Recovering abstract syntax trees from hidden representations of pre-trained language models. In ASE
José Antonio Hernández López, Martin Weyssow, Jesús Sánchez Cuadrado, and Houari A. Sahraoui. 2022 · 2022
Later among the works it cites.
ReACC: A Retrieval-Augmented Code Completion Framework. In ACL . 6227–6240
Shuai Lu, Nan Duan, Hojae Han, Daya Guo, Seung-won Hwang, and Alexey Svyatkovskiy. 2022 · 2022
Later among the works it cites.
Type4Py: Practical Deep Similarity Learning-Based Type Inference for Python. In ICSE . 2241–2252
Amir M. Mir, Evaldas Latoskinas, Sebastian Proksch, and Georgios Gousios. 2022 · 2022
Later among the works it cites.
Automatic Comment Generation via Multi-Pass Deliberation. In ASE . ACM, 14:1–14:12
Fangwen Mu, Xiao Chen, Lin Shi, Song Wang, and Qing Wang. 2022 · 2022
Later among the works it cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Later among the works it cites.
SPT-Code: Sequence-to-Sequence Pre-Training for Learning Source Code Representations. In ICSE . 1–13
Changan Niu, Chuanyi Li, Vincent Ng, Jidong Ge, Liguo Huang, and Bin Luo. 2022 · 2022
Later among the works it cites.
Synchromesh: Reliable Code Generation from Pre-trained Language Models. In ICLR
Gabriel Poesia, Alex Polozov, Vu Le, Ashish Tiwari, Gustavo Soares, Christopher Meek, and Sumit Gulwani. 2022 · 2022
Later among the works it cites.
Backdoors in Neural Models of Source Code. In ICPR . IEEE, 2892–2899
Goutham Ramakrishnan and Aws Albarghouthi. 2022 · 2022
Later among the works it cites.
Leveraging Automated Unit Tests for Unsupervised Code Translation. In ICLR
Baptiste Rozière, Jie Zhang, François Charton, Mark Harman, Gabriel Synnaeve, and Guillaume Lample. 2022 · 2022
Later among the works it cites.
An exploratory study on code attention in BERT. In ICPC . ACM, 437–448
Rishab Sharma, Fuxiang Chen, Fatemeh H. Fard, and David Lo. 2022 · 2022
Later among the works it cites.
Heterogeneous Information Networks: the Past, the Present, and the Future
Yizhou Sun, Jiawei Han, Xifeng Yan, Philip S. Yu, and Tianyi Wu. 2022b · 2022
Later among the works it cites.
AST-Trans: Code Summarization with Efficient Tree-Structured Attention. In ICSE
Ze Tang, Xiaoyu Shen, Chuanyi Li, Jidong Ge, Liguo Huang, Zheling Zhu, and Bin Luo. 2022 · 2022
Later among the works it cites.
C4: contrastive cross-language code clone detection. In ICPC . ACM, 413–424
Chenning Tao, Qi Zhan, Xing Hu, and Xin Xia. 2022 · 2022
Later among the works it cites.
CLEAR: contrastive learning for API recommendation. In ICSE . 376–387
Moshi Wei, Nima Shiri Harzevili, Yuchao Huang, Junjie Wang, and Song Wang. 2022 · 2022
Later among the works it cites.
Low-Resources Project-Specific Code Summarization. In ASE . ACM, 68:1–68:12
Rui Xie, Tianxiang Hu, Wei Ye, and Shikun Zhang. 2022 · 2022
Later among the works it cites.
A Survey on Deep Learning for Software Engineering
Yanming Yang, Xin Xia, David Lo, and John Grundy. 2022c · 2022
Later among the works it cites.
CoditT5: Pretraining for Source Code and Natural Language Editing. In 37th IEEE/ACM International Conference on Automated Software Engineering, ASE 2022, Rochester, MI, USA, October 10-14, 2022 . ACM, 22:1–22:12
Jiyang Zhang, Sheena Panthaplackel, Pengyu Nie, Junyi Jessy Li, and Milos Gligoric. 2022a · 2022
Later among the works it cites.
Adversarial Robustness of Deep Code Comment Generation
Yu Zhou, Xiaoqing Zhang, Juanjuan Shen, Tingting Han, Taolue Chen, and Harald C. Gall. 2022 · 2022
Later among the works it cites.
A Simple Retrieval-based Method for Code Comment Generation. In SANER . 1089–1100
Xiaoning Zhu, Chaofeng Sha, and Junyu Niu. 2022 · 2022
Later among the works it cites.
Suriya Gunasekar, Yi Zhang, Jyoti Aneja, Caio César Teodoro Mendes, Allie Del Giorno, Sivakanth Gopi, Mojan Javaheripi, Piero Kauffmann, Gustavo de Rosa, Olli Saarikivi, et al · 2023
Closest in time.
Enabling Programming Thinking in Large Language Models Toward Code Generation
Jia Li, Ge Li, Yongmin Li, and Zhi Jin. 2023b · 2023
Closest in time.
Large Language Model-Aware In-Context Learning for Code Generation
Jia Li, Ge Li, Chongyang Tao, Huangzhao Zhang, Fang Liu, and Zhi Jin. 2023c · 2023
Closest in time.
StarCoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, et al · 2023
Closest in time.
Improving ChatGPT Prompt for Code Generation
Chao Liu, Xuanlin Bao, Hongyu Zhang, Neng Zhang, Haibo Hu, Xiaohong Zhang, and Meng Yan. 2023 · 2023
Closest in time.
WizardCoder: Empowering Code Large Language Models with Evol-Instruct
Ziyang Luo, Can Xu, Pu Zhao, Qingfeng Sun, Xiubo Geng, Wenxiang Hu, Chongyang Tao, Jing Ma, Qingwei Lin, and Daxin Jiang. 2023 · 2023
Closest in time.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al · 2023
Closest in time.
Codet5+: Open code large language models for code understanding and generation
Yue Wang, Hung Le, Akhilesh Deepak Gotmare, Nghi DQ Bui, Junnan Li, and Steven CH Hoi. 2023 · 2023
Closest in time.
Automated program repair in the era of large pre-trained language models. In Proceedings of the 45th International Conference on Software Engineering (ICSE 2023). Association for Computing Machinery
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Closest in time.
Codegeex: A pre-trained model for code generation with multilingual evaluations on humaneval-x
Qinkai Zheng, Xiao Xia, Xu Zou, Yuxiao Dong, Shan Wang, Yufei Xue, Zihan Wang, Lei Shen, Andi Wang, Yang Li, et al · 2023
Closest in time.
Large Language Models are Few-Shot Summarizers: Multi-Intent Comment Generation via In-Context Learning
Mingyang Geng, Shangwen Wang, Dezun Dong, Haotian Wang, Ge Li, Zhi Jin, Xiaoguang Mao, and Xiangke Liao. 2024 · 2024
Closest in time.
Summarizing source code using a neural attention model. In ACL . 2073–2083
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung, and Luke Zettlemoyer. 2016 · 2083
Closest in time.
A convolutional attention network for extreme summarization of source code. In ICML . 2091–2100
Miltiadis Allamanis, Hao Peng, and Charles Sutton. 2016 · 2091
Closest in time.