Fetching the paper…
Reading the bibliography…
Code embedding is a keystone in the application of machine learning on several Software Engineering (SE) tasks.
The Program Dependence Graph and Its Use in Optimization
Jeanne Ferrante, Karl J. Ottenstein, and Joe D. Warren. 1987 · 1987
Earlier work this paper cites.
Optimization of Object-Oriented Programs Using Static Class Hierarchy Analysis. In Proceedings of the 9th European Conference on Object-Oriented Programming (ECOOP ’95) . Springer-Verlag, Berlin, Heidelberg, 77–101
Jeffrey Dean, David Grove, and Craig Chambers. 1995 · 1995
Earlier work this paper cites.
Supporting controlled experimentation with testing techniques: An infrastructure and its potential impact
Hyunsook Do, Sebastian Elbaum, and Gregg Rothermel. 2005 · 2005
Earlier work this paper cites.
Deckard: Scalable and accurate tree-based detection of code clones. In 29th International Conference on Software Engineering (ICSE’07) . IEEE, 96–105
Lingxiao Jiang, Ghassan Misherghi, Zhendong Su, and Stephane Glondu. 2007 · 2007
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini. 2008 · 2008
Earlier work this paper cites.
Why does unsupervised pre-training help deep learning?. In Proceedings of the thirteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 201–208
Dumitru Erhan, Aaron Courville, Yoshua Bengio, and Pascal Vincent. 2010 · 2010
Earlier work this paper cites.
Soot: A Java bytecode optimization framework
Raja Vallée-Rai, Phong Co, Etienne Gagnon, Laurie Hendren, Patrick Lam, and Vijay Sundaresan. 2010 · 2010
Earlier work this paper cites.
Generalization capability of artificial neural network incorporated with pruning method. In International Conference on Advanced Computing, Networking and Security . Springer, 171–178
Siddhaling Urolagin, KV Prema, and NV Subba Reddy. 2011 · 2011
Earlier work this paper cites.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing: System Demonstrations . Association for Computational Linguistics, Brussels, Belgium, 66–71
Taku Kudo and John Richardson. 2018 · 2012
Earlier work this paper cites.
Tomás Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013a · 2013
Earlier work this paper cites.
Defects4J: A database of existing faults to enable controlled testing studies for Java programs. In Proceedings of the 2014 International Symposium on Software Testing and Analysis . 437–440
René Just, Darioush Jalali, and Michael D Ernst. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Adaptive control processes
Richard E Bellman. 2015 · 2015
Earlier work this paper cites.
Gated graph sequence neural networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard Zemel. 2015 · 2015
Earlier work this paper cites.
Going deeper with convolutions. In 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 1–9
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich. 2015 · 2015
Earlier work this paper cites.
Word embedding evaluation and combination. In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC’16) . 300–305
Sahar Ghannay, Benoit Favre, Yannick Esteve, and Nathalie Camelin. 2016 · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N Kipf and Max Welling. 2016a · 2016
Earlier work this paper cites.
Variational graph auto-encoders
Thomas N Kipf and Max Welling. 2016b · 2016
Earlier work this paper cites.
Fractalnet: Ultra-deep neural networks without residuals
Gustav Larsson, Michael Maire, and Gregory Shakhnarovich. 2016 · 2016
Earlier work this paper cites.
Automatic patch generation by learning correct code. In Proceedings of the 43rd Annual ACM SIGPLAN-SIGACT Symposium on Principles of Programming Languages . 298–312
Fan Long and Martin Rinard. 2016 · 2016
Earlier work this paper cites.
Neural Machine Translation of Rare Words with Subword Units. In Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Berlin, Germany, 1715–1725
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2818–2826
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016 · 2016
Earlier work this paper cites.
Deep learning code fragments for code clone detection. In 2016 31st IEEE/ACM International Conference on Automated Software Engineering (ASE) . 87–98
Martin White, Michele Tufano, Christopher Vendome, and Denys Poshyvanyk. 2016 · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Earlier work this paper cites.
Learning to represent programs with graphs
Miltiadis Allamanis, Marc Brockschmidt, and Mahmoud Khademi. 2017 · 2017
Cited alongside, same era.
JaDA–the Java deadlock analyser
Abel Garcia and Cosimo Laneve. 2017 · 2017
Cited alongside, same era.
Source code classification using Neural Networks. In 2017 14th International Joint Conference on Computer Science and Software Engineering (JCSSE) . IEEE, 1–6
Shlok Gilda. 2017 · 2017
Cited alongside, same era.
Neural Message Passing for Quantum Chemistry
Justin Gilmer, Samuel S. Schoenholz, Patrick F. Riley, Oriol Vinyals, and George E. Dahl. 2017 · 2017
Cited alongside, same era.
Inductive representation learning on large graphs. In Proceedings of the 31st International Conference on Neural Information Processing Systems . 1025–1035
William L Hamilton, Rex Ying, and Jure Leskovec. 2017 · 2017
Mutation testing advances: an analysis and survey
Mike Papadakis, Marinos Kintis, Jie Zhang, Yue Jia, Yves Le Traon, and Mark Harman. 2019 · 2019
Later among the works it cites.
XLNet: Generalized Autoregressive Pretraining for Language Understanding. In Advances in Neural Information Processing Systems , H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alché-Buc, E. Fox, and R. Garnett (Eds.), Vol. 32. Curran Associates, Inc
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
A novel neural source code representation based on abstract syntax tree. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 783–794
Jian Zhang, Xu Wang, Hongyu Zhang, Hailong Sun, Kaixuan Wang, and Xudong Liu. 2019 · 2019
Later among the works it cites.
A Multilingual BPE Embedding Space for Universal Sentiment Lexicon Induction. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy, 3506–3517
Mengjie Zhao and Hinrich Schütze. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Densely connected convolutional networks. In Proceedings of the IEEE conference on computer vision and pattern recognition . 4700–4708
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger. 2017 · 2017
Cited alongside, same era.
Petar Veličković, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lio, and Yoshua Bengio. 2017 · 2017
Cited alongside, same era.
Neural code comprehension: A learnable representation of code semantics
Tal Ben-Nun, Alice Shoshana Jakobovits, and Torsten Hoefler. 2018 · 2018
Cited alongside, same era.
What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Melbourne, Australia, 2126–2136
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
BPEmb: Tokenization-free Pre-trained Subword Embeddings in 275 Languages. In Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018) . European Language Resources Association (ELRA), Miyazaki, Japan
Benjamin Heinzerling and Michael Strube. 2018 · 2018
Cited alongside, same era.
Improvements to context based self-supervised learning. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 9339–9348
T Nathan Mundhenk, Daniel Ho, and Barry Y Chen. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Exploring software naturalness through neural language models
Luca Buratti, Saurabh Pujar, Mihaela Bornea, Scott McCarley, Yunhui Zheng, Gaetano Rossiello, Alessandro Morari, Jim Laredo, Veronika Thost, Yufan Zhuang, et al · 2020
Later among the works it cites.
Chris Cummins, Hugh Leather, Zacharias Fisches, Tal Ben-Nun, Torsten Hoefler, and Michael O’Boyle. 2020 · 2020
Later among the works it cites.
Codebert: A pre-trained model for programming and natural languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Later among the works it cites.
Graphcodebert: Pre-training code representations with data flow
Daya Guo, Shuo Ren, Shuai Lu, Zhangyin Feng, Duyu Tang, Shujie Liu, Long Zhou, Nan Duan, Alexey Svyatkovskiy, Shengyu Fu, et al · 2020
Later among the works it cites.
CC2Vec: Distributed Representations of Code Changes. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering (Seoul, South Korea) (ICSE ’20) . Association for Computing Machinery, New York, NY, USA, 518–529
Thong Hoang, Hong Jin Kang, David Lo, and Julia Lawall. 2020 · 2020
Later among the works it cites.
Learning and evaluating contextual embedding of source code. In International Conference on Machine Learning . PMLR, 5110–5121
Aditya Kanade, Petros Maniatis, Gogul Balakrishnan, and Kensen Shi. 2020 · 2020
Later among the works it cites.
Why do Deep Neural Networks with Skip Connections and Concatenated Hidden Representations Work?. In International Conference on Neural Information Processing . Springer, 380–392
Oyebade K Oyedotun and Djamila Aouada. 2020 · 2020
Later among the works it cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
A Primer in BERTology: What We Know About How BERT Works
Anna Rogers, Olga Kovaleva, and Anna Rumshisky. 2020 · 2020
Later among the works it cites.
Ir2vec: Llvm ir based scalable program embeddings
S VenkataKeerthy, Rohit Aggarwal, Shalini Jain, Maunendra Sankar Desarkar, Ramakrishna Upadrasta, and YN Srikant. 2020 · 2020
Later among the works it cites.
Detecting code clones with graph neural network and flow-augmented abstract syntax tree. In 2020 IEEE 27th International Conference on Software Analysis, Evolution and Reengineering (SANER) . IEEE, 261–271
Wenhan Wang, Ge Li, Bo Ma, Xin Xia, and Zhi Jin. 2020a · 2020
Later among the works it cites.
Learning to Represent Programs with Heterogeneous Graphs
Wenhan Wang, Kechi Zhang, Ge Li, and Zhi Jin. 2020c · 2020
Later among the works it cites.
A comprehensive survey on graph neural networks
Zonghan Wu, Shirui Pan, Fengwen Chen, Guodong Long, Chengqi Zhang, and S Yu Philip. 2020 · 2020
Later among the works it cites.
Quantifying the Contextualization of Word Representations with Semantic Class Probing. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, Online, 1219–1234
Mengjie Zhao, Philipp Dufter, Yadollah Yaghoobzadeh, and Hinrich Schütze. 2020 · 2020
Later among the works it cites.
Graph neural networks: A review of methods and applications
Jie Zhou, Ganqu Cui, Shengding Hu, Zhengyan Zhang, Cheng Yang, Zhiyuan Liu, Lifeng Wang, Changcheng Li, and Maosong Sun. 2020 · 2020
Later among the works it cites.
InferCode: Self-Supervised Learning of Code Representations by Predicting Subtrees. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 1186–1197
Nghi DQ Bui, Yijun Yu, and Lingxiao Jiang. 2021 · 2021
Closest in time.
Speech & language processing
Dan Jurafsky and James H. Martin. 2021 · 2021
Closest in time.
How could Neural Networks understand Programs?
Dinglan Peng, Shuxin Zheng, Yatao Li, Guolin Ke, Di He, and Tie-Yan Liu. 2021 · 2021
Closest in time.
Project CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks
Ruchir Puri, David S Kung, Geert Janssen, Wei Zhang, Giacomo Domeniconi, Vladmir Zolotov, Julian Dolby, Jie Chen, Mihir Choudhury, Lindsey Decker, et al · 2021
Closest in time.
Automated classification of overfitting patches with statically extracted code features
He Ye, Jian Gu, Matias Martinez, Thomas Durieux, and Martin Monperrus. 2021 · 2021
Closest in time.