Fetching the paper…
Reading the bibliography…
Statistical language modeling techniques have successfully been applied to source code, yielding a variety of new software development tools, such as tools for code suggestion and improving readability.
CACHECA: A Cache Language Model Based Code Suggestion Tool.. In ICSE (2) , Antonia Bertolino, Gerardo Canfora, and Sebastian G. Elbaum (Eds.). IEEE Computer Society, 705–708
Christine Franks, Zhaopeng Tu, Premkumar T. Devanbu, and Vincent Hellendoorn. 2015 · 1934
Earlier work this paper cites.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
A New Algorithm for Data Compression
Philip Gage. 1994 · 1994
Earlier work this paper cites.
An Empirical Study of Smoothing Techniques for Language Modeling. In Proceedings of the 34th Annual Meeting on Association for Computational Linguistics (ACL ’96) . Association for Computational Linguistics, Stroudsburg, PA, USA, 310–318
Stanley F. Chen and Joshua Goodman. 1996 · 1996
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Speech and language processing
D. Jurafsky and J.H. Martin. 2000 · 2000
Earlier work this paper cites.
Modelling Out-of-vocabulary Words for Robust Speech Recognition
Issam Bazzi. 2002 · 2002
Earlier work this paper cites.
Contributing to Eclipse: Principles, Patterns, and Plugins
Erich Gamma and Kent Beck. 2003 · 2003
Earlier work this paper cites.
Learning from Examples to Improve Code Completion Systems. In Proceedings of the the 7th Joint Meeting of the European Software Engineering Conference and the ACM SIGSOFT Symposium on The Foundations of Software Engineering (ESEC/FSE ’09) . ACM, New York, NY, USA, 213–222
Marcel Bruch, Martin Monperrus, and Mira Mezini. 2009 · 2009
Earlier work this paper cites.
Mining source code to automatically split identifiers for software analysis
Eric Enslen, Emily Hill, Lori Pollock, and K Vijay-Shanker. 2009 · 2009
Earlier work this paper cites.
A Study of the Uniqueness of Source Code. In Proceedings of the Eighteenth ACM SIGSOFT International Symposium on Foundations of Software Engineering (FSE ’10) . ACM, New York, NY, USA, 147–156
Mark Gabel and Zhendong Su. 2010 · 2010
Earlier work this paper cites.
Recurrent neural network based language model.. In INTERSPEECH , Takao Kobayashi, Keikichi Hirose, and Satoshi Nakamura (Eds.). ISCA, 1045–1048
Tomas Mikolov, Martin Karafiát, Lukás Burget, Jan Cernocký, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Generating Text with Recurrent Neural Networks. In Proceedings of the 28th International Conference on International Conference on Machine Learning (ICML’11) . Omnipress, USA, 1017–1024
Ilya Sutskever, James Martens, and Geoffrey Hinton. 2011 · 2011
Earlier work this paper cites.
Noise-contrastive Estimation of Unnormalized Statistical Models, with Applications to Natural Image Statistics
Michael U. Gutmann and Aapo Hyvärinen. 2012 · 2012
Earlier work this paper cites.
On the Naturalness of Software. In Proceedings of the 34th International Conference on Software Engineering (ICSE ’12) . IEEE Press, Piscataway, NJ, USA, 837–847
Abram Hindle, Earl T. Barr, Zhendong Su, Mark Gabel, and Premkumar Devanbu. 2012 · 2012
Earlier work this paper cites.
SUBWORD LANGUAGE MODELING WITH NEURAL NETWORKS
Tomáš Mikolov, Ilya Sutskever, Anoop Deoras, Le Hai Son, Stefan Kombrink, and Jaň Cernock. 2012 · 2012
Earlier work this paper cites.
Mining Source Code Repositories at Massive Scale Using Language Modeling. In Proceedings of the 10th Working Conference on Mining Software Repositories (MSR ’13) . IEEE Press, Piscataway, NJ, USA, 207–216
Miltiadis Allamanis and Charles Sutton. 2013 · 2013
Earlier work this paper cites.
One Billion Word Benchmark for Measuring Progress in Statistical Language Modeling
Ciprian Chelba, Tomas Mikolov, Mike Schuster, Qi Ge, Thorsten Brants, and Phillipp Koehn. 2013 · 2013
Earlier work this paper cites.
Statistical NLP for computer program source code: An information theoretic perspective on programming lanuguage verbosity
Sergey Dudoladov. 2013 · 2013
Earlier work this paper cites.
Training and Analyzing Deep Recurrent Neural Networks. In Proceedings of the 26th International Conference on Neural Information Processing Systems - Volume 1 (NIPS’13) . Curran Associates Inc., USA, 190–198
Michiel Hermans and Benjamin Schrauwen. 2013 · 2013
Earlier work this paper cites.
Better Word Representations with Recursive Neural Networks for Morphology. In Proceedings of the Seventeenth Conference on Computational Natural Language Learning . Association for Computational Linguistics, 104–113
Thang Luong, Richard Socher, and Christopher Manning. 2013 · 2013
Earlier work this paper cites.
Distributed Representations of Words and Phrases and Their Compositionality. In Proceedings of the 26th International Conference on Neural Information Processing Systems - Volume 2 (NIPS’13) . Curran Associates Inc., USA, 3111–3119
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Lexical Statistical Machine Translation for Language Migration. In Proceedings of the 2013 9th Joint Meeting on Foundations of Software Engineering (ESEC/FSE 2013) . ACM, New York, NY, USA, 651–654
Anh Tuan Nguyen, Tung Thanh Nguyen, and Tien N. Nguyen. 2013a · 2013
Earlier work this paper cites.
A Statistical Semantic Language Model for Source Code. In Proceedings of the 2013 9th Joint Meeting on Foundations of Software Engineering (ESEC/FSE 2013) . ACM, New York, NY, USA, 532–542
Tung Thanh Nguyen, Anh Tuan Nguyen, Hoan Anh Nguyen, and Tien N. Nguyen. 2013b · 2013
Earlier work this paper cites.
Learning Natural Coding Conventions. In Proceedings of the 22Nd ACM SIGSOFT International Symposium on Foundations of Software Engineering (FSE 2014) . ACM, New York, NY, USA, 281–293
Miltiadis Allamanis, Earl T. Barr, Christian Bird, and Charles Sutton. 2014 · 2014
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Syntax Errors Just Aren’t Natural: Improving Error Reporting with Language Models. In Proceedings of the 11th Working Conference on Mining Software Repositories (MSR 2014) . ACM, New York, NY, USA, 252–261
Joshua Charles Campbell, Abram Hindle, and José Nelson Amaral. 2014 · 2014
Cited alongside, same era.
On the Properties of Neural Machine Translation: Encoder-Decoder Approaches
KyungHyun Cho, Bart van Merrienboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014a · 2014
Cited alongside, same era.
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014b · 2014
Cited alongside, same era.
Program Synthesis using Natural Language
Aditya Desai, Sumit Gulwani, Vineet Hingorani, Nidhi Jain, Amey Karkare, Mark Marron, Sailesh R, and Subhajit Roy. 2016 · 2016
Later among the works it cites.
Efficient softmax approximation for GPUs
Edouard Grave, Armand Joulin, Moustapha Cissé, David Grangier, and Hervé Jégou. 2016 · 2016
Later among the works it cites.
T2API: Synthesizing API Code Usage Templates from English Texts with Statistical Translation. In Proceedings of the 2016 24th ACM SIGSOFT International Symposium on Foundations of Software Engineering (FSE 2016) . ACM, New York, NY, USA, 1013–1017
Thanh Nguyen, Peter C. Rigby, Anh Tuan Nguyen, Mark Karanfil, and Tien N. Nguyen. 2016 · 2016
Later among the works it cites.
SWIM: Synthesizing What I Mean: Code Search and Idiomatic Snippet Synthesis. In Proceedings of the 38th International Conference on Software Engineering (ICSE ’16) . ACM, New York, NY, USA, 357–367
Mukund Raghothaman, Yi Wei, and Youssef Hamadi. 2016 · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Junyoung Chung, Çaglar Gülçehre, KyungHyun Cho, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
An empirical study of identifier splitting techniques
Emily Hill, David Binkley, Dawn Lawrie, Lori Pollock, and K Vijay-Shanker. 2014 · 2014
Cited alongside, same era.
Phrase-Based Statistical Translation of Programming Languages. In Proceedings of the 2014 ACM International Symposium on New Ideas, New Paradigms, and Reflections on Programming & Software (Onward! 2014) . ACM, New York, NY, USA, 173–184
Svetoslav Karaivanov, Veselin Raychev, and Martin Vechev. 2014 · 2014
Cited alongside, same era.
Source Code-Based Recommendation Systems
Kim Mens and Angela Lozano. 2014 · 2014
Cited alongside, same era.
Code Completion with Statistical Language Models. In Proceedings of the 35th ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI ’14) . ACM, New York, NY, USA, 419–428
Veselin Raychev, Martin Vechev, and Eran Yahav. 2014 · 2014
Cited alongside, same era.
On the localness of software. In Proceedings of the 22nd Symposium on the Foundations of Software Engineering . ACM, 269–280
Zhaopeng Tu, Zhendong Su, and Premkumar T. Devanbu. 2014 · 2014
Cited alongside, same era.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dan Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2015 · 2015
Cited alongside, same era.
Suggesting Accurate Method and Class Names. In Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering (ESEC/FSE 2015) . ACM, New York, NY, USA, 38–49
Miltiadis Allamanis, Earl T. Barr, Christian Bird, and Charles Sutton. 2015a · 2015
Cited alongside, same era.
Later among the works it cites.
On the "Naturalness" of Buggy Code. In Proceedings of the 38th International Conference on Software Engineering (ICSE ’16) . ACM, New York, NY, USA, 428–439
Baishakhi Ray, Vincent Hellendoorn, Saheel Godhane, Zhaopeng Tu, Alberto Bacchelli, and Premkumar Devanbu. 2016 · 2016
Later among the works it cites.
Deep Learning Code Fragments for Code Clone Detection. In Proceedings of the 31st IEEE/ACM International Conference on Automated Software Engineering (ASE 2016) . ACM, New York, NY, USA, 87–98
Martin White, Michele Tufano, Christopher Vendome, and Denys Poshyvanyk. 2016 · 2016
Later among the works it cites.
RobustFill: Neural Program Learning under Noisy I/O
Jacob Devlin, Jonathan Uesato, Surya Bhupatiraju, Rishabh Singh, Abdel-rahman Mohamed, and Pushmeet Kohli. 2017 · 2017
Later among the works it cites.
Program synthesis
Sumit Gulwani, Oleksandr Polozov, Rishabh Singh, et al · 2017
Later among the works it cites.
DeepFix: Fixing common C language errors by deep learning. In National Conference on Artificial Intelligence (AAAI)
Rahul Gupta, Soham Pal, Aditya Kanade, and Shirish Shevade. 2017 · 2017
Later among the works it cites.
Are Deep Neural Networks the Best Choice for Modeling Source Code?. In Proceedings of the 2017 11th Joint Meeting on Foundations of Software Engineering (ESEC/FSE 2017) . ACM, New York, NY, USA, 763–773
Vincent J. Hellendoorn and Premkumar Devanbu. 2017 · 2017
Later among the works it cites.
Learning to create and reuse words in open-vocabulary neural language modeling
Kazuya Kawakami, Chris Dyer, and Phil Blunsom. 2017 · 2017
Later among the works it cites.
On the State of the Art of Evaluation in Neural Language Models
G. Melis, C. Dyer, and P. Blunsom. 2017 · 2017
Later among the works it cites.
Attention is all you need. In Advances in Neural Information Processing Systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
A Survey of Machine Learning for Big Code and Naturalness
Miltiadis Allamanis, Earl T. Barr, Premkumar Devanbu, and Charles Sutton. 2018 · 2018
Later among the works it cites.
Context2Name: A Deep Learning-Based Approach to Infer Natural Variable Names from Usage Contexts
Rohan Bavishi, Michael Pradel, and Koushik Sen. 2018 · 2018
Later among the works it cites.
Leveraging Grammar and Reinforcement Learning for Neural Program Synthesis
Rudy Bunel, Matthew J. Hausknecht, Jacob Devlin, Rishabh Singh, and Pushmeet Kohli. 2018 · 2018
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
Deep Code Comment Generation. In Proceedings of the 26th Conference on Program Comprehension (ICPC ’18) . ACM, New York, NY, USA, 200–210
Xing Hu, Ge Li, Xin Xia, David Lo, and Zhi Jin. 2018 · 2018
Later among the works it cites.
Code Completion with Neural Attention and Pointer Networks. In Proceedings of the Twenty-Seventh International Joint Conference on Artificial Intelligence, IJCAI-18 . International Joint Conferences on Artificial Intelligence Organization, 4159–4165
Jian Li, Yue Wang, Michael R. Lyu, and Irwin King. 2018 · 2018
Later among the works it cites.
Spell Once, Summon Anywhere: A Two-Level Open-Vocabulary Language Model
Sebastian J. Mielke and Jason Eisner. 2018 · 2018
Later among the works it cites.
Deep Contextualized Word Representations. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) . Association for Computational Linguistics, 2227–2237
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
DeepBugs: A Learning Approach to Name-based Bug Detection
Michael Pradel and Koushik Sen. 2018 · 2018
Later among the works it cites.
From Characters to Words to in Between: Do We Capture Morphology?. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, 2016–2027
Clara Vania and Adam Lopez. 2017 · 2027
Closest in time.
Summarizing Source Code using a Neural Attention Model. In Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Berlin, Germany, 2073–2083
Srinivasan Iyer, Ioannis Konstas, Alvin Cheung, and Luke Zettlemoyer. 2016 · 2083
Closest in time.