Fetching the paper…
Reading the bibliography…
Neural Language Models of Code, or Neural Code Models (NCMs), are rapidly progressing from research prototypes to commercial developer tools.
S. Hochreiter and J. Schmidhuber, “Long Short-Term Memory,” Neural Computation , vol. 9, no. 8, pp. 1735–1780, 11 1997. [Online]. Available: https://doi.org/10.1162/neco.1997.9.8.1735
1997
Earlier work this paper cites.
——, “Causality: models, reasoning, and inference, by judea pearl, cambridge university press, 2000,” Econometric Theory , vol. 19, no. 4, p. 675–685, 2003
2003
Earlier work this paper cites.
Y. Bengio, R. Ducharme, and P. Vincent, “A neural probabilistic language model,” Advances in Neural Information Processing Systems , vol. 3, pp. 1137–1155, 2003
2003
Earlier work this paper cites.
G. C. Murphy, M. Kersten, and L. Findlater, “How are java software developers using the eclipse ide?” IEEE Softw. , vol. 23, no. 4, p. 76–83, Jul. 2006. [Online]. Available: https://doi.org/10.1109/MS.2006.105
2006
Earlier work this paper cites.
T. Menzies, J. Greenwald, and A. Frank, “Data mining static code attributes to learn defect predictors,” IEEE Transactions on Software Engineering , vol. 33, no. 1, pp. 2–13, 2007
2007
Earlier work this paper cites.
2009
Earlier work this paper cites.
2010
Earlier work this paper cites.
A. Hindle, E. T. Barr et al. , “On the naturalness of software,” in Proceedings of the 34th International Conference on Software Engineering , ser. ICSE ’12. IEEE Press, 2012, p. 837–847
2012
Earlier work this paper cites.
T. Nguyen, A. Nguyen, and H. Nguyen, “A statistical semantic language model for source code,” in ESEC/FSE 2013 , 2013
2013
Earlier work this paper cites.
V. Raychev, M. T. Vechev, and E. Yahav, “Code completion with statistical language models,” Proceedings of the 35th ACM SIGPLAN Conference on Programming Language Design and Implementation , 2014
2014
Earlier work this paper cites.
Z. Tu, Z. Su, and P. Devanbu, “On the localness of software,” in Proceedings of the 22nd ACM SIGSOFT International Symposium on Foundations of Software Engineering , ser. FSE 2014. New York, NY, USA: Association for Computing Machinery, 2014, p. 269–280. [Online]. Available: https://doi-org.proxy.wm.edu/10.1145/2635868.2635875
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
M. White, C. Vendome et al. , “Toward Deep Learning Software Repositories,” in Proceedings of the 12th IEEE Working Conference on Mining Software Repositories (MSR’15) , ser. MSR ’15. Piscataway, NJ, USA: IEEE Press, 2015, pp. 334–345. [Online]. Available: http://dl.acm.org/citation.cfm?id=2820518.2820559
2015
Earlier work this paper cites.
A. T. Nguyen and T. N. Nguyen, “Graph-based statistical language model for code,” in Proceedings of the 37th International Conference on Software Engineering - Volume 1 , ser. ICSE ’15. IEEE Press, 2015, p. 858–868
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Svajlenko and C. K. Roy, “Evaluating clone detection tools with BigCloneBench,” 2015 IEEE 31st International Conference on Software Maintenance and Evolution, ICSME 2015 - Proceedings , pp. 131–140, 2015
2015
Earlier work this paper cites.
M. Abadi, A. Agarwal et al. , “TensorFlow: Large-scale machine learning on heterogeneous systems,” 2015, software available from tensorflow.org. [Online]. Available: https://www.tensorflow.org/
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
J. Pearl, M. Glymour, and N. P.Jewell, Causal Inference in Statistics, A Primer , 2016
2016
Earlier work this paper cites.
M. White, M. Tufano et al. , “Deep learning code fragments for code clone detection,” in 2016 31st IEEE/ACM International Conference on Automated Software Engineering (ASE) , 2016, pp. 87–98
2016
Earlier work this paper cites.
“tree-sitter/tree-sitter-python,” Dec. 2023, original-date: 2016-12-15T06:22:25Z. [Online]. Available: https://github.com/tree-sitter/tree-sitter-python
2016
Earlier work this paper cites.
B. Ray, V. Hellendoorn et al. , “On the ”naturalness” of buggy code,” in Proceedings of the 38th International Conference on Software Engineering , ser. ICSE ’16. New York, NY, USA: Association for Computing Machinery, 2016, p. 428–439. [Online]. Available: https://doi.org/10.1145/2884781.2884848
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer et al. , “Attention is all you need,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , ser. NeurIPs’17. Red Hook, NY, USA: Curran Associates Inc., 2017, p. 6000–6010
2017
Earlier work this paper cites.
V. J. Hellendoorn and P. Devanbu, “Are deep neural networks the best choice for modeling source code?” ESEC/FSE 2017: Proceedings of the 2017 11th Joint Meeting on Foundations of Software Engineering , pp. 763–773, 2017
2017
Earlier work this paper cites.
J. Pearl and E. Bareinboim, “Transportability of Causal and Statistical Relations: A Formal Approach.”
2017
Earlier work this paper cites.
M. Tufano, C. Watson et al. , “Deep learning similarities from different representations of source code,” in 2018 IEEE/ACM 15th International Conference on Mining Software Repositories (MSR) , 2018, pp. 542–553
2018
Earlier work this paper cites.
——, “Considerations for Evaluation and Generalization in Interpretable Machine Learning,” pp. 3–17, 2018
2018
Cited alongside, same era.
J. Pearl and D. Mackenzie, The book of why: The New Science of Cause and Effect , 2018
2018
Cited alongside, same era.
X. Hu, G. Li et al. , “Deep code comment generation,” in Proceedings of the 26th Conference on Program Comprehension , ser. ICPC ’18. New York, NY, USA: Association for Computing Machinery, 2018, p. 200–210. [Online]. Available: https://doi-org.proxy.wm.edu/10.1145/3196321.3196334
2018
Cited alongside, same era.
M. Tufano, C. Watson et al. , “An empirical investigation into learning bug-fixing patches in the wild via neural machine translation,” in Proceedings of the 33rd ACM/IEEE International Conference on Automated Software Engineering , ser. ASE 2018. New York, NY, USA: ACM, 2018, pp. 832–837. [Online]. Available: http://doi.acm.org/10.1145/3238147.3240732
2018
Cited alongside, same era.
Y. Hussain, Z. Huang et al. , “Deep transfer learning for source code modeling,” Int. J. Softw. Eng. Knowl. Eng. , vol. 30, pp. 649–668, 2020
2020
Later among the works it cites.
Z. Feng, D. Guo et al. , “CodeBERT: A pre-trained model for programming and natural languages,” in Findings of the Association for Computational Linguistics: EMNLP 2020 . Online: Association for Computational Linguistics, Nov. 2020, pp. 1536–1547. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.139
2020
Later among the works it cites.
github, “Github,” 2020. [Online]. Available: https://github.com/
2020
Later among the works it cites.
T. Wolf, L. Debut et al. , “Transformers: State-of-the-art natural language processing,” in Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations . Online: Association for Computational Linguistics, Oct. 2020, pp. 38–45. [Online]. Available: https://www.aclweb.org/anthology/2020.emnlp-demos.6
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Allamanis, M. Brockschmidt, and M. Khademi, “Learning to represent programs with graphs,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=BJOFETxR-
2018
Cited alongside, same era.
P. Rajpurkar, R. Jia, and P. Liang, “Know what you don’t know: Unanswerable questions for SQuAD,” in Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) . Melbourne, Australia: Association for Computational Linguistics, Jul. 2018, pp. 784–789. [Online]. Available: https://aclanthology.org/P18-2124
2018
Cited alongside, same era.
U. Khandelwal, H. He et al. , “Sharp nearby, fuzzy far away: How neural language models use context,” in Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Melbourne, Australia: Association for Computational Linguistics, Jul. 2018, pp. 284–294. [Online]. Available: https://aclanthology.org/P18-1027
2018
Cited alongside, same era.
B. Kim, M. Wattenberg et al. , “Interpretability beyond feature attribution: Quantitative Testing with Concept Activation Vectors (TCAV),” 35th International Conference on Machine Learning, ICML 2018 , vol. 6, pp. 4186–4195, 2018
2018
Cited alongside, same era.
M. White, M. Tufano et al. , “Sorting and transforming program repair ingredients via deep learning code similarities,” in 2019 IEEE 26th International Conference on Software Analysis, Evolution and Reengineering (SANER) , 2019, pp. 479–490
2019
Cited alongside, same era.
Z. Chen, S. J. Kommrusch et al. , “Sequencer: Sequence-to-sequence learning for end-to-end program repair,” IEEE Transactions on Software Engineering , pp. 1–1, 2019
2019
Cited alongside, same era.
T. Wu, M. T. Ribeiro et al. , “Errudite: Scalable, reproducible, and testable error analysis,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Florence, Italy: Association for Computational Linguistics, Jul. 2019, pp. 747–763. [Online]. Available: https://aclanthology.org/P19-1073
2019
Cited alongside, same era.
C. Molnar, Interpretable Machine Learning , 2019, https://christophm.github.io/interpretable-ml-book/
2019
Cited alongside, same era.
Later among the works it cites.
R. M. Karampatsis, H. Babii et al. , “Big code != big vocabulary: Open-vocabulary models for source code,” Proceedings - International Conference on Software Engineering , pp. 1073–1085, 2020
2020
Later among the works it cites.
M. Ciniselli, N. Cooper et al. , “An empirical study on the usage of transformer models for code completion,” 2021
2021
Later among the works it cites.
A. Mastropaolo, S. Scalabrino et al. , “Studying the Usage of Text-To-Text Transfer Transformer to Support Code-Related Tasks,” pp. 336–347, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
W. Zaremba, G. Brockman, and OpenAI, “Openai codex,” Aug 2021. [Online]. Available: https://openai.com/blog/openai-codex/
2021
Later among the works it cites.
M. Chen, J. Tworek et al. , “Evaluating large language models trained on code,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
E. M. Bender, T. Gebru et al. , “On the dangers of stochastic parrots: Can language models be too big?” in Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency , ser. FAccT ’21. New York, NY, USA: Association for Computing Machinery, 2021, p. 610–623. [Online]. Available: https://doi-org.proxy.wm.edu/10.1145/3442188.3445922
2021
Later among the works it cites.
D. Guo, S. Ren et al. , “Graphcode{bert}: Pre-training code representations with data flow,” in International Conference on Learning Representations , 2021. [Online]. Available: https://openreview.net/forum?id=jLoC4ez43PZ
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
M. R. I. Rabin, N. D. Bui et al. , “On the generalizability of neural program models with respect to semantic-preserving program transformations,” Information and Software Technology , vol. 135, p. 106552, 2021
2021
Later among the works it cites.
T. Wu, M. T. Ribeiro et al. , “Polyjuice: Generating counterfactuals for explaining, evaluating, and improving models,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Online: Association for Computational Linguistics, Aug. 2021, pp. 6707–6723. [Online]. Available: https://aclanthology.org/2021.acl-long.523
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Sharma, V. Syrgkanis et al. , “DoWhy : Addressing Challenges in Expressing and Validating Causal Assumptions,” 2021
2021
Later among the works it cites.
2022
Later among the works it cites.
J. Cito, I. Dillig et al. , “Counterfactual explanations for models of code,” in Proceedings of the 44th International Conference on Software Engineering: Software Engineering in Practice , ser. ICSE-SEIP ’22. New York, NY, USA: Association for Computing Machinery, 2022, p. 125–134. [Online]. Available: https://doi.org/10.1145/3510457.3513081
2022
Later among the works it cites.
Y. Hu and J. Tian, “Neuron dependency graphs: A causal abstraction of neural networks,” in Proceedings of the 39th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, K. Chaudhuri, S. Jegelka et al. , Eds., vol. 162. PMLR, 17–23 Jul 2022, pp. 9020–9040. [Online]. Available: https://proceedings.mlr.press/v162/hu22b.html
2022
Later among the works it cites.
D. Rodriguez-Cardenas, D. N. Palacio et al. , “Benchmarking causal study to interpret large language models for source code,” in 2023 IEEE International Conference on Software Maintenance and Evolution (ICSME) , 2023, pp. 329–334
2023
Closest in time.
2023
Closest in time.
“WM-SEMERU/CausalSE,” Mar. 2024, original-date: 2024-03-06T23:51:39Z. [Online]. Available: https://github.com/WM-SEMERU/CausalSE
2024
Closest in time.