Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have recently been widely used for code generation.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Joseph L Fleiss. 1971 · 1971
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting of the Association for Computational Linguistics . 311–318
Kishore Papineni et al · 2002
Earlier work this paper cites.
Reliability in content analysis: Some common misconceptions and recommendations
Klaus Krippendorff. 2004 · 2004
Earlier work this paper cites.
Saliency map
Ernst Niebur. 2007 · 2007
Earlier work this paper cites.
Interrater reliability: the kappa statistic
Mary L McHugh. 2012 · 2012
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2013 · 2013
Earlier work this paper cites.
Extraction of salient sentences from labelled documents
Misha Denil, Alban Demiraj, and Nando De Freitas. 2014 · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks. In Computer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part I 13 . Springer, 818–833
Matthew D Zeiler and Rob Fergus. 2014 · 2014
Earlier work this paper cites.
Guided alignment training for topic-aware neural machine translation
Wenhu Chen, Evgeny Matusov, Shahram Khadivi, and Jan-Thorsten Peter. 2016 · 2016
Earlier work this paper cites.
Joint attention in autonomous driving (JAAD)
Iuliia Kotseruba, Amir Rasouli, and John K Tsotsos. 2016 · 2016
Earlier work this paper cites.
Understanding neural networks through representation erasure
Jiwei Li, Will Monroe, and Dan Jurafsky. 2016 · 2016
Earlier work this paper cites.
" Why should i trust you?" Explaining the predictions of any classifier. In Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining . 1135–1144
Marco Tulio Ribeiro et al · 2016
Earlier work this paper cites.
Learning deep features for discriminative localization. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2921–2929
Bolei Zhou, Aditya Khosla, Agata Lapedriza, Aude Oliva, and Antonio Torralba. 2016 · 2016
Earlier work this paper cites.
Component-based synthesis of table consolidation and transformation tasks from examples
Yu Feng, Ruben Martins, Jacob Van Geffen, Isil Dillig, and Swarat Chaudhuri. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization. In Proceedings of the IEEE international conference on computer vision . 618–626
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Learning important features through propagating activation differences. In International conference on machine learning . PMLR, 3145–3153
Avanti Shrikumar, Peyton Greenside, and Anshul Kundaje. 2017 · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks. In International conference on machine learning . PMLR, 3319–3328
Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017 · 2017
Earlier work this paper cites.
Using human brain activity to guide machine learning
Ruth C Fong, Walter J Scheirer, and David D Cox. 2018 · 2018
Earlier work this paper cites.
Evaluating feature importance estimates
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans, and Been Kim. 2018 · 2018
Earlier work this paper cites.
Biometric recognition through eye movements using a recurrent neural network. In 2018 IEEE International Conference on Big Knowledge (ICBK) . IEEE, 57–64
Shaohua Jia et al · 2018
Earlier work this paper cites.
Content analysis: An introduction to its methodology
Klaus Krippendorff. 2018 · 2018
Earlier work this paper cites.
Nlize: A perturbation-driven visual interrogation tool for analyzing and interpreting natural language inference models
Shusen Liu et al · 2018
Earlier work this paper cites.
Object detection and localization with artificial foveal visual attention. In 2018 Joint IEEE 8th international conference on development and learning and epigenetic robotics (ICDL-EpiRob) . IEEE, 101–106
Cristina Melício et al · 2018
Earlier work this paper cites.
What Does BERT Look at? An Analysis of BERT’s Attention. In Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP . 276–286
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D Manning. 2019 · 2019
Earlier work this paper cites.
Fine-grained analysis of propaganda in news article. In Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP) . Association for Computational Linguistics, 5636–5646
Giovanni Da San Martino et al · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics . 4171–4186
Jacob Devlin et al · 2019
Earlier work this paper cites.
Towards Complex Text-to-SQL in Cross-Domain Database with Intermediate Representation. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . 4524–4535
Jiaqi Guo et al · 2019
Earlier work this paper cites.
Spoc: Search-based pseudocode to code
Sumith Kulal, Panupong Pasupat, Kartik Chandra, Mina Lee, Oded Padon, Alex Aiken, and Percy S Liang. 2019 · 2019
Earlier work this paper cites.
Attention interpretability across nlp tasks
Shikhar Vashishth, Shyam Upadhyay, Gaurav Singh Tomar, and Manaal Faruqui. 2019 · 2019
Earlier work this paper cites.
" Do you trust me?" Increasing user-trust by integrating virtual agents in explainable AI interaction design. In Proceedings of the 19th ACM International Conference on Intelligent Virtual Agents . 7–9
Katharina Weitz et al · 2019
Earlier work this paper cites.
Adversarial robustness for code. In International Conference on Machine Learning . PMLR, 896–907
Pavol Bielik and Martin Vechev. 2020 · 2020
Earlier work this paper cites.
CodeBERT: A Pre-Trained Model for Programming and Natural Language Processing. In Proceedings of the 28th ACM International Conference on Information and Knowledge Management . 307–316
Zhangyin Feng et al · 2020
Earlier work this paper cites.
Attention in natural language processing
Andrea Galassi, Marco Lippi, and Paolo Torroni. 2020 · 2020
Cited alongside, same era.
Graphcodebert: Pre-training code representations with data flow
Daya Guo et al · 2020
Cited alongside, same era.
Aligning AI with Shared Human Values
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song, and Jacob Steinhardt. 2020 · 2020
Cited alongside, same era.
An Empirical Study on the Usage of Transformer Models for Code Completion. In 2020 IEEE International Conference on Software Maintenance and Evolution (ICSME) . IEEE, 408–418
Xiaodong Liu, Ying Xia, and David Lo. 2020 · 2020
Cited alongside, same era.
Interpretable machine learning
Christoph Molnar. 2020 · 2020
Cited alongside, same era.
Taking Flight with Copilot: Early insights and opportunities of AI-powered pair-programming tools
Christian Bird et al · 2022
Later among the works it cites.
Shared interest: Measuring human-ai alignment to identify recurring patterns in model behavior. In Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems . 1–17
Angie Boggust, Benjamin Hoover, Arvind Satyanarayan, and Hendrik Strobelt. 2022 · 2022
Later among the works it cites.
NatGen: generative pre-training by “naturalizing” source code. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 18–30
Saikat Chakraborty et al · 2022
Later among the works it cites.
Codet: Code generation with generated tests
Bei Chen et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning to search for objects in images from human gaze sequences. In Image Analysis and Recognition: 17th International Conference . Springer, 280–292
Afonso Nunes, Rui Figueiredo, and Plinio Moreno. 2020 · 2020
Cited alongside, same era.
Codebleu: a method for automatic evaluation of code synthesis
Shuo Ren et al · 2020
Cited alongside, same era.
Generating Adversarial Computer Programs using Optimized Obfuscations. In International Conference on Learning Representations
Shashank Srikant et al · 2020
Cited alongside, same era.
RAT-SQL: Relation-Aware Schema Encoding and Linking for Text-to-SQL Parsers. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 7567–7578
Bailin Wang et al · 2020
Cited alongside, same era.
Perturbed Masking: Parameter-free Probing for Analyzing and Interpreting BERT. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 4166–4176
Zhiyong Wu, Yun Chen, Ben Kao, and Qun Liu. 2020 · 2020
Cited alongside, same era.
Adversarial examples for models of code
Noam Yefet, Uri Alon, and Eran Yahav. 2020 · 2020
Cited alongside, same era.
RECPARSER: A Recursive Semantic Parsing Framework for Text-to-SQL Task.. In IJCAI . 3644–3650
Yu Zeng et al · 2020
Cited alongside, same era.
To what extent do deep learning-based code recommenders generate predictions by cloning code from the training set?. In Proceedings of the 19th International Conference on Mining Software Repositories . 167–178
Matteo Ciniselli, Luca Pascarella, and Gabriele Bavota. 2022 · 2022
Later among the works it cites.
Aligning eyes between humans and deep neural network through interactive attention alignment
Yuyang Gao et al · 2022
Later among the works it cites.
Transformer feed-forward layers build predictions by promoting concepts in the vocabulary space
Mor Geva, Avi Caciularu, Kevin Ro Wang, and Yoav Goldberg. 2022 · 2022
Later among the works it cites.
Importance is in your attention: agent importance prediction for autonomous driving. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2532–2535
Christopher Hazard et al · 2022
Later among the works it cites.
Can NMT understand me? towards perturbation-based evaluation of NMT models for code generation. In 2022 IEEE/ACM 1st International Workshop on Natural Language-Based Software Engineering (NLBSE) . IEEE, 59–66
Pietro Liguori et al · 2022
Later among the works it cites.
Asleep at the keyboard? assessing the security of github copilot’s code contributions. In 2022 IEEE Symposium on Security and Privacy (SP) . IEEE, 754–768
Hammond Pearce et al · 2022
Later among the works it cites.
Incorporating domain knowledge through task augmentation for front-end JavaScript code generation. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1533–1543
Sijie Shen, Xiang Zhu, Yihong Dong, Qizhi Guo, Yankun Zhen, and Ge Li. 2022 · 2022
Later among the works it cites.
An Empirical Study of Code Smells in Transformer-based Code Generation Techniques. In 2022 IEEE 22nd International Working Conference on Source Code Analysis and Manipulation (SCAM) . IEEE, 71–82
Mohammed Latif Siddiq et al · 2022
Later among the works it cites.
Thirdeye: Attention maps for safe autonomous driving systems. In Proceedings of the 37th IEEE/ACM International Conference on Automated Software Engineering . 1–12
Andrea Stocco et al · 2022
Later among the works it cites.
Natural Language Processing with Transformers: Building Language Applications with Hugging Face
Lewis Tunstall, Leandro von Werra, and Thomas Wolf. 2022 · 2022
Later among the works it cites.
Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models. In CHI conference on human factors in computing systems extended abstracts . 1–7
Priyan Vaithilingam et al · 2022
Later among the works it cites.
What do they capture? a structural analysis of pre-trained language models for source code. In Proceedings of the 44th International Conference on Software Engineering . 2377–2388
Yao Wan, Wei Zhao, Hongyu Zhang, Yulei Sui, Guandong Xu, and Hai Jin. 2022 · 2022
Later among the works it cites.
A systematic evaluation of large language models of code. In Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming . 1–10
Frank F Xu, Uri Alon, Graham Neubig, and Vincent Josua Hellendoorn. 2022 · 2022
Later among the works it cites.
Assessing the quality of GitHub copilot’s code generation. In Proceedings of the 18th International Conference on Predictive Models and Data Analytics in Software Engineering . 62–71
Burak Yetistiren, Isik Ozsoy, and Eray Tuzun. 2022 · 2022
Later among the works it cites.
CERT: Continual Pre-training on Sketches for Library-oriented Code Generation. In Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence, IJCAI 2022, Vienna, Austria, 23-29 July 2022 , Luc De Raedt (Ed.). ijcai.org, 2369–2375
Daoguang Zan, Bei Chen, et al · 2022
Later among the works it cites.
An extensive study on pre-trained models for program understanding and generation. In Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis . 39–51
Zhengran Zeng, Hanzhuo Tan, Haotian Zhang, Jing Li, Yuqun Zhang, and Lingming Zhang. 2022 · 2022
Later among the works it cites.
What does Transformer learn about source code?
Kechi Zhang, Ge Li, and Zhi Jin. 2022a · 2022
Later among the works it cites.
GPT-4 Parameters: Unlimited guide NLP’s Game-Changer
2023 · 2023
Closest in time.
Towards Modeling Human Attention from Eye Movements for Neural Source Code Summarization
Aakash Bansal, Bonita Sharif, and Collin McMillan. 2023 · 2023
Closest in time.
Grounded copilot: How programmers interact with code-generating models
Shraddha Barke, Michael B James, and Nadia Polikarpova. 2023 · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck et al · 2023
Closest in time.
Giuseppe Destefanis, Silvia Bartolucci, and Marco Ortu. 2023 · 2023
Closest in time.
InCoder: A Generative Model for Code Infilling and Synthesis. In The Eleventh International Conference on Learning Representations
Daniel Fried et al · 2023
Closest in time.
SkCoder: A Sketch-based Approach for Automatic Code Generation
Jia Li, Yongmin Li, Ge Li, Zhi Jin, Yiyang Hao, and Xing Hu. 2023 · 2023
Closest in time.
On the Reliability and Explainability of Automated Code Generation Approaches
Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, and Li Li. 2023 · 2023
Closest in time.
CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis. In The Eleventh International Conference on Learning Representations
Erik Nijkamp, Bo Pang, et al · 2023
Closest in time.
ToxiSpanSE: An Explainable Toxicity Detection in Code Review Comments. In 2023 ACM/IEEE International Symposium on Empirical Software Engineering and Measurement (ESEM) . IEEE, 1–12
Jaydeb Sarker, Sayma Sultana, Steven R Wilson, and Amiangshu Bosu. 2023 · 2023
Closest in time.
On Robustness of Prompt-based Semantic Parsing with Large Pre-trained Language Model: An Empirical Study on Codex. In Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics . 1090–1102
Terry Yue Zhuo, Zhuang Li, Yujin Huang, Fatemeh Shiri, Weiqing Wang, Gholamreza Haffari, and Yuan-Fang Li. 2023 · 2023
Closest in time.
Attention-Alignment-Empirical-Study
Bonan Kou. 2024 · 2024
Closest in time.