Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have brought a paradigm shift to the field of code generation, offering the potential to enhance the software development process.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
Biometry the principles and practice of statistics in biological research
F Rohlf et al · 1981
Earlier work this paper cites.
A style analysis of C programs
R E Berry and B AE Meekings. 1985 · 1985
Earlier work this paper cites.
A taxonomy for programming style. In Proceedings of the 1990 ACM annual conference on Cooperation . 244–250
Paul W Oman and Curtis R Cook. 1990 · 1990
Earlier work this paper cites.
Open coding
Shahedul Huq Khandkar. 2009 · 2009
Earlier work this paper cites.
Measuring the stylistic inconsistency in software projects using hierarchical agglomerative clustering. In Proceedings of the The 12th International Conference on Predictive Models and Data Analytics in Software Engineering . 1–10
Qing Mi, Jacky Keung, and Yang Yu. 2016 · 2016
Earlier work this paper cites.
Towards a universal code formatter through machine learning. In Proceedings of the 2016 ACM SIGPLAN International Conference on Software Language Engineering . 137–151
Terence Parr and Jurgen Vinju. 2016 · 2016
Earlier work this paper cites.
Python–the fastest growing programming language
KR Srinath. 2017 · 2017
Earlier work this paper cites.
Python coding style compliance on stack overflow. In 2019 IEEE/ACM 16th International Conference on Mining Software Repositories (MSR) . IEEE, 210–214
Nikolaos Bafatakis, Niels Boecker, Wenjie Boon, Martin Cabello Salazar, Jens Krinke, Gazi Oznacar, and Robert White. 2019 · 2019
Earlier work this paper cites.
STYLE-ANALYZER: fixing code style inconsistencies with interpretable unsupervised algorithms. In 2019 IEEE/ACM 16th International Conference on Mining Software Repositories (MSR) . IEEE, 468–478
Vadim Markovtsev, Waren Long, Hugo Mougard, Konstantin Slavnov, and Egor Bulychev. 2019 · 2019
Earlier work this paper cites.
Program synthesis with large language models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, et al · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde De Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Exploring dynamic selection of branch expansion orders for code generation
Hui Jiang, Chulun Zhou, Fandong Meng, Biao Zhang, Jie Zhou, Degen Huang, Qingqiang Wu, and Jinsong Su. 2021 · 2021
Earlier work this paper cites.
OpenAI Code
OpenAI. 2021 · 2021
Earlier work this paper cites.
Generating adversarial source programs using important tokens-based structural transformations. In 2022 26th International Conference on Engineering of Complex Computer Systems (ICECCS) . IEEE, 173–182
Penglong Chen, Zhen Li, Yu Wen, and Lili Liu. 2022 · 2022
Earlier work this paper cites.
Coderl: Mastering code generation through pretrained models and deep reinforcement learning
Hung Le, Yue Wang, Akhilesh Deepak Gotmare, Silvio Savarese, and Steven Chu Hong Hoi. 2022 · 2022
Earlier work this paper cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Earlier work this paper cites.
Exploring the Impact of Code Style in Identifying Good Programmers
Rafed Muhammad Yasir and Dr Ahmedul Kabir. 2022 · 2022
Earlier work this paper cites.
Towards robustness of deep program processing models—detection, estimation, and enhancement
Huangzhao Zhang, Zhiyi Fu, Ge Li, Lei Ma, Zhehao Zhao, Hua’an Yang, Yizhe Sun, Yang Liu, and Zhi Jin. 2022 · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Codeplan: Repository-level coding using llms and planning
Ramakrishna Bairi, Atharv Sonwane, Aditya Kanade, Arun Iyer, Suresh Parthasarathy, Sriram Rajamani, B Ashok, Shashank Shet, et al · 2023
Earlier work this paper cites.
Improving code generation by training with natural language feedback
Angelica Chen, Jérémy Scheurer, Tomasz Korbak, Jon Ander Campos, Jun Shern Chan, Samuel R Bowman, Kyunghyun Cho, and Ethan Perez. 2023 · 2023
Earlier work this paper cites.
DUETCS: Code Style Transfer through Generation and Retrieval. In 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 2362–2373
Binger Chen and Ziawasch Abedjan. 2023 · 2023
Earlier work this paper cites.
Evaluation of chatgpt model for vulnerability detection
Anton Cheshkov, Pavel Zadorozhny, and Rodion Levichev. 2023 · 2023
Earlier work this paper cites.
Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks
Domenico Cotroneo, Cristina Improta, Pietro Liguori, and Roberto Natella. 2023 · 2023
Earlier work this paper cites.
Self-collaboration Code Generation via ChatGPT
Yihong Dong, Xue Jiang, Zhi Jin, and Ge Li. 2023 · 2023
Earlier work this paper cites.
Classeval: A manually-crafted benchmark for evaluating llms on class-level code generation
Xueying Du, Mingwei Liu, Kaixin Wang, Hanlin Wang, Junwei Liu, Yixuan Chen, Jiayi Feng, Chaofeng Sha, Xin Peng, and Yiling Lou. 2023 · 2023
Earlier work this paper cites.
Enhancing Large Language Models in Coding Through Multi-Perspective Self-Consistency
Baizhou Huang, Shuai Lu, Weizhu Chen, Xiaojun Wan, and Nan Duan. 2023 · 2023
Earlier work this paper cites.
LLM-Assisted Code Cleaning For Training Accurate Code Generators
Naman Jain, Tianjun Zhang, Wei-Lin Chiang, Joseph E Gonzalez, Koushik Sen, and Ion Stoica. 2023 · 2023
Earlier work this paper cites.
SelfEvolve: A Code Evolution Framework via Large Language Models
Shuyang Jiang, Yuhao Wang, and Yu Wang. 2023b · 2023
Earlier work this paper cites.
Self-planning code generation with large language model
Xue Jiang, Yihong Dong, Lecheng Wang, Qiwei Shang, and Ge Li. 2023a · 2023
Cited alongside, same era.
Structured Chain-of-Thought Prompting for Code Generation
Jia Li, Ge Li, Yongmin Li, and Zhi Jin. 2023b · 2023
Cited alongside, same era.
Large Language Model-Aware In-Context Learning for Code Generation
Jia Li, Ge Li, Chongyang Tao, Huangzhao Zhang, Fang Liu, and Zhi Jin. 2023d · 2023
Cited alongside, same era.
Skcoder: A sketch-based approach for automatic code generation
Jia Li, Yongmin Li, Ge Li, Zhi Jin, Yiyang Hao, and Xing Hu. 2023c · 2023
Cited alongside, same era.
Replication Package
2024 · 2024
Closest in time.
Automatic semantic augmentation of language model prompts (for code summarization). In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering . 1–13
Toufique Ahmed, Kunal Suresh Pai, Premkumar Devanbu, and Earl Barr. 2024 · 2024
Closest in time.
CoSQA+: Enhancing Code Search Dataset with Matching Code
Jing Gong, Yanghui Wu, Linxi Liang, Zibin Zheng, and Yanlin Wang. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Y Wu, YK Li, et al · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jia Li, Yunfei Zhao, Yongmin Li, Ge Li, and Zhi Jin. 2023f · 2023
Cited alongside, same era.
Starcoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, Jia Li, Jenny Chim, et al · 2023
Cited alongside, same era.
Think Outside the Code: Brainstorming Boosts Large Language Models in Code Generation
Xin-Ye Li, Jiang-Tian Xue, Zheng Xie, and Ming Li. 2023e · 2023
Cited alongside, same era.
Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation. In Advances in Neural Information Processing Systems 36: Annual Conference on Neural Information Processing Systems 2023, NeurIPS 2023, New Orleans, LA, USA, December 10 - 16, 2023 , Alice Oh, Tristan Naumann, Amir Globerson, Kate Saenko, Moritz Hardt, and Sergey Levine (Eds.)
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2023a · 2023
Cited alongside, same era.
CodeGen4Libs: A Two-Stage Approach for Library-Oriented Code Generation. In 38th IEEE/ACM International Conference on Automated Software Engineering, ASE 2023, Luxembourg, September 11-15, 2023 . IEEE, 434–445
Mingwei Liu, Tianyong Yang, Yiling Lou, Xueying Du, Ying Wang, and Xin Peng. 2023b · 2023
Cited alongside, same era.
ClarifyGPT: Empowering LLM-based Code Generation with Intention Clarification
Fangwen Mu, Lin Shi, Song Wang, Zhuohao Yu, Binquan Zhang, Chenxue Wang, Shichao Liu, and Qing Wang. 2023 · 2023
Cited alongside, same era.
Lever: Learning to verify language-to-code generation with execution. In International Conference on Machine Learning . PMLR, 26106–26128
Ansong Ni, Srini Iyer, Dragomir Radev, Veselin Stoyanov, Wen-tau Yih, Sida Wang, and Xi Victoria Lin. 2023 · 2023
Cited alongside, same era.
Sanghak Oh, Kiho Lee, Seonhye Park, Doowon Kim, and Hyoungshick Kim. 2023 · 2023
Cited alongside, same era.
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Y Wu, YK Li, et al · 2024
Closest in time.
Analyzing the performance of large language models on code summarization
Rajarshi Haldar and Julia Hockenmaier. 2024 · 2024
Closest in time.
Tackling Long Code Search with Splitting, Encoding, and Aggregating. In Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) . 15500–15510
Fan Hu, Yanlin Wang, Lun Du, Hongyu Zhang, Dongmei Zhang, and Xirong Li. 2024 · 2024
Closest in time.
KareCoder: A New Knowledge-Enriched Code Generation System. In Proceedings of the 2024 IEEE/ACM 46th International Conference on Software Engineering: Companion Proceedings . 270–271
Tao Huang, Zhihong Sun, Zhi Jin, Ge Li, and Chen Lyu. 2024 · 2024
Closest in time.
Improving Repository-level Code Search with Text Conversion. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 4: Student Research Workshop) . 130–137
Mizuki Kondo, Daisuke Kawahara, and Toshiyuki Kurabayashi. 2024 · 2024
Closest in time.
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
Jia Li, Ge Li, Xuanming Zhang, Yihong Dong, and Zhi Jin. 2024a · 2024
Closest in time.
DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
Jia Li, Ge Li, Yunfei Zhao, Yongmin Li, Huanyu Liu, Hao Zhu, Lecheng Wang, Kaibo Liu, Zheng Fang, Lanshen Wang, et al · 2024
Closest in time.
ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search
Zehan Li, Jianfei Zhang, Chuantao Yin, Yuanxin Ouyang, and Wenge Rong. 2024e · 2024
Closest in time.
STALL+: Boosting LLM-based Repository-level Code Completion with Static Analysis
Junwei Liu, Yixuan Chen, Mingwei Liu, Xin Peng, and Yiling Lou. 2024 · 2024
Closest in time.
Commit Messages in the Age of Large Language Models
Cristina V Lopes, Vanessa I Klotzman, Iris Ma, and Iftekar Ahmed. 2024 · 2024
Closest in time.
COMET: Generating Commit Messages using Delta Graph Context Representation
Abhinav Reddy Mandli, Saurabhsingh Rajput, and Tushar Sharma. 2024 · 2024
Closest in time.
RepoHyper: Better Context Retrieval Is All You Need for Repository-Level Code Completion
Huy N Phan, Hoang N Phan, Tien N Nguyen, and Nghi DQ Bui. 2024 · 2024
Closest in time.
Distilled GPT for source code summarization
Chia-Yi Su and Collin McMillan. 2024 · 2024
Closest in time.
Karl Tamberg and Hayretdin Bahsi. 2024 · 2024
Closest in time.
KADEL: Knowledge-Aware Denoising Learning for Commit Message Generation
Wei Tao, Yucheng Zhou, Yanlin Wang, Hongyu Zhang, Haofen Wang, and Wenqiang Zhang. 2024 · 2024
Closest in time.
Structcoder: Structure-aware transformer for code generation
Sindhu Tipirneni, Ming Zhu, and Chandan K Reddy. 2024 · 2024
Closest in time.
Improving llm code generation with grammar augmentation
Shubham Ugare, Tarun Suresh, Hangoo Kang, Sasa Misailovic, and Gagandeep Singh. 2024 · 2024
Closest in time.
SparseCoder: Identifier-Aware Sparse Transformer for File-Level Code Summarization
Yanlin Wang, Yanxian Huang, Daya Guo, Hongyu Zhang, and Zibin Zheng. 2024a · 2024
Closest in time.
Rlcoder: Reinforcement learning for repository-level code completion
Yanlin Wang, Yanli Wang, Daya Guo, Jiachi Chen, Ruikai Zhang, Yuchi Ma, and Zibin Zheng. 2024c · 2024
Closest in time.
M2CVD: Multi-Model Collaboration for Code Vulnerability Detection
Ziliang Wang, Ge Li, Jia Li, Yingfei Xiong, and Zhi Jin. 2024b · 2024
Closest in time.
Security Vulnerability Detection with Multitask Self-Instructed Fine-Tuning of Large Language Models
Aidan Z. H. Yang, Haoye Tian, He Ye, Ruben Martins, and Claire Le Goues. 2024 · 2024
Closest in time.
Codereval: A benchmark of pragmatic code generation with generative pre-trained models. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–12
Hao Yu, Bo Shen, Dezhi Ran, Jiaxin Zhang, Qi Zhang, Yuchi Ma, Guangtai Liang, Ying Li, Qianxiang Wang, and Tao Xie. 2024 · 2024
Closest in time.
Imam Nur Bani Yusuf and Lingxiao Jiang. 2024 · 2024
Closest in time.
Automatic commit message generation: A critical review and directions for future work
Yuxia Zhang, Zhiqing Qiu, Klaas-Jan Stol, Wenhui Zhu, Jiaxin Zhu, Yingchen Tian, and Hui Liu. 2024 · 2024
Closest in time.
Hot or Cold? Adaptive Temperature Sampling for Code Generation with Large Language Models. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 38. 437–445
Yuqi Zhu, Jia Li, Ge Li, YunFei Zhao, Zhi Jin, and Hong Mei. 2024 · 2024
Closest in time.