Fetching the paper…
Reading the bibliography…
Software systems have been evolving rapidly and inevitably introducing bugs at an increasing rate, leading to significant losses in resources consumed by software maintenance.
Ropgen: Towards robust code authorship attribution via automatic coding style transformation. In Proceedings of the 44th International Conference on Software Engineering . 1906–1918
Zhen Li, Guenevere Chen, Chen Chen, Yayi Zou, and Shouhuai Xu. 2022a · 1918
Earlier work this paper cites.
Individual comparisons by ranking methods
Frank Wilcoxon. 1992 · 1992
Earlier work this paper cites.
Dominance statistics: Ordinal analyses to answer ordinal questions
Norman Cliff. 1993 · 1993
Earlier work this paper cites.
A comparison of document clustering techniques. In KDD Workshop on Text Mining, 2000
Steinbach Michael. 2000 · 2000
Earlier work this paper cites.
Maximum likelihood estimation of Dirichlet distribution parameters
Jonathan Huang. 2005 · 2005
Earlier work this paper cites.
How long will it take to fix this bug?. In fourth international workshop on mining software repositories (MSR’07: ICSE Workshops 2007) . IEEE, 1–1
Cathrin Weiss, Rahul Premraj, Thomas Zimmermann, and Andreas Zeller. 2007 · 2007
Earlier work this paper cites.
Reversible debugging software
Tom Britton, Lisa Jeng, Graham Carver, Paul Cheak, and Tomer Katzenellenbogen. 2013 · 2013
Earlier work this paper cites.
Defects4J: A database of existing faults to enable controlled testing studies for Java programs. In Proceedings of the 2014 international symposium on software testing and analysis . 437–440
René Just, Darioush Jalali, and Michael D Ernst. 2014 · 2014
Earlier work this paper cites.
De-anonymizing programmers via code stylometry. In 24th USENIX security symposium (USENIX Security 15) . 255–270
Aylin Caliskan-Islam, Richard Harang, Andrew Liu, Arvind Narayanan, Clare Voss, Fabian Yamaguchi, and Rachel Greenstadt. 2015 · 2015
Earlier work this paper cites.
The ManyBugs and IntroClass benchmarks for automated repair of C programs
Claire Le Goues, Neal Holtschulte, Edward K Smith, Yuriy Brun, Premkumar Devanbu, Stephanie Forrest, and Westley Weimer. 2015 · 2015
Earlier work this paper cites.
Practitioners’ expectations on automated fault localization. In Proceedings of the 25th international symposium on software testing and analysis . 165–176
Pavneet Singh Kochhar, Xin Xia, David Lo, and Shanping Li. 2016 · 2016
Earlier work this paper cites.
Automatic clustering of code changes. In Proceedings of the 13th International Conference on Mining Software Repositories . 61–72
Patrick Kreutzer, Georg Dotzler, Matthias Ring, Bjoern M Eskofier, and Michael Philippsen. 2016 · 2016
Earlier work this paper cites.
Evaluating code complexity triggers, use of complexity measures and the influence of code complexity on maintenance time
Vard Antinyan, Miroslaw Staron, and Anna Sandberg. 2017 · 2017
Earlier work this paper cites.
Communication-efficient learning of deep networks from decentralized data. In Artificial intelligence and statistics . PMLR, 1273–1282
Brendan McMahan, Eider Moore, Daniel Ramage, Seth Hampson, and Blaise Aguera y Arcas. 2017 · 2017
Earlier work this paper cites.
user2code2vec: Embeddings for profiling students based on distributional representations of source code. In Proceedings of the 9th International Conference on Learning Analytics & Knowledge . 86–95
David Azcona, Piyush Arora, I-Han Hsiao, and Alan Smeaton. 2019 · 2019
Earlier work this paper cites.
University of cambridge study: Failure to adopt reverse debugging costs global economy $41 billion annually
CO Boulder. 2019 · 2019
Earlier work this paper cites.
On reliability of patch correctness assessment. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 524–535
Xuan-Bach D Le, Lingfeng Bao, David Lo, Xin Xia, Shanping Li, and Corina Pasareanu. 2019 · 2019
Earlier work this paper cites.
Decoupled Weight Decay Regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Earlier work this paper cites.
Borda count in collective decision making: a summary of recent results. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 9830–9836
Jörg Rothe. 2019 · 2019
Earlier work this paper cites.
An empirical study on learning bug-fixing patches in the wild via neural machine translation
Michele Tufano, Cody Watson, Gabriele Bavota, Massimiliano Di Penta, Martin White, and Denys Poshyvanyk. 2019 · 2019
Earlier work this paper cites.
Replicating novices’ struggles with coding style. In 2019 IEEE/ACM 27th International Conference on Program Comprehension (ICPC) . IEEE, 13–18
Eliane S Wiese, Anna N Rafferty, Daniel M Kopta, and Jacqulyn M Anderson. 2019 · 2019
Earlier work this paper cites.
Hybridalpha: An efficient approach for privacy-preserving federated learning. In Proceedings of the 12th ACM workshop on artificial intelligence and security . 13–23
Runhua Xu, Nathalie Baracaldo, Yi Zhou, Ali Anwar, and Heiko Ludwig. 2019 · 2019
Earlier work this paper cites.
Codebert: A pre-trained model for programming and natural languages
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, et al · 2020
Earlier work this paper cites.
The non-iid data quagmire of decentralized machine learning. In International Conference on Machine Learning . PMLR, 4387–4398
Kevin Hsieh, Amar Phanishayee, Onur Mutlu, and Phillip Gibbons. 2020 · 2020
Earlier work this paper cites.
Building implicit vector representations of individual coding style. In Proceedings of the IEEE/ACM 42nd International Conference on Software Engineering Workshops . 117–124
Vladimir Kovalenko, Egor Bogomolov, Timofey Bryksin, and Alberto Bacchelli. 2020 · 2020
Earlier work this paper cites.
Federated optimization in heterogeneous networks
Tian Li, Anit Kumar Sahu, Manzil Zaheer, Maziar Sanjabi, Ameet Talwalkar, and Virginia Smith. 2020 · 2020
Earlier work this paper cites.
Large-scale machine learning systems in real-world industrial settings: A review of challenges and solutions
Lucy Ellen Lwakatare, Aiswarya Raj, Ivica Crnkovic, Jan Bosch, and Helena Holmström Olsson. 2020 · 2020
Earlier work this paper cites.
Personalized federated learning with moreau envelopes
Canh T Dinh, Nguyen Tran, and Josh Nguyen. 2020 · 2020
Earlier work this paper cites.
Automated patch correctness assessment: How far are we?. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 968–980
Shangwen Wang, Ming Wen, Bo Lin, Hongjun Wu, Yihao Qin, Deqing Zou, Xiaoguang Mao, and Hai Jin. 2020 · 2020
Earlier work this paper cites.
Program synthesis with large language models
Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, et al · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde De Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Making Pre-trained Language Models Better Few-shot Learners. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , Chengqing Zong, Fei Xia, Wenjie Li, and Roberto Navigli (Eds.). Association for Computational Linguistics, Online, 3816–3830
Tianyu Gao, Adam Fisch, and Danqi Chen. 2021 · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Earlier work this paper cites.
Cure: Code-aware neural machine translation for automatic program repair. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 1161–1173
Nan Jiang, Thibaud Lutellier, and Lin Tan. 2021 · 2021
Earlier work this paper cites.
Advances and open problems in federated learning
Peter Kairouz, H Brendan McMahan, Brendan Avent, Aurélien Bellet, Mehdi Bennis, Arjun Nitin Bhagoji, Kallista Bonawitz, Zachary Charles, Graham Cormode, Rachel Cummings, et al · 2021
Earlier work this paper cites.
Fedbn: Federated learning on non-iid features via local batch normalization
Xiaoxiao Li, Meirui Jiang, Xiaofei Zhang, Michael Kamp, and Qi Dou. 2021b · 2021
Earlier work this paper cites.
Fednlp: Benchmarking federated learning methods for natural language processing tasks
Bill Yuchen Lin, Chaoyang He, Zihang Zeng, Hulin Wang, Yufen Huang, Christophe Dupuy, Rahul Gupta, Mahdi Soltanolkotabi, Xiang Ren, and Salman Avestimehr. 2021 · 2021
Earlier work this paper cites.
No fear of heterogeneity: Classifier calibration for federated learning with non-iid data
Mi Luo, Fei Chen, Dapeng Hu, Yifan Zhang, Jian Liang, and Jiashi Feng. 2021 · 2021
Cited alongside, same era.
Applying codebert for automated program repair of java simple bugs. In 2021 IEEE/ACM 18th International Conference on Mining Software Repositories (MSR) . IEEE, 505–509
Ehsan Mashhadi and Hadi Hemmati. 2021 · 2021
Cited alongside, same era.
Program comprehension and code complexity metrics: An fmri study. In 2021 IEEE/ACM 43rd International Conference on Software Engineering (ICSE) . IEEE, 524–536
Norman Peitek, Sven Apel, Chris Parnin, André Brechmann, and Janet Siegmund. 2021 · 2021
Cited alongside, same era.
Adaptive Federated Optimization. In International Conference on Learning Representations
Sashank J. Reddi, Zachary Charles, Manzil Zaheer, Zachary Garrett, Keith Rush, Jakub Konečný, Sanjiv Kumar, and Hugh Brendan McMahan. 2021 · 2021
Cited alongside, same era.
Personalized federated learning with contextualized generalization
FedID: Federated Interactive Distillation for Large-Scale Pretraining Language Models. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 8566–8577
Xinge Ma, Jiangming Liu, Jin Wang, and Xuejie Zhang. 2023 · 2023
Later among the works it cites.
Enhancing automated program repair through fine-tuning and prompt engineering
Rishov Paul, Md Mohib Hossain, Mohammed Latif Siddiq, Masum Hasan, Anindya Iqbal, and Joanna Santos. 2023 · 2023
Later among the works it cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Romain Sauvestre, Tal Remez, et al · 2023
Later among the works it cites.
An empirical evaluation of using large language models for automated unit test generation
Max Schäfer, Sarah Nadi, Aryaz Eghbali, and Frank Tip. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xueyang Tang, Song Guo, and Jingcai Guo. 2021 · 2021
Cited alongside, same era.
Automated classification of overfitting patches with statically extracted code features
He Ye, Jian Gu, Matias Martinez, Thomas Durieux, and Martin Monperrus. 2021 · 2021
Cited alongside, same era.
Multilingual training for software engineering. In Proceedings of the 44th International Conference on Software Engineering . 1443–1455
Toufique Ahmed and Premkumar Devanbu. 2022 · 2022
Cited alongside, same era.
Improving generalization in federated learning by seeking flat minima. In European Conference on Computer Vision . Springer, 654–672
Debora Caldarola, Barbara Caputo, and Marco Ciccone. 2022 · 2022
Cited alongside, same era.
Can pre-trained code embeddings improve model performance? Revisiting the use of code embeddings in software engineering tasks
Zishuo Ding, Heng Li, Weiyi Shang, and Tse-Hsun Peter Chen. 2022 · 2022
Cited alongside, same era.
VulRepair: a T5-based automated software vulnerability repair. In Proceedings of the 30th ACM joint european software engineering conference and symposium on the foundations of software engineering . 935–947
Michael Fu, Chakkrit Tantithamthavorn, Trung Le, Van Nguyen, and Dinh Phung. 2022 · 2022
Cited alongside, same era.
Multi-level branched regularization for federated learning. In International Conference on Machine Learning . PMLR, 11058–11073
Jinkyu Kim, Geeho Kim, and Bohyung Han. 2022 · 2022
Cited alongside, same era.
Federated learning on non-iid data silos: An experimental study. In 2022 IEEE 38th international conference on data engineering (ICDE) . IEEE, 965–978
Qinbin Li, Yiqun Diao, Quan Chen, and Bingsheng He. 2022b · 2022
Cited alongside, same era.
André Silva, Sen Fang, and Martin Monperrus. 2023 · 2023
Later among the works it cites.
Fedbpt: Efficient federated black-box prompt tuning for large language models
Jingwei Sun, Ziyue Xu, Hongxu Yin, Dong Yang, Daguang Xu, Yiran Chen, and Holger R Roth. 2023 · 2023
Later among the works it cites.
FedSea: Federated Learning via Selective Feature Alignment for Non-IID Multimodal Data
Min Tan, Yinfu Feng, Lingqiang Chu, Jingcheng Shi, Rong Xiao, Haihong Tang, and Jun Yu. 2023 · 2023
Later among the works it cites.
CodeStylist: a system for performing code style transfer using neural networks. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 37. 16485–16487
Chih-Kai Ting, Karl Munson, Serenity Wade, Anish Savla, Kiran Kate, and Kavitha Srinivas. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Can Public Large Language Models Help Private Cross-device Federated Learning?
Boxin Wang, Yibo Jacky Zhang, Yuan Cao, Bo Li, H Brendan McMahan, Sewoong Oh, Zheng Xu, and Manzil Zaheer. 2023b · 2023
Later among the works it cites.
CodeT5+: Open Code Large Language Models for Code Understanding and Generation. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 1069–1088
Yue Wang, Hung Le, Akhilesh Gotmare, Nghi Bui, Junnan Li, and Steven Hoi. 2023a · 2023
Later among the works it cites.
Copiloting the copilots: Fusing large language models with completion engines for automated program repair. In Proceedings of the 31st ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 172–184
Yuxiang Wei, Chunqiu Steven Xia, and Lingming Zhang. 2023 · 2023
Later among the works it cites.
Personalized federated learning under mixture of distributions. In International Conference on Machine Learning . PMLR, 37860–37879
Yue Wu, Shuaicheng Zhang, Wenchao Yu, Yanchi Liu, Quanquan Gu, Dawei Zhou, Haifeng Chen, and Wei Cheng. 2023 · 2023
Later among the works it cites.
Automated program repair in the era of large pre-trained language models. In 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 1482–1494
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Later among the works it cites.
Wizardlm: Empowering large language models to follow complex instructions
Can Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng, Pu Zhao, Jiazhan Feng, Chongyang Tao, and Daxin Jiang. 2023 · 2023
Later among the works it cites.
Dynamic personalized federated learning with adaptive differential privacy
Xiyuan Yang, Wenke Huang, and Mang Ye. 2023 · 2023
Later among the works it cites.
Fedjudge: Federated legal large language model
Linan Yue, Qi Liu, Yichao Du, Weibo Gao, Ye Liu, and Fangzhou Yao. 2023 · 2023
Later among the works it cites.
A survey of learning-based automated program repair
Quanjun Zhang, Chunrong Fang, Yuxiang Ma, Weisong Sun, and Zhenyu Chen. 2023a · 2023
Later among the works it cites.
FedSlice: Protecting Federated Learning Models from Malicious Participants with Model Slicing. In 2023 IEEE/ACM 45th International Conference on Software Engineering (ICSE) . IEEE, 460–472
Ziqi Zhang, Yuanchun Li, Bingyan Liu, Yifeng Cai, Ding Li, Yao Guo, and Xiangqun Chen. 2023b · 2023
Later among the works it cites.
Input reconstruction attack against vertical federated large language models
Fei Zheng. 2023 · 2023
Later among the works it cites.
Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x. In Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining . 5673–5684
Qinkai Zheng, Xiao Xia, Xu Zou, Yuxiao Dong, Shan Wang, Yufei Xue, Lei Shen, Zihan Wang, Andi Wang, Yang Li, et al · 2023
Later among the works it cites.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
Daya Guo, Qihao Zhu, Dejian Yang, Zhenda Xie, Kai Dong, Wentao Zhang, Guanting Chen, Xiao Bi, Yu Wu, YK Li, et al · 2024
Closest in time.
Federatedscope-llm: A comprehensive package for fine-tuning large language models in federated learning. In Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining . 5260–5271
Weirui Kuang, Bingchen Qian, Zitao Li, Daoyuan Chen, Dawei Gao, Xuchen Pan, Yuexiang Xie, Yaliang Li, Bolin Ding, and Jingren Zhou. 2024 · 2024
Closest in time.
Code Summarization without Direct Access to Code-Towards Exploring Federated LLMs for Software Engineering. In Proceedings of the 28th International Conference on Evaluation and Assessment in Software Engineering . 100–109
Jahnavi Kumar and Sridhar Chimalakonda. 2024 · 2024
Closest in time.
Fedcir: Client-invariant representation learning for federated non-iid features
Zijian Li, Zehong Lin, Jiawei Shao, Yuyi Mao, and Jun Zhang. 2024 · 2024
Closest in time.
Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2024 · 2024
Closest in time.
Position: Will we run out of data? Limits of LLM scaling based on human-generated data. In Forty-first International Conference on Machine Learning
Pablo Villalobos, Anson Ho, Jaime Sevilla, Tamay Besiroglu, Lennart Heim, and Marius Hobbhahn. 2024 · 2024
Closest in time.
Federated fine-tuning of llms on the very edge: The good, the bad, the ugly. In Proceedings of the Eighth Workshop on Data Management for End-to-End Machine Learning . 39–50
Herbert Woisetschläger, Alexander Erben, Shiqiang Wang, Ruben Mayer, and Hans-Arno Jacobsen. 2024 · 2024
Closest in time.
ConDefects: A Complementary Dataset to Address the Data Leakage Concern for LLM-Based Fault Localization and Program Repair. In Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering . 642–646
Yonghao Wu, Zheng Li, Jie M Zhang, and Yong Liu. 2024 · 2024
Closest in time.
Multi-Objective Fine-Tuning for Enhanced Program Repair with LLMs
Boyang Yang, Haoye Tian, Jiadong Ren, Hongyu Zhang, Jacques Klein, Tegawendé F Bissyandé, Claire Le Goues, and Shunfu Jin. 2024d · 2024
Closest in time.
Federated Learning for Software Engineering: A Case Study of Code Clone Detection and Defect Prediction
Yanming Yang, Xing Hu, Zhipeng Gao, Jinfu Chen, Chao Ni, Xin Xia, and David Lo. 2024a · 2024
Closest in time.
Appt: Boosting automated patch correctness prediction via fine-tuning pre-trained models
Quanjun Zhang, Chunrong Fang, Weisong Sun, Yan Liu, Tieke He, Xiaodong Hao, and Zhenyu Chen. 2024a · 2024
Closest in time.
A Systematic Literature Review on Large Language Models for Automated Program Repair
Quanjun Zhang, Chunrong Fang, Yang Xie, YuXiang Ma, Weisong Sun, and Yun Yang Zhenyu Chen. 2024b · 2024
Closest in time.
Llm-based federated recommendation
Jujia Zhao, Wenjie Wang, Chen Xu, Zhaochun Ren, See-Kiong Ng, and Tat-Seng Chua. 2024 · 2024
Closest in time.
Improving automated program repair with domain adaptation
Armin Zirak and Hadi Hemmati. 2024 · 2024
Closest in time.