Fetching the paper…
Reading the bibliography…
In recent years, large language models (LLMs) have seen rapid advancements, significantly impacting various fields such as computer vision, natural language processing, and software engineering.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
The measurement of observer agreement for categorical data
J Richard Landis and Gary G Koch. 1977 · 1977
Earlier work this paper cites.
Qualitative methods in empirical studies of software engineering
Carolyn B. Seaman. 1999 · 1999
Earlier work this paper cites.
A field study of API learning obstacles
Martin P Robillard and Robert DeLine. 2011 · 2011
Earlier work this paper cites.
An empirical study of bugs in machine learning systems. In 2012 IEEE 23rd International Symposium on Software Reliability Engineering . IEEE, 271–280
Ferdian Thung, Shaowei Wang, David Lo, and Lingxiao Jiang. 2012 · 2012
Earlier work this paper cites.
Improving API usability
Brad A Myers and Jeffrey Stylos. 2016 · 2016
Earlier work this paper cites.
What are mobile developers asking about? a large scale study using stack overflow
Christoffer Rosen and Emad Shihab. 2016 · 2016
Earlier work this paper cites.
What security questions do developers ask? a large-scale study of stack overflow posts
Xin-Li Yang, David Lo, Xin Xia, Zhi-Yuan Wan, and Jian-Ling Sun. 2016 · 2016
Earlier work this paper cites.
What do concurrency developers ask about? a large-scale study using stack overflow. In Proceedings of the 12th ACM/IEEE international symposium on empirical software engineering and measurement . 1–10
Syed Ahmed and Mehdi Bagherzadeh. 2018 · 2018
Earlier work this paper cites.
An empirical study on tensorflow program bugs. In Proceedings of the 27th ACM SIGSOFT international symposium on software testing and analysis . 129–140
Yuhao Zhang, Yifan Chen, Shing-Chi Cheung, Yingfei Xiong, and Lu Zhang. 2018 · 2018
Earlier work this paper cites.
Software documentation issues unveiled. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 1199–1210
Emad Aghajani, Csaba Nagy, Olga Lucero Vega-Márquez, Mario Linares-Vásquez, Laura Moreno, Gabriele Bavota, and Michele Lanza. 2019 · 2019
Earlier work this paper cites.
Why is developing machine learning applications challenging? a study on stack overflow posts. In 2019 acm/ieee international symposium on empirical software engineering and measurement (esem) . IEEE, 1–11
Moayad Alshangiti, Hitesh Sapkota, Pradeep K Murukannaiah, Xumin Liu, and Qi Yu. 2019 · 2019
Earlier work this paper cites.
Going big: a large-scale study on what big data developers ask. In Proceedings of the 2019 27th ACM joint meeting on european software engineering conference and symposium on the foundations of software engineering . 432–442
Mehdi Bagherzadeh and Raffi Khatchadourian. 2019 · 2019
Earlier work this paper cites.
An empirical study towards characterizing deep learning development and deployment across different frameworks and platforms. In 2019 34th IEEE/ACM International Conference on Automated Software Engineering (ASE) . IEEE, 810–822
Qianyu Guo, Sen Chen, Xiaofei Xie, Lei Ma, Qiang Hu, Hongtao Liu, Yang Liu, Jianjun Zhao, and Xiaohong Li. 2019 · 2019
Earlier work this paper cites.
A comprehensive study on deep learning bug characteristics. In Proceedings of the 2019 27th ACM joint meeting on european software engineering conference and symposium on the foundations of software engineering . 510–520
Md Johirul Islam, Giang Nguyen, Rangeet Pan, and Hridesh Rajan. 2019 · 2019
Earlier work this paper cites.
An empirical study of common challenges in developing deep learning applications. In 2019 IEEE 30th International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 104–115
Tianyi Zhang, Cuiyun Gao, Lei Ma, Michael Lyu, and Miryung Kim. 2019 · 2019
Earlier work this paper cites.
Challenges in chatbot development: A study of stack overflow posts. In Proceedings of the 17th international conference on mining software repositories . 174–185
Ahmad Abdellatif, Diego Costa, Khaled Badran, Rabe Abdalkareem, and Emad Shihab. 2020 · 2020
Earlier work this paper cites.
A comprehensive study on challenges in deploying deep learning based software. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 750–762
Zhenpeng Chen, Yanbin Cao, Yuanqiang Liu, Haoyu Wang, Tao Xie, and Xuanzhe Liu. 2020 · 2020
Earlier work this paper cites.
Taxonomy of real faults in deep learning systems. In Proceedings of the ACM/IEEE 42nd international conference on software engineering . 1110–1121
Nargiz Humbatova, Gunel Jahangirova, Gabriele Bavota, Vincenzo Riccio, Andrea Stocco, and Paolo Tonella. 2020 · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Earlier work this paper cites.
Understanding build issue resolution in practice: symptoms and fix patterns. In Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 617–628
Yiling Lou, Zhenpeng Chen, Yanbin Cao, Dan Hao, and Lu Zhang. 2020 · 2020
Earlier work this paper cites.
Challenges in developing desktop web apps: a study of stack overflow and github. In 2021 IEEE/ACM 18th International Conference on Mining Software Repositories (MSR) . IEEE, 271–282
Gian Luca Scoccia, Patrizio Migliarini, and Marco Autili. 2021 · 2021
Earlier work this paper cites.
A comprehensive study of deep learning compiler bugs. In Proceedings of the 29th ACM Joint meeting on european software engineering conference and symposium on the foundations of software engineering . 968–980
Qingchao Shen, Haoyang Ma, Junjie Chen, Yongqiang Tian, Shing-Chi Cheung, and Xiang Chen. 2021 · 2021
Earlier work this paper cites.
An empirical study on challenges of application development in serverless computing. In Proceedings of the 29th ACM joint meeting on European software engineering conference and symposium on the foundations of software engineering . 416–428
Jinfeng Wen, Zhenpeng Chen, Yi Liu, Yiling Lou, Yun Ma, Gang Huang, Xin Jin, and Xuanzhe Liu. 2021 · 2021
Earlier work this paper cites.
Understanding performance problems in deep learning systems. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 357–369
Junming Cao, Bihuan Chen, Chao Sun, Longjie Hu, Shuaihong Wu, and Xin Peng. 2022 · 2022
Earlier work this paper cites.
A survey on in-context learning
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Zhiyong Wu, Baobao Chang, Xu Sun, Jingjing Xu, and Zhifang Sui. 2022 · 2022
Cited alongside, same era.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Cited alongside, same era.
Automating code review activities by large-scale pre-training. In Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1035–1047
Zhiyu Li, Shuai Lu, Daya Guo, Nan Duan, Shailesh Jannu, Grant Jenks, Deep Majumder, Jared Green, Alexey Svyatkovskiy, Shengyu Fu, et al · 2022
Cited alongside, same era.
Barriers in front-end web development. In 2022 IEEE Symposium on Visual Languages and Human-Centric Computing (VL/HCC) . IEEE, 1–11
David I Samudio and Thomas D LaToza. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
No more manual tests? evaluating and improving chatgpt for unit test generation
Zhiqiang Yuan, Yiling Lou, Mingwei Liu, Shiji Ding, Kaixin Wang, Yixuan Chen, and Xin Peng. 2023 · 2023
Later among the works it cites.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Later among the works it cites.
Automatic semantic augmentation of language model prompts (for code summarization). In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering . 1–13
Toufique Ahmed, Kunal Suresh Pai, Premkumar Devanbu, and Earl Barr. 2024 · 2024
Closest in time.
De-hallucinator: Iterative grounding for llm-based code completion
Aryaz Eghbali and Michael Pradel. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
Creating trustworthy llms: Dealing with hallucinations in healthcare ai
Muhammad Aurangzeb Ahmad, Ilker Yaramis, and Taposh Dutta Roy. 2023 · 2023
Cited alongside, same era.
Is ChatGPT leading generative AI? What is beyond expectations?
Ömer Aydın and Enis Karaarslan. 2023 · 2023
Cited alongside, same era.
When chatgpt meets smart contract vulnerability detection: How far are we?
Chong Chen, Jianzhong Su, Jiachi Chen, Yanlin Wang, Tingting Bi, Yanli Wang, Xingwei Lin, Ting Chen, and Zibin Zheng. 2023b · 2023
Cited alongside, same era.
Toward understanding deep learning framework bugs
Junjie Chen, Yihua Liang, Qingchao Shen, Jiajun Jiang, and Shuochuan Li. 2023a · 2023
Cited alongside, same era.
Frugalgpt: How to use large language models while reducing cost and improving performance
Lingjiao Chen, Matei Zaharia, and James Zou. 2023c · 2023
Cited alongside, same era.
Towards understanding the capability of large language models on code clone detection: a survey
Shihan Dou, Junjie Shan, Haoxiang Jia, Wenhao Deng, Zhiheng Xi, Wei He, Yueming Wu, Tao Gui, Yang Liu, and Xuanjing Huang. 2023 · 2023
Cited alongside, same era.
Large language models for software engineering: Survey and open problems
Angela Fan, Beliz Gokkaya, Mark Harman, Mitya Lyubarskiy, Shubho Sengupta, Shin Yoo, and Jie M Zhang. 2023 · 2023
Cited alongside, same era.
Large language models are few-shot summarizers: Multi-intent comment generation via in-context learning. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Mingyang Geng, Shangwen Wang, Dezun Dong, Haotian Wang, Ge Li, Zhi Jin, Xiaoguang Mao, and Xiangke Liao. 2024 · 2024
Closest in time.
On the effectiveness of large language models in domain-specific code generation
Xiaodong Gu, Meng Chen, Yalan Lin, Yuhan Hu, Hongyu Zhang, Chengcheng Wan, Zhao Wei, Yong Xu, and Juhong Wang. 2024 · 2024
Closest in time.
A deep dive into large language models for automated bug localization and repair
Soneya Binta Hossain, Nan Jiang, Qiang Zhou, Xiaopeng Li, Wen-Hao Chiang, Yingjun Lyu, Hoan Nguyen, and Omer Tripp. 2024 · 2024
Closest in time.
A quantitative and qualitative evaluation of LLM-based explainable fault localization
Sungmin Kang, Gabin An, and Shin Yoo. 2024 · 2024
Closest in time.
Exploring and evaluating hallucinations in llm-powered code generation
Fang Liu, Yang Liu, Lin Shi, Houkun Huang, Ruifeng Wang, Zhen Yang, Li Zhang, Zhongqi Li, and Yuchi Ma. 2024b · 2024
Closest in time.
Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2024d · 2024
Closest in time.
Refining chatgpt-generated code: Characterizing and mitigating code quality issues
Yue Liu, Thanh Le-Cong, Ratnadira Widyasari, Chakkrit Tantithamthavorn, Li Li, Xuan-Bach D Le, and David Lo. 2024a · 2024
Closest in time.
On the reliability and explainability of language models for program generation
Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, and Li Li. 2024c · 2024
Closest in time.
GRACE: Empowering LLM-based software vulnerability detection with graph structure and in-context learning
Guilong Lu, Xiaolin Ju, Xiang Chen, Wenlong Pei, and Zhilong Cai. 2024 · 2024
Closest in time.
Common challenges of deep reinforcement learning applications development: an empirical study
Mohammad Mehdi Morovati, Florian Tambon, Mina Taraghi, Amin Nikanjam, and Foutse Khomh. 2024 · 2024
Closest in time.
Domain knowledge matters: Improving prompts with fix templates for repairing python type errors. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Yun Peng, Shuzheng Gao, Cuiyun Gao, Yintong Huo, and Michael Lyu. 2024 · 2024
Closest in time.
Code-Aware Prompting: A Study of Coverage-Guided Test Generation in Regression Setting using LLM
Gabriel Ryan, Siddhartha Jain, Mingyue Shang, Shiqi Wang, Xiaofei Ma, Murali Krishna Ramanathan, and Baishakhi Ray. 2024 · 2024
Closest in time.
Dataflow analysis-inspired deep learning for efficient vulnerability detection. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–13
Benjamin Steenhoek, Hongyang Gao, and Wei Le. 2024 · 2024
Closest in time.
Source code summarization in the era of large language models
Weisong Sun, Yun Miao, Yuekang Li, Hongyu Zhang, Chunrong Fang, Yi Liu, Gelei Deng, Yang Liu, and Zhenyu Chen. 2024 · 2024
Closest in time.
Software testing with large language models: Survey, landscape, and vision
Junjie Wang, Yuchao Huang, Chunyang Chen, Zhe Liu, Song Wang, and Qing Wang. 2024 · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan. 2024 · 2024
Closest in time.
Fine-tuning large language models to improve accuracy and comprehensibility of automated code review
Yongda Yu, Guoping Rong, Haifeng Shen, He Zhang, Dong Shao, Min Wang, Zhao Wei, Yong Xu, and Juhong Wang. 2024 · 2024
Closest in time.
Assessing the Code Clone Detection Capability of Large Language Models. In 2024 4th International Conference on Code Quality (ICCQ) . IEEE, 75–83
Zixian Zhang and Takfarinas Saber. 2024 · 2024
Closest in time.
Llm hallucinations in practical code generation: Phenomena, mechanism, and mitigation
Ziyao Zhang, Yanlin Wang, Chong Wang, Jiachi Chen, and Zibin Zheng. 2024 · 2024
Closest in time.
Automatic smart contract comment generation via large language models and in-context learning
Junjie Zhao, Xiang Chen, Guang Yang, and Yiheng Shen. 2024 · 2024
Closest in time.