Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have made significant progress in code generation, offering developers groundbreaking automated programming support.
Learning Source Phrase Representations for Neural Machine Translation
Xu, H.; van Genabith, J.; Xiong, D.; Liu, Q.; and Zhang, J. 2020 · 2020
Earlier work this paper cites.
Program Synthesis with Large Language Models
Austin, J.; Odena, A.; Nye, M. I.; Bosma, M.; Michalewski, H.; Dohan, D.; Jiang, E.; Cai, C. J.; Terry, M.; Le, Q. V.; and Sutton, C. 2021 · 2021
Earlier work this paper cites.
Evaluating Large Language Models Trained on Code
Chen, M.; Tworek, J.; Jun, H.; Yuan, Q.; de Oliveira Pinto, H. P.; Kaplan, J.; Edwards, H.; Burda, Y.; Joseph, N.; Brockman, G.; Ray, A.; Puri, R.; Krueger, G.; Petrov, M.; Khlaaf, H.; Sastry, G.; Mishkin, P.; Chan, B.; Gray, S.; Ryder, N.; Pavlov, M.; Power, A.; Kaiser, L.; Bavarian, M.; Winter, C.; Tillet, P.; Such, F. P.; Cummings, D.; Plappert, M.; Chantzis, F.; Barnes, E.; Herbert-Voss, A.; Guss, W. H.; Nichol, A.; Paino, A.; Tezak, N.; Tang, J.; Babuschkin, I.; Balaji, S.; Jain, S.; Saunders, W.; Hesse, C.; Carr, A. N.; Leike, J.; Achiam, J.; Misra, V.; Morikawa, E.; Radford, A.; Knight, M.; Brundage, M.; Murati, M.; Mayer, K.; Welinder, P.; McGrew, B.; Amodei, D.; McCandlish, S.; Sutskever, I.; and Zaremba, W. 2021 · 2021
Earlier work this paper cites.
Glm: General language model pretraining with autoregressive blank infilling
Du, Z.; Qian, Y.; Liu, X.; Ding, M.; Qiu, J.; Yang, Z.; and Tang, J. 2021 · 2021
Earlier work this paper cites.
Measuring coding challenge competence with apps
Hendrycks, D.; Basart, S.; Kadavath, S.; Mazeika, M.; Arora, A.; Guo, E.; Burns, C.; Puranik, S.; He, H.; Song, D.; et al. 2021 · 2021
Earlier work this paper cites.
Truthfulqa: Measuring how models mimic human falsehoods
Lin, S.; Hilton, J.; and Evans, O. 2021 · 2021
Earlier work this paper cites.
Recent Advances in Intelligent Source Code Generation: A Survey on Natural Language Based Studies
Yang, C.; Liu, Y.; and Yin, C. 2021 · 2021
Earlier work this paper cites.
Learning to break the loop: Analyzing and mitigating repetitions for neural text generation
Xu, J.; Liu, X.; Yan, J.; Cai, D.; Li, H.; and Li, J. 2022 · 2022
Earlier work this paper cites.
WhyGen: explaining ML-powered code generation by referring to training examples
Yan, W.; and Li, Y. 2022 · 2022
Earlier work this paper cites.
Bai, J.; Bai, S.; Chu, Y.; Cui, Z.; Dang, K.; Deng, X.; Fan, Y.; Ge, W.; Han, Y.; Huang, F.; et al. 2023 · 2023
Earlier work this paper cites.
Introducing ERNIE 3.5: Baidu’s Knowledge-Enhanced Foundation Model Takes a Giant Leap Forward
Baidu. 2023 · 2023
Earlier work this paper cites.
Evaluating hallucinations in chinese large language models
Cheng, Q.; Sun, T.; Zhang, W.; Wang, S.; Liu, X.; Zhang, M.; He, J.; Huang, M.; Yin, Z.; Chen, K.; et al. 2023 · 2023
Earlier work this paper cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; et al. 2023 · 2023
Earlier work this paper cites.
Halo: Estimation and Reduction of Hallucinations in Open-Source Weak Large Language Models
Elaraby, M.; Lu, M.; Dunn, J.; Zhang, X.; Wang, Y.; and Liu, S. 2023 · 2023
Earlier work this paper cites.
Gemini: a family of highly capable multimodal models
Gemini. 2023 · 2023
Cited alongside, same era.
An Empirical Study on Fine-Tuning Large Language Models of Code for Automated Program Repair
Huang, K.; Meng, X.; Zhang, J.; Liu, Y.; Wang, W.; Li, S.; and Zhang, Y. 2023 · 2023
Cited alongside, same era.
Large Language Models and Simple, Stupid Bugs
Jesse, K.; Ahmed, T.; Devanbu, P. T.; and Morgan, E. 2023 · 2023
Cited alongside, same era.
Survey of hallucination in natural language generation
Ji, Z.; Lee, N.; Frieske, R.; Yu, T.; Su, D.; Xu, Y.; Ishii, E.; Bang, Y. J.; Madotto, A.; and Fung, P. 2023 · 2023
Cited alongside, same era.
Jiang, A. Q.; Sablayrolles, A.; Mensch, A.; Bamford, C.; Chaplot, D. S.; Casas, D. d. l.; Bressand, F.; Lengyel, G.; Lample, G.; Saulnier, L.; et al. 2023 · 2023
Cited alongside, same era.
Yan, W.; Liu, H.; Wang, Y.; Li, Y.; Chen, Q.; Wang, W.; Lin, T.; Zhao, W.; Zhu, L.; Deng, S.; et al. 2023 · 2023
Later among the works it cites.
Zhai, B.; Yang, S.; Zhao, X.; Xu, C.; Shen, S.; Zhao, D.; Keutzer, K.; Li, M.; Yan, T.; and Fan, X. 2023 · 2023
Later among the works it cites.
Siren’s song in the AI ocean: a survey on hallucination in large language models
Zhang, Y.; Li, Y.; Cui, L.; Cai, D.; Liu, L.; Fu, T.; Huang, X.; Zhao, E.; Zhang, Y.; Chen, Y.; et al. 2023 · 2023
Later among the works it cites.
Codegeex: A pre-trained model for code generation with multilingual evaluations on humaneval-x
Zheng, Q.; Xia, X.; Zou, X.; Dong, Y.; Wang, S.; Xue, Y.; Wang, Z.; Shen, L.; Wang, A.; Li, Y.; et al. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Swe-bench: Can language models resolve real-world github issues?
Jimenez, C. E.; Yang, J.; Wettig, A.; Yao, S.; Pei, K.; Press, O.; and Narasimhan, K. 2023 · 2023
Cited alongside, same era.
Aligning large multi-modal model with robust instruction tuning
Liu, F.; Lin, K.; Li, L.; Wang, J.; Yacoob, Y.; and Wang, L. 2023 · 2023
Cited alongside, same era.
Wizardcoder: Empowering code large language models with evol-instruct
Luo, Z.; Xu, C.; Zhao, P.; Sun, Q.; Geng, X.; Hu, W.; Tao, C.; Ma, J.; Lin, Q.; and Jiang, D. 2023 · 2023
Cited alongside, same era.
GPT-4 Technical Report
OpenAI. 2023 · 2023
Cited alongside, same era.
Understanding the effectiveness of large language models in code translation
Pan, R.; Ibrahimzada, A. R.; Krishna, R.; Sankar, D.; Wassi, L. P.; Merler, M.; Sobolev, B.; Pavuluri, R.; Sinha, S.; and Jabbarvand, R. 2023 · 2023
Cited alongside, same era.
Peng, B.; Galley, M.; He, P.; Cheng, H.; Xie, Y.; Hu, Y.; Huang, Q.; Liden, L.; Yu, Z.; Chen, W.; et al. 2023 · 2023
Cited alongside, same era.
Code llama: Open foundation models for code
Roziere, B.; Gehring, J.; Gloeckle, F.; Sootla, S.; Gat, I.; Tan, X. E.; Adi, Y.; Liu, J.; Remez, T.; Rapin, J.; et al. 2023 · 2023
Cited alongside, same era.
The Claude 3 Model Family: Opus, Sonnet, Haiku
Anthropic. 2024 · 2024
Closest in time.
Sora Detector: A Unified Hallucination Detection for Large Text-to-Video Models
Chu, Z.; Zhang, L.; Sun, Y.; Xue, S.; Wang, Z.; Qin, Z.; and Ren, K. 2024 · 2024
Closest in time.
DeepSeek-Coder: When the Large Language Model Meets Programming–The Rise of Code Intelligence
Guo, D.; Zhu, Q.; Yang, D.; Xie, Z.; Dong, K.; Zhang, W.; Chen, G.; Bi, X.; Wu, Y.; Li, Y.; et al. 2024 · 2024
Closest in time.
Visual Hallucinations of Multi-modal Large Language Models
Huang, W.; Liu, H.; Guo, M.; and Gong, N. Z. 2024 · 2024
Closest in time.
MMCode: Evaluating Multi-Modal Code Large Language Models with Visually Rich Programming Problems
Li, K.; Tian, Y.; Hu, Q.; Luo, Z.; and Ma, J. 2024 · 2024
Closest in time.
A survey on hallucination in large vision-language models
Liu, H.; Xue, W.; Chen, Y.; Chen, D.; Zhao, X.; Wang, K.; Hou, L.; Li, R.; and Peng, W. 2024 · 2024
Closest in time.
Gemma: Open models based on gemini research and technology
Team, G.; Mesnard, T.; Hardin, C.; Dadashi, R.; Bhupatiraju, S.; Pathak, S.; Sifre, L.; Rivière, M.; Kale, M. S.; Love, J.; et al. 2024 · 2024
Closest in time.
Where Do Large Language Models Fail When Generating Code?
Wang, Z.; Zhou, Z.; Song, D.; Huang, Y.; Chen, S.; Ma, L.; and Zhang, T. 2024 · 2024
Closest in time.
HiRoPE: Length Extrapolation for Code Models
Zhang, K.; Li, G.; Zhang, H.; and Jin, Z. 2024 · 2024
Closest in time.