Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have shown great potential in automating significant aspects of coding by producing natural code from informal natural language (NL) intent.
S. G. Hart and L. E. Staveland, “Development of nasa-tlx (task load index): Results of empirical and theoretical research,” in Advances in psychology . Elsevier, 1988, vol. 52, pp. 139–183
1988
Earlier work this paper cites.
Y. Benjamini and Y. Hochberg, “Controlling the false discovery rate: a practical and powerful approach to multiple testing,” Journal of the Royal statistical society: series B (Methodological) , vol. 57, no. 1, pp. 289–300, 1995
1995
Earlier work this paper cites.
S. G. Hart, “Nasa-task load index (nasa-tlx); 20 years later,” in Proceedings of the human factors and ergonomics society annual meeting , vol. 50, no. 9. Sage publications Sage CA: Los Angeles, CA, 2006, pp. 904–908
2006
Earlier work this paper cites.
T. Lau, “Why programming-by-demonstration systems fail: Lessons learned for usable ai,” AI Magazine , vol. 30, no. 4, pp. 65–65, 2009
2009
Earlier work this paper cites.
A. Solar-Lezama, “The sketching approach to program synthesis,” in Programming Languages and Systems , Z. Hu, Ed. Berlin, Heidelberg: Springer Berlin Heidelberg, 2009, pp. 4–13
2009
Earlier work this paper cites.
S. Jha, S. Gulwani, S. A. Seshia, and A. Tiwari, “Oracle-guided component-based program synthesis,” in Proceedings of the 32nd ACM/IEEE International Conference on Software Engineering - Volume 1, ICSE 2010, Cape Town, South Africa, 1-8 May 2010 , J. Kramer, J. Bishop, P. T. Devanbu, and S. Uchitel, Eds. ACM, 2010, pp. 215–224. [Online]. Available: https://doi.org/10.1145/1806799.1806833
2010
Earlier work this paper cites.
S. Gulwani, “Automating string processing in spreadsheets using input-output examples,” in PoPL’11, January 26-28, 2011, Austin, Texas, USA , January 2011. [Online]. Available: https://www.microsoft.com/en-us/research/publication/automating-string-processing-spreadsheets-using-input-output-examples/
2011
Earlier work this paper cites.
M. Mayer, G. Soares, M. Grechkin, V. Le, M. Marron, O. Polozov, R. Singh, B. Zorn, and S. Gulwani, “User interaction models for disambiguation in programming by example,” in Proceedings of the 28th Annual ACM Symposium on User Interface Software & Technology , 2015, pp. 291–301
2015
Earlier work this paper cites.
S. Gulwani, O. Polozov, and R. Singh, “Program synthesis,” Found. Trends Program. Lang. , vol. 4, no. 1-2, pp. 1–119, 2017. [Online]. Available: https://doi.org/10.1561/2500000010
2017
Earlier work this paper cites.
S. Jha and S. A. Seshia, “A theory of formal synthesis via inductive learning,” Acta Informatica , vol. 54, no. 7, pp. 693–726, 2017. [Online]. Available: https://doi.org/10.1007/s00236-017-0294-5
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
T. Zhang, L. Lowmanstone, X. Wang, and E. L. Glassman, “Interactive program synthesis by augmented examples,” in Proceedings of the 33rd Annual ACM Symposium on User Interface Software and Technology , 2020, pp. 627–648
2020
Earlier work this paper cites.
R. Ji, J. Liang, Y. Xiong, L. Zhang, and Z. Hu, “Question selection for interactive program synthesis,” in Proceedings of the 41st ACM SIGPLAN Conference on Programming Language Design and Implementation , ser. PLDI 2020. New York, NY, USA: Association for Computing Machinery, 2020, p. 1143–1158. [Online]. Available: https://doi.org/10.1145/3385412.3386025
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
K. Rahmani, M. Raza, S. Gulwani, V. Le, D. Morris, A. Radhakrishna, G. Soares, and A. Tiwari, “Multi-modal program inference: a marriage of pre-trained language models and component-based synthesis,” Proc. ACM Program. Lang. , vol. 5, no. OOPSLA, pp. 1–29, 2021. [Online]. Available: https://doi.org/10.1145/3485535
2021
Earlier work this paper cites.
2021
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Later among the works it cites.
P. Vaithilingam, T. Zhang, and E. L. Glassman, “Expectation vs. experience: Evaluating the usability of code generation tools powered by large language models,” in Extended Abstracts of the 2022 CHI Conference on Human Factors in Computing Systems , ser. CHI EA ’22. New York, NY, USA: Association for Computing Machinery, 2022. [Online]. Available: https://doi.org/10.1145/3491101.3519665
2022
Later among the works it cites.
N. Nguyen and S. Nadi, “An empirical evaluation of github copilot’s code suggestions,” in Proceedings of the 19th International Conference on Mining Software Repositories , 2022, pp. 1–5
2022
Later among the works it cites.
S. Imai, “Is github copilot a substitute for human pair-programming? an empirical study,” in Proceedings of the ACM/IEEE 44th International Conference on Software Engineering: Companion Proceedings , 2022, pp. 319–321
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
F. F. Xu, U. Alon, G. Neubig, and V. J. Hellendoorn, “A systematic evaluation of large language models of code,” in Proceedings of the 6th ACM SIGPLAN International Symposium on Machine Programming , ser. MAPS 2022. New York, NY, USA: Association for Computing Machinery, 2022, p. 1–10. [Online]. Available: https://doi.org/10.1145/3520312.3534862
2022
Cited alongside, same era.
GitHub, “Github copilot,” 2022, accessed August 5, 2022. https://github.com/features/copilot/
2022
Cited alongside, same era.
A. Ziegler, E. Kalliamvakou, X. A. Li, A. Rice, D. Rifkin, S. Simister, G. Sittampalam, and E. Aftandilian, “Productivity assessment of neural code completion,” in MAPS@PLDI 2022: 6th ACM SIGPLAN International Symposium on Machine Programming, San Diego, CA, USA, 13 June 2022 , S. Chaudhuri and C. Sutton, Eds. ACM, 2022, pp. 21–29. [Online]. Available: https://doi.org/10.1145/3520312.3534864
2022
Cited alongside, same era.
F. F. Xu, B. Vasilescu, and G. Neubig, “In-ide code generation from natural language: Promise and challenges,” ACM Transactions on Software Engineering and Methodology (TOSEM) , vol. 31, no. 2, pp. 1–47, 2022
2022
Cited alongside, same era.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” Advances in Neural Information Processing Systems , vol. 35, pp. 27 730–27 744, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
C. Bird, D. Ford, T. Zimmermann, N. Forsgren, E. Kalliamvakou, T. Lowdermilk, and I. Gazit, “Taking flight with copilot: Early insights and opportunities of ai-powered pair-programming tools,” Queue , vol. 20, no. 6, pp. 35–57, 2022
2022
Cited alongside, same era.
2022
Later among the works it cites.
2023
Later among the works it cites.
C. Lemieux, J. P. Inala, S. K. Lahiri, and S. Sen, “Codamosa: Escaping coverage plateaus in test generation with pre-trained large language models,” in 45th International Conference on Software Engineering, ser. ICSE , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
A. M. McNutt, C. Wang, R. A. Deline, and S. M. Drucker, “On the design of ai-powered code assistants for notebooks,” in Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems , 2023, pp. 1–16
2023
Later among the works it cites.
S. Barke, M. B. James, and N. Polikarpova, “Grounded copilot: How programmers interact with code-generating models,” Proceedings of the ACM on Programming Languages , vol. 7, no. OOPSLA1, pp. 85–111, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.