Fetching the paper…
Reading the bibliography…
While code generation has been widely used in various software development scenarios, the quality of the generated code is not guaranteed.
Toward automatic program synthesis
Manna, Z., and Waldinger, R. J · 1971
Earlier work this paper cites.
Application of theorem proving to problem solving
Green, C · 1981
Earlier work this paper cites.
Learning bayesian networks: The combination of knowledge and statistical data
Heckerman, D., Geiger, D., and Chickering, D. M · 1995
Earlier work this paper cites.
Causation, prediction, and search
Spirtes, P., Glymour, C. N., Scheines, R., and Heckerman, D · 2000
Earlier work this paper cites.
Learning equivalence classes of bayesian-network structures
Chickering, D. M · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
The max-min hill-climbing bayesian network structure learning algorithm
Tsamardinos, I., Brown, L. E., and Aliferis, C. F · 2006
Earlier work this paper cites.
Program synthesis by sketching
Solar-Lezama, A · 2008
Earlier work this paper cites.
Cognitively motivated features for readability assessment
Feng, L., Elhadad, N., and Huenerfauth, M · 2009
Earlier work this paper cites.
Causality
Pearl, J · 2009
Earlier work this paper cites.
Automatic analysis of syntactic complexity in second language writing
Lu, X · 2010
Earlier work this paper cites.
Search based software test data generation for structural testing: a perspective
Varshney, S., and Mehrotra, M · 2013
Earlier work this paper cites.
Computational assessment of text readability: A survey of current and future research
Collins-Thompson, K · 2014
Earlier work this paper cites.
The stanford corenlp natural language processing toolkit
Manning, C. D., Surdeanu, M., Bauer, J., Finkel, J. R., Bethard, S., and McClosky, D · 2014
Earlier work this paper cites.
Learning bayesian networks with thousands of variables
Scanagatta, M., de Campos, C. P., Corani, G., and Zaffalon, M · 2015
Earlier work this paper cites.
Weakly supervised part-of-speech tagging using eye-tracking data
Barrett, M., Bingel, J., Keller, F., and Søgaard, A · 2016
Earlier work this paper cites.
Double/debiased machine learning for treatment and causal parameters
Chernozhukov, V., Chetverikov, D., Demirer, M., Duflo, E., Hansen, C., Newey, W., and Robins, J · 2016
Earlier work this paper cites.
Bayesian network structure learning with integer programming: Polytopes, facets and complexity
Cussens, J., Järvisalo, M., Korhonen, J. H., and Bartlett, M · 2017
Earlier work this paper cites.
Program synthesis
Gulwani, S., Polozov, O., Singh, R., et al · 2017
Earlier work this paper cites.
Asyncclock: Scalable inference of asynchronous event causality
Hsiao, C.-H., Narayanasamy, S., Khan, E. M. I., Pereira, C. L., and Pokam, G. A · 2017
Earlier work this paper cites.
Elements of causal inference: foundations and learning algorithms
Peters, J., Janzing, D., and Schölkopf, B · 2017
Earlier work this paper cites.
On software productivity analysis with propensity score matching
Tsunoda, M., and Amasaki, S · 2017
Earlier work this paper cites.
A syntactic neural model for general-purpose code generation
Yin, P., and Neubig, G · 2017
Earlier work this paper cites.
A call for clarity in reporting BLEU scores
Post, M · 2018
Earlier work this paper cites.
Dags with no tears: Continuous optimization for structure learning
Zheng, X., Aragam, B., Ravikumar, P. K., and Xing, E. P · 2018
Earlier work this paper cites.
Codesearchnet challenge: Evaluating the state of semantic code search
Husain, H., Wu, H.-H., Gazit, T., Allamanis, M., and Brockschmidt, M · 2019
Earlier work this paper cites.
A grammar-based structural cnn decoder for code generation
Sun, Z., Zhu, Q., Mou, L., Xiong, Y., Li, G., and Zhang, L · 2019
Earlier work this paper cites.
Code generation as a dual task of code summarization
Wei, B., Li, G., Xia, X., Fu, Z., and Jin, Z · 2019
Earlier work this paper cites.
Text readability assessment for second language learners
Xia, M., Kochmar, E., and Briscoe, T · 2019
Cited alongside, same era.
Dag-gnn: Dag structure learning with graph neural networks
Yu, Y., Chen, J., Gao, T., and Yu, M · 2019
Cited alongside, same era.
Fine-tuning language models from human preferences
Ziegler, D. M., Stiennon, N., Wu, J., Brown, T. B., Radford, A., Amodei, D., Christiano, P., and Irving, G · 2019
Cited alongside, same era.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D · 2020
Cited alongside, same era.
Causality-guided adaptive interventional debugging
Fariha, A., Nath, S., and Meliou, A · 2020
Chain of thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Chi, E., Le, Q., and Zhou, D · 2022
Later among the works it cites.
Uncertainty quantification with pre-trained language models: A large-scale empirical analysis
Xiao, Y., Liang, P. P., Bhatt, U., Neiswanger, W., Salakhutdinov, R., and Morency, L.-P · 2022
Later among the works it cites.
Adaptive fairness improvement based on causality analysis
Zhang, M., and Sun, J · 2022
Later among the works it cites.
Large language models are human-level prompt engineers
Zhou, Y., Muresanu, A. I., Han, Z., Paster, K., Pitis, S., Chan, H., and Ba, J · 2022
Later among the works it cites.
https://chat.openai.com/chat , 2023
ChatGPT · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Causal testing: understanding defects’ root causes
Johnson, B., Brun, Y., and Meliou, A · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2020
Cited alongside, same era.
Codebleu: a method for automatic evaluation of code synthesis
Ren, S., Guo, D., Lu, S., Zhou, L., Liu, S., Tang, D., Sundaresan, N., Zhou, M., Blanco, A., and Ma, S · 2020
Cited alongside, same era.
Program synthesis with large language models
Austin, J., Odena, A., Nye, M., Bosma, M., Michalewski, H., Dohan, D., Jiang, E., Cai, C., Terry, M., Le, Q., et al · 2021
Cited alongside, same era.
Evaluating large language models trained on code
Chen, M., Tworek, J., Jun, H., Yuan, Q., Pinto, H. P. d. O., Kaplan, J., Edwards, H., Burda, Y., Joseph, N., Brockman, G., et al · 2021
Cited alongside, same era.
Measuring coding challenge competence with apps
Hendrycks, D., Basart, S., Kadavath, S., Mazeika, M., Arora, A., Guo, E., Burns, C., Puranik, S., He, H., Song, D., and Steinhardt, J · 2021
Cited alongside, same era.
Pushing on text readability assessment: A transformer meets handcrafted linguistic features
Lee, B. W., Jang, Y. S., and Lee, J · 2021
Cited alongside, same era.
https://quillbot.com/ , 2023
Rephrasing Tool - QuillBot AI · 2023
Closest in time.
https://anonymous.4open.science/r/CALL-E7F2/ , 2023
Research artifact · 2023
Closest in time.
Language models can explain neurons in language models
Bills, S., Cammarata, N., Mossing, D., Tillman, H., Gao, L., Goh, G., Sutskever, I., Leike, J., Wu, J., and Saunders, W · 2023
Closest in time.
Improving factuality and reasoning in language models through multiagent debate
Du, Y., Li, S., Torralba, A., Tenenbaum, J. B., and Mordatch, I · 2023
Closest in time.
Automated repair of programs from large language models
Fan, Z., Gao, X., Mirchev, M., Roychoudhury, A., and Tan, S. H · 2023
Closest in time.
Furia, C. A., Torkar, R., and Feldt, R · 2023
Closest in time.
Evaluating the robustness of discrete prompts
Ishibashi, Y., Bollegala, D., Sudoh, K., and Nakamura, S · 2023
Closest in time.
Cc: Causality-aware coverage criterion for deep neural networks
Ji, Z., Ma, P., Yuan, Y., and Wang, S · 2023
Closest in time.
Is chatgpt a good translator? a preliminary study
Jiao, W., Wang, W., Huang, J.-t., Wang, X., and Tu, Z · 2023
Closest in time.
Kuhn, L., Gal, Y., and Farquhar, S · 2023
Closest in time.
LFTK: Handcrafted features in computational linguistics
Lee, B. W., and Lee, J · 2023
Closest in time.
Cctest: Testing and repairing code completion systems, 2023
Li, Z., Wang, C., Liu, Z., Wang, H., Wang, S., and Gao, C · 2023
Closest in time.
Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing
Liu, P., Yuan, W., Fu, J., Jiang, Z., Hayashi, H., and Neubig, G · 2023
Closest in time.
On the reliability and explainability of automated code generation approaches
Liu, Y., Tantithamthavorn, C., Liu, Y., and Li, L · 2023
Closest in time.
Gpt-4 technical report
OpenAI, R · 2023
Closest in time.
Llm is like a box of chocolates: the non-determinism of chatgpt in code generation
Ouyang, S., Zhang, J. M., Harman, M., and Wang, M · 2023
Closest in time.
Automatic prompt optimization with” gradient descent” and beam search
Pryzant, R., Iter, D., Li, J., Lee, Y. T., Zhu, C., and Zeng, M · 2023
Closest in time.
Applications of statistical causal inference in software engineering
Siebert, J · 2023
Closest in time.
Large language models encode clinical knowledge
Singhal, K., Azizi, S., Tu, T., Mahdavi, S. S., Wei, J., Chung, H. W., Scales, N., Tanwani, A., Cole-Lewis, H., Pfohl, S., et al · 2023
Closest in time.
Large language models as optimizers
Yang, C., Wang, X., Lu, Y., Liu, H., Le, Q. V., Zhou, D., and Chen, X · 2023
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Yao, S., Yu, D., Zhao, J., Shafran, I., Griffiths, T. L., Cao, Y., and Narasimhan, K · 2023
Closest in time.
Cumulative reasoning with large language models
Zhang, Y., Yang, J., Yuan, Y., and Yao, A. C.-C · 2023
Closest in time.