Fetching the paper…
Reading the bibliography…
We introduce DS-1000, a code generation benchmark with a thousand data science problems spanning seven Python libraries, such as NumPy and Pandas.
Learning to parse database queries using inductive logic programming
Zelle, M. and Mooney, R. J · 1996
Earlier work this paper cites.
Online learning of relaxed ccg grammars for parsing to logical form
Zettlemoyer, L. and Collins, M · 2007
Earlier work this paper cites.
Codebleu: a method for automatic evaluation of code synthesis
Ren, S., Guo, D., Lu, S., Zhou, L., Liu, S., Tang, D., Sundaresan, N., Zhou, M., Blanco, A., and Ma, S · 2009
Earlier work this paper cites.
Semantic parsing on freebase from question-answer pairs
Berant, J., Chou, A., Frostig, R., and Liang, P · 2013
Earlier work this paper cites.
Data mining in education
Romero, C. and Ventura, S · 2013
Earlier work this paper cites.
A big data guide to understanding climate change: The case for theory-guided data science
Faghmous, J. H. and Kumar, V · 2014
Earlier work this paper cites.
Cosette: An automated prover for sql
Chu, S., Wang, C., Weitz, K., and Cheung, A · 2017
Earlier work this paper cites.
Learning Dependency-Based Compositional Semantics
Liang, P., Jordan, M. I., and Klein, D · 2017
Earlier work this paper cites.
Learning to mine aligned code and natural language pairs from stack overflow
Yin, P., Deng, B., Chen, E., Vasilescu, B., and Neubig, G · 2018
Earlier work this paper cites.
Spider: A large-scale human-labeled dataset for complex and cross-domain semantic parsing and text-to-SQL task
Yu, T., Zhang, R., Yang, K., Yasunaga, M., Wang, D., Li, Z., Ma, J., Li, I., Yao, Q., Roman, S., Zhang, Z., and Radev, D. R · 2018
Earlier work this paper cites.
JuICe: A large scale distantly supervised dataset for open domain context-based code generation
Agashe, R., Iyer, S., and Zettlemoyer, L · 2019
Cited alongside, same era.
Reproducible, interactive, scalable and extensible microbiome data science using qiime 2 (vol 37, pg 852, 2019)
Bolyen, E., Rideout, J. R., Dillon, M. R., Bokulich, N. A., Abnet, C. C., Al-Ghalith, G. A., Alexander, H., Alm, E. J., Arumugam, M., et al · 2019
Cited alongside, same era.
Unit test case generation with transformers and focal context
Tufano, M., Drain, D., Svyatkovskiy, A., Deng, S. K., and Sundaresan, N · 2020
Cited alongside, same era.
Semantic evaluation for text-to-SQL with distilled test suites
Zhong, R., Yu, T., and Klein, D · 2020
Cited alongside, same era.
Program synthesis with large language models
Austin, J., Odena, A., Nye, M., Bosma, M., Michalewski, H., Dohan, D., Jiang, E., Cai, C., Terry, M., Le, Q., et al · 2021
Cm3: A causal masked multimodal model of the internet
Aghajanyan, A., Huang, B., Ross, C., Karpukhin, V., Xu, H., Goyal, N., Okhonko, D., Joshi, M., Ghosh, G., Lewis, M., et al · 2022
Closest in time.
Efficient training of language models to fill in the middle
Bavarian, M., Jun, H., Tezak, N., Schulman, J., McLeavey, C., Tworek, J., and Chen, M · 2022
Closest in time.
GPT-NeoX-20B: An open-source autoregressive language model
Black, S., Biderman, S., Hallahan, E., Anthony, Q., Gao, L., Golding, L., He, H., Leahy, C., McDonell, K., Phang, J., Pieler, M., Prashanth, U. S., Purohit, S., Reynolds, L., Tow, J., Wang, B., and Weinbach, S · 2022
Closest in time.
Incoder: A generative model for code infilling and synthesis
Fried, D., Aghajanyan, A., Lin, J., Wang, S., Wallace, E., Shi, F., Zhong, R., Yih, W., Zettlemoyer, L., and Lewis, M · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Extracting training data from large language models
Carlini, N., Tramèr, F., Wallace, E., Jagielski, M., Herbert-Voss, A., Lee, K., Roberts, A., Brown, T. B., Song, D., Erlingsson, Ú., Oprea, A., and Raffel, C · 2021
Cited alongside, same era.
Memorization vs. generalization : Quantifying data leakage in NLP performance evaluation
Elangovan, A., He, J., and Verspoor, K · 2021
Cited alongside, same era.
Measuring coding challenge competence with apps
Hendrycks, D., Basart, S., Kadavath, S., Mazeika, M., Arora, A., Guo, E., Burns, C., Puranik, S., He, H., Song, D., and Steinhardt, J · 2021
Cited alongside, same era.
PICARD: Parsing incrementally for constrained auto-regressive decoding from language models
Scholak, T., Schucher, N., and Bahdanau, D · 2021
Cited alongside, same era.
Calibrate before use: Improving few-shot performance of language models
Zhao, Z., Wallace, E., Feng, S., Klein, D., and Singh, S · 2021
Cited alongside, same era.
Extracting training data from large language models
Carlini, N., Tramèr, F., Wallace, E., Jagielski, M., Herbert-Voss, A., Lee, K., Roberts, A., Brown, T., Song, D., Erlingsson, Ú., Oprea, A., and Raffel, C
Cited in the paper.
Training and evaluating a jupyter notebook data science assistant
Chandel, S., Clement, C. B., Serrato, G., and Sundaresan, N
Cited in the paper.
Li, Y., Choi, D. H., Chung, J., Kushman, N., Schrittwieser, J., Leblond, R., Eccles, T., Keeling, J., Gimeno, F., Lago, A. D., Hubert, T., Choy, P., de Masson d’Autume, C., Babuschkin, I., Chen, X., Huang, P., Welbl, J., Gowal, S., Cherepanov, A., Molloy, J., Mankowitz, D. J., Robson, E. S., Kohli, P., de Freitas, N., Kavukcuoglu, K., and Vinyals, O · 2022
Closest in time.
A conversational paradigm for program synthesis
Nijkamp, E., Pang, B., Hayashi, H., Tu, L., Wang, H., Zhou, Y., Savarese, S., and Xiong, C · 2022
Closest in time.
Synchromesh: Reliable code generation from pre-trained language models
Poesia, G., Polozov, A., Le, V., Tiwari, A., Soares, G., Meek, C., and Gulwani, S · 2022
Closest in time.
Natural language to code translation with execution
Shi, F., Fried, D., Ghazvininejad, M., Zettlemoyer, L., and Wang, S. I · 2022
Closest in time.
Unifying language learning paradigms
Tay, Y., Dehghani, M., Tran, V. Q., Garcia, X., Bahri, D., Schuster, T., Zheng, H. S., Houlsby, N., and Metzler, D · 2022
Closest in time.
A systematic evaluation of large language models of code
Xu, F. F., Alon, U., Neubig, G., and Hellendoorn, V. J · 2022
Closest in time.