Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have sparked significant interest in their generative capabilities, leading to the development of various commercial applications.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020) · 1901
Earlier work this paper cites.
A comparative analysis of selection schemes used in genetic algorithms
Goldberg, D. E. and Deb, K. (1991) · 1991
Earlier work this paper cites.
Efficient progressive sampling
Provost, F., Jensen, D., and Oates, T. (1999) · 1999
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. (2004) · 2004
Earlier work this paper cites.
A systematic characterization of sampling algorithms for open-ended language generation
Nadeem, M., He, T., Cho, K., and Glass, J. (2020) · 2009
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
Bergstra, J. S., Bardenet, R., Bengio, Y., and Kégl, B. (2011) · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
Bergstra, J. and Bengio, Y. (2012) · 2012
Earlier work this paper cites.
Concentration inequalities for sampling without replacement
Bardenet, R. and Maillard, O.-A. (2015) · 2015
Earlier work this paper cites.
Hyperband: A novel bandit-based approach to hyperparameter optimization
Li, L., Jamieson, K., DeSalvo, G., Rostamizadeh, A., and Talwalkar, A. (2017) · 2017
Earlier work this paper cites.
Narayan, S., Cohen, S. B., and Lapata, M. (2018) · 2018
Earlier work this paper cites.
Hyperparameter optimization
Feurer, M. and Hutter, F. (2019) · 2019
Earlier work this paper cites.
Gpt-3 creative fiction
Branwen, G. (2020) · 2020
Cited alongside, same era.
Green ai
Schwartz, R., Dodge, J., Smith, N. A., and Etzioni, O. (2020) · 2020
Cited alongside, same era.
Reproducible and efficient benchmarks for hyperparameter optimization of neural machine translation systems
Zhang, X. and Duh, K. (2020) · 2020
Cited alongside, same era.
Evaluating large language models trained on code
Chen, M., Tworek, J., Jun, H., Yuan, Q., Pinto, H. P. d. O., Kaplan, J., Edwards, H., Burda, Y., Joseph, N., Brockman, G., et al. (2021) · 2021
Cited alongside, same era.
Training verifiers to solve math word problems
Cobbe, K., Kosaraju, V., Bavarian, M., Chen, M., Jun, H., Kaiser, L., Plappert, M., Tworek, J., Hilton, J., Nakano, R., et al. (2021) · 2021
Cited alongside, same era.
Leveraging natural supervision for language representation learning and generation
Chen, M. (2022) · 2022
Later among the works it cites.
Holistic evaluation of language models
Liang, P., Bommasani, R., Lee, T., Tsipras, D., Soylu, D., Yasunaga, M., Zhang, Y., Narayanan, D., Wu, Y., Kumar, A., et al. (2022) · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C. L., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., et al. (2022) · 2022
Later among the works it cites.
Synchromesh: Reliable code generation from pre-trained language models
Poesia, G., Polozov, O., Le, V., Tiwari, A., Soares, G., Meek, C., and Gulwani, S. (2022) · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Liu, P., Yuan, W., Fu, J., Jiang, Z., Hayashi, H., and Neubig, G. (2021) · 2021
Cited alongside, same era.
An empirical study on hyperparameter optimization for fine-tuning pre-trained language models
Liu, X. and Wang, C. (2021) · 2021
Cited alongside, same era.
Gender and representation bias in gpt-3 generated stories
Lucy, L. and Bamman, D. (2021) · 2021
Cited alongside, same era.
Reframing instructional prompts to gptk’s language
Mishra, S., Khashabi, D., Baral, C., Choi, Y., and Hajishirzi, H. (2021) · 2021
Cited alongside, same era.
Economic hyperparameter optimization with blended search strategy
Wang, C., Wu, Q., Huang, S., and Saied, A. (2021) · 2021
Cited alongside, same era.
Frugal optimization for cost-related hyperparameters
Wu, Q., Wang, C., and Huang, S. (2021) · 2021
Cited alongside, same era.
Measuring coding challenge competence with apps
Hendrycks, D., Basart, S., Kadavath, S., Mazeika, M., Arora, A., Guo, E., Burns, C., Puranik, S., He, H., Song, D., et al. (2021a)
Cited in the paper.
Shieh, J. (2022) · 2022
Later among the works it cites.
Codexdb: Synthesizing code for query processing from natural language instructions using gpt-3 codex
Trummer, I. (2022) · 2022
Later among the works it cites.
Super-naturalinstructions: Generalization via declarative instructions on 1600+ nlp tasks
Wang, Y., Mishra, S., Alipoormolabashi, P., Kordi, Y., Mirzaei, A., Arunkumar, A., Ashok, A., Dhanasekaran, A. S., Naik, A., Stap, D., et al. (2022) · 2022
Later among the works it cites.
Don’t be so sure! boosting asr decoding via confidence relaxation
Wullach, T. and Chazan, S. E. (2022) · 2022
Later among the works it cites.
Solving math word problems concerning systems of equations with gpt-3
Zong, M. and Krishnamachari, B. (2022) · 2022
Later among the works it cites.
OpenAI API
(2023) · 2023
Closest in time.
Gpt-4 technical report
OpenAI (2023) · 2023
Closest in time.