Fetching the paper…
Reading the bibliography…
We explore an evolutionary search strategy for scaling inference time compute in Large Language Models.
The Nature of Human Intelligence
J. P. Guilford · 1967
Earlier work this paper cites.
Computers and Intractability: A Guide to the Theory of NP Completeness
M. R. Garey and D. S. Johnson · 1979
Earlier work this paper cites.
Genetic Algorithms in Search, Optimization, and Machine Learning
D. E. Golberg · 1989
Earlier work this paper cites.
Distributed genetic algorithms for function optimization
R. Tanese · 1989
Earlier work this paper cites.
A note on Boltzmann tournament selection for genetic algorithms and population-oriented simulated annealing
D. E. Goldberg · 1990
Earlier work this paper cites.
Adaptation in Natural and Artificial Systems
J. H. Holland · 1992
Earlier work this paper cites.
A survey of parallel genetic algorithms
E. Cantú-Paz et al · 1998
Earlier work this paper cites.
An Introduction to Genetic Algorithms
M. Mitchell · 1998
Earlier work this paper cites.
Hide and seek: An introduction to steganography
N. Provos and P. Honeyman · 2003
Earlier work this paper cites.
Levenshtein distance, sequence comparison and biological database search
B. Berger, M. S. Waterman, and Y. W. Yu · 2020
Earlier work this paper cites.
Training verifiers to solve math word problems
K. Cobbe, V. Kosaraju, M. Bavarian, M. Chen, H. Jun, L. Kaiser, M. Plappert, J. Tworek, J. Hilton, R. Nakano, et al · 2021
Earlier work this paper cites.
Constitutional AI: Harmlessness from AI feedback
Y. Bai, S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, et al · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa · 2022
Earlier work this paper cites.
CodeRL: Mastering code generation through pretrained models and deep reinforcement learning
H. Le, Y. Wang, A. D. Gotmare, S. Savarese, and S. C. H. Hoi · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou, et al · 2022
Earlier work this paper cites.
Promptbreeder: Self-referential self-improvement via prompt evolution
C. Fernando, D. Banarse, H. Michalewski, S. Osindero, and T. Rocktäschel · 2023
Cited alongside, same era.
Connecting large language models with evolutionary algorithms yields powerful prompt optimizers
Q. Guo, R. Wang, J. Guo, B. Li, K. Song, X. Tan, G. Liu, J. Bian, and Y. Yang · 2023
Cited alongside, same era.
Language models can solve computer tasks
G. Kim, P. Baldi, and S. McAleer · 2023
Cited alongside, same era.
Evolution through large models
J. Lehman, J. Gordon, S. Jain, K. Ndousse, C. Yeh, and K. O. Stanley · 2023
Cited alongside, same era.
H. Lightman, V. Kosaraju, Y. Burda, H. Edwards, B. Baker, T. Lee, J. Leike, J. Schulman, I. Sutskever, and K. Cobbe · 2023
Evolving code with a large language model
E. Hemberg, S. Moskal, and U.-M. O’Reilly · 2024
Later among the works it cites.
Prover-verifier games improve legibility of LLM outputs
J. H. Kirchner, Y. Chen, H. Edwards, J. Leike, N. McAleese, and Y. Burda · 2024
Later among the works it cites.
Improving LLM reasoning through scaling inference computation with collaborative verification
Z. Liang, Y. Liu, T. Niu, X. Zhang, Y. Zhou, and S. Yavuz · 2024
Later among the works it cites.
Large language models as evolutionary optimizers
S. Liu, C. Chen, X. Qu, K. Tang, and Y.-S. Ong · 2024
Later among the works it cites.
Self-refine: Iterative refinement with self-feedback
A. Madaan, N. Tandon, P. Gupta, S. Hallinan, L. Gao, S. Wiegreffe, U. Alon, N. Dziri, S. Prabhumoye, Y. Yang, et al · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Fully autonomous programming with large language models
V. Liventsev, A. Grishina, A. Härmä, and L. Moonen · 2023
Cited alongside, same era.
Generative agents: Interactive simulacra of human behavior
J. S. Park, J. O’Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein · 2023
Cited alongside, same era.
Self-consistency improves chain of thought reasoning in language models
X. Wang, J. Wei, D. Schuurmans, Q. V. Le, E. H. Chi, S. Narang, A. Chowdhery, and D. Zhou · 2023
Cited alongside, same era.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao, D. Yu, J. Zhao, I. Shafran, T. L. Griffiths, Y. Cao, and K. Narasimhan · 2023
Cited alongside, same era.
ALGO: Synthesizing algorithmic programs with LLM-generated oracle verifiers
K. Zhang, D. Wang, J. Xia, W. Y. Wang, and L. Li · 2023
Cited alongside, same era.
Large language model-based evolutionary optimizer: Reasoning with elitism
S. Brahmachary, S. M. Joshi, A. Panda, K. Koneripalli, A. K. Sagotra, H. Patel, A. Sharma, A. D. Jagtap, and K. Kalyanaraman · 2024
Cited alongside, same era.
Large language monkeys: Scaling inference compute with repeated sampling
B. Brown, J. Juravsky, R. Ehrlich, R. Clark, Q. V. Le, C. Ré, and A. Mirhoseini · 2024
Cited alongside, same era.
Mathematical discoveries from program search with large language models
B. Romera-Paredes, M. Barekatain, A. Novikov, M. Balog, M. P. Kumar, E. Dupont, F. J. Ruiz, J. S. Ellenberg, P. Wang, O. Fawzi, et al · 2024
Later among the works it cites.
Rewarding progress: Scaling automated process verifiers for LLM reasoning
A. Setlur, C. Nagpal, A. Fisch, X. Geng, J. Eisenstein, R. Agarwal, A. Agarwal, J. Berant, and A. Kumar · 2024
Later among the works it cites.
Reflexion: Language agents with verbal reinforcement learning
N. Shinn, F. Cassano, A. Gopinath, K. Narasimhan, and S. Yao · 2024
Later among the works it cites.
Scaling LLM test-time compute optimally can be more effective than scaling model parameters
C. Snell, J. Lee, K. Xu, and A. Kumar · 2024
Later among the works it cites.
Z. Wang, Y. Li, Y. Wu, L. Luo, L. Hou, H. Yu, and J. Shang · 2024
Later among the works it cites.
Travelplanner: A benchmark for real-world planning with language agents
J. Xie, K. Zhang, J. Chen, T. Zhu, R. Lou, Y. Tian, Y. Xiao, and Y. Su · 2024
Later among the works it cites.
ReEvo: Large language models as hyper-heuristics with reflective evolution
H. Ye, J. Wang, Z. Cao, F. Berto, C. Hua, H. Kim, J. Park, and G. Song · 2024
Later among the works it cites.
EvoAgent: Towards automatic multi-agent generation via evolutionary algorithms
S. Yuan, K. Song, J. Chen, X. Tan, D. Li, and D. Yang · 2024
Later among the works it cites.
NATURAL PLAN: Benchmarking LLMs on natural language planning
H. S. Zheng, S. Mishra, H. Zhang, X. Chen, M. Chen, A. Nova, L. Hou, H.-T. Cheng, Q. V. Le, E. H. Chi, et al · 2024
Later among the works it cites.