Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated remarkable abilities across various language tasks, but solving complex reasoning problems remains a significant challenge.
A cognitive process theory of writing
Flower, L · 1981
Earlier work this paper cites.
A Theoretical Framework , pp. 65–96
Amabile, T. M · 1983
Earlier work this paper cites.
Heuristic and analytic processes in reasoning
Evans, J. S. B. T · 1984
Earlier work this paper cites.
The Architecture of Complexity , pp. 457–476
Simon, H. A · 1991
Earlier work this paper cites.
Monte-carlo tree search: a new framework for game ai
Chaslot, G., Bakkes, S., Szita, I., and Spronck, P · 2008
Earlier work this paper cites.
Thinking, Fast and Slow
Kahneman, D · 2011
Earlier work this paper cites.
Information-theoretic regret bounds for gaussian process optimization in the bandit setting
Srinivas, N., Krause, A., Kakade, S. M., and Seeger, M. W · 2011
Earlier work this paper cites.
A survey of monte carlo tree search methods
Browne, C. B., Powley, E., Whitehouse, D., Lucas, S. M., Cowling, P. I., Rohlfshagen, P., Tavener, S., Perez, D., Samothrakis, S., and Colton, S · 2012
Earlier work this paper cites.
Large language models are zero-shot reasoners
Kojima, T., Gu, S. S., Reid, M., Matsuo, Y., and Iwasawa, Y · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Xia, F., Chi, E., Le, Q. V., Zhou, D., et al · 2022
Earlier work this paper cites.
Automatic chain of thought prompting in large language models
Zhang, Z., Zhang, A., Li, M., and Smola, A · 2022
Earlier work this paper cites.
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., et al · 2023
Earlier work this paper cites.
Chen, W., Ma, X., Wang, X., and Cohen, W. W · 2023
Cited alongside, same era.
Jiang, A. Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D. S., de las Casas, D., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., Lavaud, L. R., Lachaux, M.-A., Stock, P., Scao, T. L., Lavril, T., Wang, T., Lacroix, T., and Sayed, W. E · 2023
Cited alongside, same era.
Large language models are zero-shot reasoners
Kojima, T., Gu, S. S., Reid, M., Matsuo, Y., and Iwasawa, Y · 2023
Cited alongside, same era.
Let’s verify step by step, 2023
Lightman, H., Kosaraju, V., Burda, Y., Edwards, H., Baker, B., Lee, T., Leike, J., Schulman, J., Sutskever, I., and Cobbe, K · 2023
Cited alongside, same era.
Introducing meta llama 3: The most capable openly available llm to date
at Meta, A · 2024
Closest in time.
Everything of thoughts: Defying the law of penrose triangle for thought generation, 2024
Ding, R., Zhang, C., Wang, L., Xu, Y., Ma, M., Zhang, W., Qin, S., Rajmohan, S., Lin, Q., and Zhang, D · 2024
Closest in time.
Chatglm: A family of large language models from glm-130b to glm-4 all tools
GLM, T., :, Zeng, A., Xu, B., Wang, B., Zhang, C., Yin, D., Zhang, D., Rojas, D., Feng, G., Zhao, H., Lai, H., Yu, H., Wang, H., Sun, J., Zhang, J., Cheng, J., Gui, J., Tang, J., Zhang, J., Sun, J., Li, J., Zhao, L., Wu, L., Zhong, L., Liu, M., Huang, M., Zhang, P., Zheng, Q., Lu, R., Duan, S., Zhang, S., Cao, S., Yang, S., Tam, W. L., Zhao, W., Liu, X., Xia, X., Zhang, X., Gu, X., Lv, X., Liu, X., Liu, X., Yang, X., Song, X., Zhang, X., An, Y., Xu, Y., Niu, Y., Yang, Y., Li, Y., Bai, Y., Dong, Y., Qi, Z., Wang, Z., Yang, Z., Du, Z., Hou, Z., and Wang, Z · 2024
Closest in time.
Chain of code: Reasoning with a language model-augmented code emulator
Li, C., Liang, J., Zeng, A., Chen, X., Hausman, K., Sadigh, D., Levine, S., Fei-Fei, L., Xia, F., and Ichter, B · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Madaan, A., Tandon, N., Gupta, P., Hallinan, S., Gao, L., Wiegreffe, S., Alon, U., Dziri, N., Prabhumoye, S., Yang, Y., Gupta, S., Majumder, B. P., Hermann, K., Welleck, S., Yazdanbakhsh, A., and Clark, P · 2023
Cited alongside, same era.
Tree prompting: Efficient task adaptation without fine-tuning
Morris, J. X., Singh, C., Rush, A. M., Gao, J., and Deng, Y · 2023
Cited alongside, same era.
Nguyen, H. H., Liu, Y., Zhang, C., Zhang, T., and Yu, P. S · 2023
Cited alongside, same era.
Skeleton-of-thought: Large language models can do parallel decoding
Ning, X., Lin, Z., Zhou, Z., Wang, Z., Yang, H., and Wang, Y · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozière, B., Goyal, N., Hambro, E., Azhar, F., et al · 2023
Cited alongside, same era.
Self-consistency improves chain of thought reasoning in language models
Wang, X., Wei, J., Schuurmans, D., Le, Q., Chi, E., Narang, S., Chowdhery, A., and Zhou, D · 2023
Cited alongside, same era.
Verify-and-edit: A knowledge-enhanced chain-of-thought framework
Zhao, R., Li, X., Joty, S., Qin, C., and Bing, L · 2023
Cited alongside, same era.
Least-to-most prompting enables complex reasoning in large language models
Zhou, D., Schärli, N., Hou, L., Wei, J., Scales, N., Wang, X., Schuurmans, D., Cui, C., Bousquet, O., Le, Q., and Chi, E · 2023
Cited alongside, same era.
Closest in time.
Openai o1 system card
OpenAI · 2024
Closest in time.
Algorithm of thoughts: Enhancing exploration of ideas in large language models
Sel, B., Al-Tawaha, A., Khattar, V., Jia, R., and Jin, M · 2024
Closest in time.
Scaling llm test-time compute optimally can be more effective than scaling model parameters
Snell, C., Lee, J., Xu, K., and Kumar, A · 2024
Closest in time.
Buffer of thoughts: Thought-augmented reasoning with large language models
Yang, L., Yu, Z., Zhang, T., Cao, S., Xu, M., Zhang, W., Gonzalez, J. E., and Cui, B · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Yao, S., Yu, D., Zhao, J., Shafran, I., Griffiths, T., Cao, Y., and Narasimhan, K · 2024
Closest in time.
Zhang, D., Huang, X., Zhou, D., Li, Y., and Ouyang, W · 2024
Closest in time.
Aime 2024
AI-MO · 2025
Closest in time.
rstar-math: Small llms can master math reasoning with self-evolved deep thinking, 2025
Guan, X., Zhang, L. L., Liu, Y., Shang, N., Sun, Y., Zhu, Y., Yang, F., and Yang, M · 2025
Closest in time.