Fetching the paper…
Reading the bibliography…
Recent advances in Large Language Models (LLMs) have intensified the debate surrounding the fundamental nature of their reasoning capabilities.
2017
Earlier work this paper cites.
D. Gunning and D. Aha, Darpa’s explainable artificial intelligence (xai) program, AI magazine 40
2019
Earlier work this paper cites.
S. Mei, T. Misiakiewicz, and A. Montanari, Mean-field theory of two-layers neural networks: dimension-free bounds and kernel limit, in Conference on learning theory (PMLR, 2019) pp. 2388–2464
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Blumenfeld, D. Gilboa, and D. Soudry, A mean field theory of quantized deep networks: The quantization-depth trade-off, Advances in Neural Information Processing Systems 32
2019
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
B. Wang, X. Jiang, G. Huo, C. Su, D. Yan, and Z. Zheng, Key-point interpolation: A sparse data interpolation algorithm based on b-splines, in Journal of Physics: Conference Series , Vol. 2068 (IOP Publishing, 2021) p. 012010
2021
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou, et al. , Chain-of-thought prompting elicits reasoning in large language models, Advances in neural information processing systems 35
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
K. Valmeekam, A. Olmo, S. Sreedharan, and S. Kambhampati, Large language models still can’t plan (a benchmark for llms on planning and reasoning about change), in NeurIPS 2022 Foundation Models for Decision Making Workshop (2022)
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
R. OpenAI, Gpt-4 technical report. arxiv 2303.08774, View in Article 2
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
N. Dziri, X. Lu, M. Sclar, X. L. Li, L. Jiang, B. Y. Lin, S. Welleck, P. West, C. Bhagavatula, R. Le Bras, et al. , Faith and fate: Limits of transformers on compositionality, Advances in Neural Information Processing Systems 36
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
S. Tovey, S. Krippendorf, K. Nikolaou, and C. Holm, Towards a phenomenological understanding of neural networks: data, Machine Learning: Science and Technology 4
2023
Cited alongside, same era.
2023
Cited alongside, same era.
C. Zheng, H. Zhou, F. Meng, J. Zhou, and M. Huang, Large language models are not robust multiple choice selectors, in The Twelfth International Conference on Learning Representations (2023)
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
A. Golgoon, K. Filom, and A. Ravi Kannan, Mechanistic interpretability of large language models with applications to the financial services industry, in Proceedings of the 5th ACM International Conference on AI in Finance (2024) pp. 660–668
2024
Later among the works it cites.
2024
Later among the works it cites.
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan, Tree of thoughts: Deliberate problem solving with large language models, Advances in Neural Information Processing Systems 36
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.