Fetching the paper…
Reading the bibliography…
Chain-of-Though (CoT) represents a common strategy for reasoning in Large Language Models (LLMs) by decomposing complex tasks into intermediate inference steps.
The formalization of mathematics
Hao Wang. 1954 · 1954
Earlier work this paper cites.
A machine-oriented logic based on the resolution principle
John Alan Robinson. 1965 · 1965
Earlier work this paper cites.
Explanatory unification
Philip Kitcher. 1981 · 1981
Earlier work this paper cites.
Conditional reasoning and causation
Denise D Cummins, Todd Lubart, Olaf Alksnis, and Robert Rist. 1991 · 1991
Earlier work this paper cites.
Reasoning in explanation-based decision making
Nancy Pennington and Reid Hastie. 1993 · 1993
Earlier work this paper cites.
Program induction by rationale generation: Learning to solve and explain algebraic word problems
Wang Ling, Dani Yogatama, Chris Dyer, and Phil Blunsom. 2017 · 2017
Earlier work this paper cites.
Logical reasoning in formal and everyday reasoning tasks
Hugo Bronkhorst, Gerrit Roorda, Cor Suhre, and Martin Goedhart. 2019 · 2019
Earlier work this paper cites.
DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs
Dheeru Dua, Yizhong Wang, Pradeep Dasigi, Gabriel Stanovsky, Sameer Singh, and Matt Gardner. 2019 · 2019
Earlier work this paper cites.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2019 · 2019
Earlier work this paper cites.
Premise selection in natural language mathematical texts
Deborah Ferreira and André Freitas. 2020 · 2020
Earlier work this paper cites.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, et al. 2021 · 2021
Earlier work this paper cites.
Unification-based reconstruction of multi-hop explanations for science questions
Marco Valentino, Mokanarangan Thayaparan, and André Freitas. 2021 · 2021
Earlier work this paper cites.
Neural methods for logical reasoning over knowledge graphs
Alfonso Amayuelas, Shuai Zhang, Xi Susie Rao, and Ce Zhang. 2022 · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Earlier work this paper cites.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Earlier work this paper cites.
Automatic chain of thought prompting in large language models
Zhuosheng Zhang, Aston Zhang, Mu Li, and Alex Smola. 2022 · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Cited alongside, same era.
Large language models of code fail at completing code with potential bugs
Tuan Dinh, Jinman Zhao, Samson Tan, Renato Negrinho, Leonard Lausen, Sheng Zha, and George Karypis. 2023 · 2023
Cited alongside, same era.
Comparative analysis of logic reasoning and graph neural networks for ontology-mediated query answering with a covering axiom
Olga Gerasimova, Nikita Severin, and Ilya Makarov. 2023 · 2023
Cited alongside, same era.
Faithful chain-of-thought reasoning
Qing Lyu, Shreya Havaldar, Adam Stein, Li Zhang, Delip Rao, Eric Wong, Marianna Apidianaki, and Chris Callison-Burch. 2023 · 2023
Cited alongside, same era.
Introduction to Mathematical Language Processing: Informal Proofs, Word Problems, and Supporting Tasks
Improve mathematical reasoning in language models by automated process supervision
Liangchen Luo, Yinxiao Liu, Rosanne Liu, Samrat Phatale, Harsh Lara, Yunxuan Li, Lei Shu, Yun Zhu, Lei Meng, Jiao Sun, et al. 2024 · 2024
Later among the works it cites.
Gsm-symbolic: Understanding the limitations of mathematical reasoning in large language models
Iman Mirzadeh, Keivan Alizadeh, Hooman Shahrokhi, Oncel Tuzel, Samy Bengio, and Mehrdad Farajtabar. 2024 · 2024
Later among the works it cites.
Self-refine instruction-tuning for aligning reasoning in language models
Leonardo Ranaldi and Andre Freitas. 2024b · 2024
Later among the works it cites.
Empowering multi-step reasoning across languages via program-aided language models
Leonardo Ranaldi, Giulia Pucci, Barry Haddow, and Alexandra Birch. 2024a · 2024
Later among the works it cites.
A tree-of-thoughts to broaden multi-step reasoning across languages
Leonardo Ranaldi, Giulia Pucci, Federico Ranaldi, Elena Sofia Ruzzetti, and Fabio Massimo Zanzotto. 2024b · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jordan Meadows and André Freitas. 2023 · 2023
Cited alongside, same era.
A symbolic framework for systematic evaluation of mathematical reasoning with transformers
Jordan Meadows, Marco Valentino, Damien Teney, and Andre Freitas. 2023 · 2023
Cited alongside, same era.
Logic-lm: Empowering large language models with symbolic solvers for faithful logical reasoning
Liangming Pan, Alon Albalak, Xinyi Wang, and William Yang Wang. 2023 · 2023
Cited alongside, same era.
Gpqa: A graduate-level google-proof q&a benchmark
David Rein, Betty Li Hou, Asa Cooper Stickland, Jackson Petty, Richard Yuanzhe Pang, Julien Dirani, Julian Michael, and Samuel R. Bowman. 2023 · 2023
Cited alongside, same era.
Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models
Lei Wang, Wanyu Xu, Yihuai Lan, Zhiqiang Hu, Yunshi Lan, Roy Ka-Wei Lee, and Ee-Peng Lim. 2023 · 2023
Cited alongside, same era.
Least-to-most prompting enables complex reasoning in large language models
Denny Zhou, Nathanael Schärli, Le Hou, Jason Wei, Nathan Scales, Xuezhi Wang, Dale Schuurmans, Claire Cui, Olivier Bousquet, Quoc Le, and Ed Chi. 2023 · 2023
Cited alongside, same era.
Aaron Grattafiori et al. 2024 · 2024
Cited alongside, same era.
Flare: Faithful logic-aided reasoning and exploration
Erik Arakelyan, Pasquale Minervini, Pat Verga, Patrick Lewis, and Isabelle Augenstein. 2024 · 2024
Cited alongside, same era.
Later among the works it cites.
Zhi Rui Tam, Cheng-Kuang Wu, Yi-Lin Tsai, Chieh-Yen Lin, Hung-yi Lee, and Yun-Nung Chen. 2024 · 2024
Later among the works it cites.
Language models don’t always say what they think: unfaithful explanations in chain-of-thought prompting
Miles Turpin, Julian Michael, Ethan Perez, and Samuel Bowman. 2024 · 2024
Later among the works it cites.
On the nature of explanation: An epistemological-linguistic perspective for explanation-based natural language inference
Marco Valentino and André Freitas. 2024 · 2024
Later among the works it cites.
Faithful logical reasoning via symbolic chain-of-thought
Jundong Xu, Hao Fei, Liangming Pan, Qian Liu, Mong-Li Lee, and Wynne Hsu. 2024 · 2024
Later among the works it cites.
An Yang, Baosong Yang, Binyuan Hui, Bo Zheng, Bowen Yu, Chang Zhou, Chengpeng Li, Chengyuan Li, Dayiheng Liu, Fei Huang, et al. 2024 · 2024
Later among the works it cites.
Satlm: Satisfiability-aided language models using declarative prompting
Xi Ye, Qiaochu Chen, Isil Dillig, and Greg Durrett. 2024 · 2024
Later among the works it cites.
Dissociation of faithful and unfaithful reasoning in LLMs
Evelyn Yee, Alice Li, Chenyu Tang, Yeon Ho Jung, Ramamohan Paturi, and Leon Bergen. 2024 · 2024
Later among the works it cites.
Take a step back: Evoking reasoning via abstraction in large language models
Huaixiu Steven Zheng, Swaroop Mishra, Xinyun Chen, Heng-Tze Cheng, Ed H. Chi, Quoc V Le, and Denny Zhou. 2024 · 2024
Later among the works it cites.
Achieving> 97% on gsm8k: Deeply understanding the problems makes LLMs perfect reasoners
Qihuang Zhong, Kang Wang, Ziyang Xu, Juhua Liu, Liang Ding, Bo Du, and Dacheng Tao. 2024 · 2024
Later among the works it cites.
Multilingual reasoning via self-training
Leonardo Ranaldi and Giulia Pucci. 2025 · 2025
Closest in time.
Eliciting critical reasoning in retrieval-augmented generation via contrastive explanations
Leonardo Ranaldi, Marco Valentino, and Andre Freitas. 2025b · 2025
Closest in time.
Are NLP models really able to solve simple math word problems?
Arkil Patel, Satwik Bhattamishra, and Navin Goyal. 2021 · 2094
Closest in time.