Fetching the paper…
Reading the bibliography…
Large language models (LLMs) have demonstrated limitations in handling combinatorial optimization problems involving long-range reasoning, partially due to causal hallucinations and huge search space.
Situations, actions, and causal laws
John McCarthy et al · 1963
Earlier work this paper cites.
Causal analysis
David R Heise · 1975
Earlier work this paper cites.
Time complexity of the towers of hanoi problem
Colin Gerety and Paul Cull · 1986
Earlier work this paper cites.
The computational complexity of propositional strips planning
Tom Bylander · 1994
Earlier work this paper cites.
Artificial intelligence: A modern approach
Stuart Russell and Peter Norvig · 1995
Earlier work this paper cites.
The art of causal conjecture
Glenn Shafer · 1996
Earlier work this paper cites.
Model predictive control
Eduardo F. Camacho and Carlos Bordons (eds.) · 1999
Earlier work this paper cites.
On tests of the overall treatment effect in meta-analysis with normally distributed responses
Joachim Hartung and Guido Knapp · 2001
Earlier work this paper cites.
Theory of games and economic behavior
Robert J. Leonard · 2006
Earlier work this paper cites.
Causal reasoning through intervention
York Hagmayer, Steven A Sloman, David A Lagnado, and Michael R Waldmann · 2007
Earlier work this paper cites.
Essentials of game theory: A concise multidisciplinary introduction
Kevin Leyton-Brown and Yoav Shoham · 2008
Earlier work this paper cites.
Causality
Judea Pearl · 2009
Earlier work this paper cites.
Causal inference in statistics, social, and biomedical sciences
Guido W Imbens and Donald B Rubin · 2015
Earlier work this paper cites.
Program induction by rationale generation: Learning to solve and explain algebraic word problems
Wang Ling, Dani Yogatama, Chris Dyer, and Phil Blunsom · 2017
Earlier work this paper cites.
Elements of causal inference: foundations and learning algorithms
Jonas Peters, Dominik Janzing, and Bernhard Schölkopf · 2017
Earlier work this paper cites.
Theoretical impediments to machine learning with seven sparks from the causal revolution
Judea Pearl · 2018
Earlier work this paper cites.
Task planning in robotics: an empirical comparison of pddl-and asp-based systems
Yu-qian Jiang, Shi-qi Zhang, Piyush Khandelwal, and Peter Stone · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Qasc: A dataset for question answering via sentence composition
Tushar Khot, Peter Clark, Michal Guerquin, Peter Jansen, and Ashish Sabharwal · 2020
Earlier work this paper cites.
Collaborative storytelling with large-scale neural language models
Eric Nichols, Leo Gao, and Randy Gomez · 2020
Earlier work this paper cites.
Conformal inference of counterfactuals and individual treatment effects
Lihua Lei and Emmanuel J Candès · 2021
Earlier work this paper cites.
The individual-level surrogate threshold effect in a causal-inference setting with normally distributed endpoints
Wim Van der Elst, Ariel Alonso Abad, Hans Coppenolle, Paul Meyvisch, and Geert Molenberghs · 2021
Earlier work this paper cites.
Wenhu Chen, Xueguang Ma, Xinyi Wang, and William W Cohen · 2022
Earlier work this paper cites.
Causal inference in natural language processing: Estimation, prediction, interpretation and beyond
Amir Feder, Katherine A Keith, Emaad Manzoor, Reid Pryzant, Dhanya Sridhar, Zach Wood-Doughty, Jacob Eisenstein, Justin Grimmer, Roi Reichart, Margaret E Roberts, et al · 2022
Earlier work this paper cites.
Inner monologue: Embodied reasoning through planning with language models
Wenlong Huang, Fei Xia, Ted Xiao, Harris Chan, Jacky Liang, Pete Florence, Andy Zeng, Jonathan Tompson, Igor Mordatch, Yevgen Chebotar, et al · 2022
Cited alongside, same era.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Cited alongside, same era.
Solving quantitative reasoning problems with language models
Aitor Lewkowycz, Anders Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, et al · 2022
Cited alongside, same era.
Grokking: Generalization beyond overfitting on small algorithmic datasets
Alethea Power, Yuri Burda, Harri Edwards, Igor Babuschkin, and Vedant Misra · 2022
Cited alongside, same era.
Progprompt: Generating situated robot task plans using large language models
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg · 2023
Later among the works it cites.
Introducing mpt-7b: A new standard for open-source, commercially usable llms, 2023
MN Team et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
Beyond chain-of-thought, effective graph-of-thought reasoning in large language models
Yao Yao, Zuchao Li, and Hai Zhao · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2022
Cited alongside, same era.
On the paradox of learning to reason from data
Honghua Zhang, Liunian Harold Li, Tao Meng, Kai-Wei Chang, and Guy Van den Broeck · 2022
Cited alongside, same era.
Solving math word problems via cooperative reasoning induced language models
Xinyu Zhu, Junjie Wang, Lin Zhang, Yuxiang Zhang, Yongfeng Huang, Ruyi Gan, Jiaxing Zhang, and Yujiu Yang · 2022
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al · 2023
Cited alongside, same era.
Improving image generation with better captions
James Betker, Gabriel Goh, Li Jing, Tim Brooks, Jianfeng Wang, Linjie Li, Long Ouyang, Juntang Zhuang, Joyce Lee, Yufei Guo, et al · 2023
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2023
Cited alongside, same era.
A survey of chain of thought reasoning: Advances, frontiers and future
Zheng Chu, Jingchang Chen, Qianglong Chen, Weijiang Yu, Tao He, Haotian Wang, Weihua Peng, Ming Liu, Bing Qin, and Ting Liu · 2023
Cited alongside, same era.
Task and motion planning with large language models for object rearrangement
Yan Ding, Xiaohan Zhang, Chris Paxton, and Shiqi Zhang · 2023
Cited alongside, same era.
Yining Ye, Xin Cong, Yujia Qin, Yankai Lin, Zhiyuan Liu, and Maosong Sun · 2023
Later among the works it cites.
Causal parrots: Large language models may talk causality but are not causal
Matej Zečević, Moritz Willig, Devendra Singh Dhami, and Kristian Kersting · 2023
Later among the works it cites.
Competeai: Understanding the competition behaviors in large language model-based agents
Qinlin Zhao, Jindong Wang, Yixuan Zhang, Yiqiao Jin, Kaijie Zhu, Hao Chen, and Xing Xie · 2023
Later among the works it cites.
Take a step back: Evoking reasoning via abstraction in large language models
Huaixiu Steven Zheng, Swaroop Mishra, Xinyun Chen, Heng-Tze Cheng, Ed H Chi, Quoc V Le, and Denny Zhou · 2023
Later among the works it cites.
Llms with chain-of-thought are non-causal reasoners
Guangsheng Bao, Hongbo Zhang, Linyi Yang, Cunxiang Wang, and Yue Zhang · 2024
Closest in time.
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman · 2024
Closest in time.
Pokergpt: An end-to-end lightweight solver for multi-player texas hold’em via large language model
Chenghao Huang, Yanbo Cao, Yinlong Wen, Tao Zhou, and Yanru Zhang · 2024
Closest in time.
Rot: Enhancing large language models with reflection on search trees
Wenyang Hui, Yan Wang, Kewei Tu, and Chengyue Jiang · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al · 2024
Closest in time.
Swarmbrain: Embodied agent for real-time strategy game starcraft ii via large language models
Xiao Shao, Weifu Jiang, Fei Zuo, and Mengqing Liu · 2024
Closest in time.
Reflexion: Language agents with verbal reinforcement learning
Noah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao · 2024
Closest in time.
Language models don’t always say what they think: unfaithful explanations in chain-of-thought prompting
Miles Turpin, Julian Michael, Ethan Perez, and Samuel Bowman · 2024
Closest in time.
On the planning abilities of large language models-a critical investigation
Karthik Valmeekam, Matthew Marquez, Sarath Sreedharan, and Subbarao Kambhampati · 2024
Closest in time.
Csce: Boosting llm reasoning by simultaneous enhancing of casual significance and consistency
Kangsheng Wang, Xiao Zhang, Zizheng Guo, Tianyu Hu, and Huimin Ma · 2024
Closest in time.
Vllms provide better context for emotion understanding through common sense reasoning
Alexandros Xenos, Niki Maria Foteinopoulou, Ioanna Ntinou, Ioannis Patras, and Georgios Tzimiropoulos · 2024
Closest in time.
Measuring bargaining abilities of llms: A benchmark and a buyer-enhancement method
Tian Xia, Zhiwei He, Tong Ren, Yibo Miao, Zhuosheng Zhang, Yang Yang, and Rui Wang · 2024
Closest in time.
Hainiu Xu, Runcong Zhao, Lixing Zhu, Jinhua Du, and Yulan He · 2024
Closest in time.
Magic: Investigation of large language model powered multi-agent in cognition, adaptability, rationality and collaboration
Lin Xu, Zhiyuan Hu, Daquan Zhou, Hongyu Ren, Zhen Dong, Kurt Keutzer, See-Kiong Ng, and Jiashi Feng · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan · 2024
Closest in time.
K-level reasoning with large language models
Yadong Zhang, Shaoguang Mao, Tao Ge, Xun Wang, Yan Xia, Man Lan, and Furu Wei · 2024
Closest in time.