Fetching the paper…
Reading the bibliography…
The reasoning abilities of large language models (LLMs) are the topic of a growing body of research in AI and cognitive science.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
Adversarial nli: A new benchmark for natural language understanding
Yixin Nie, Adina Williams, Emily Dinan, Mohit Bansal, Jason Weston, and Douwe Kiela. 2019 · 1910
Earlier work this paper cites.
Über den Begriff der logischen Folgerung
Alfred Tarski. 1936 · 1936
Earlier work this paper cites.
Semantical analysis of modal logic I. Normal modal propositional calculi
Saul A Kripke. 1963 · 1963
Earlier work this paper cites.
A theory of conditionals
Robert C Stalnaker. 1968 · 1968
Earlier work this paper cites.
A semantic analysis of conditional logic
Robert C Stalnaker and Richmond H Thomason. 1970 · 1970
Earlier work this paper cites.
Semantic theory
Jerrold J Katz. 1972 · 1972
Earlier work this paper cites.
Free choice permission
Hans Kamp. 1973 · 1973
Earlier work this paper cites.
Counterfactuals and two kinds of expected utility
Allan Gibbard and William L Harper. 1981 · 1981
Earlier work this paper cites.
The notional category of modality
Angelika Kratzer. 1981 · 1981
Earlier work this paper cites.
A counterexample to modus ponens
Vann McGee. 1985 · 1985
Earlier work this paper cites.
On the complexity of conditional logics
Nir Friedman and Joseph Y Halpern. 1994 · 1994
Earlier work this paper cites.
On conditionals
Dorothy Edgington. 1995 · 1995
Earlier work this paper cites.
Reasoning about knowledge
Ronald Fagin, Joseph Y Halpern, Yoram Moses, and Moshe Y Vardi. 1995 · 1995
Earlier work this paper cites.
Entailment, intensionality and text understanding
Cleo Condoravdi, Dick Crouch, Valeria de Paiva, Reinhard Stolle, and Daniel G Bobrow. 2003 · 2003
Earlier work this paper cites.
If: Supposition, pragmatics, and dual processes
Jonathan St BT Evans and David E Over. 2004 · 2004
Earlier work this paper cites.
Dividing things up: The semantics of or and the modal/ or interaction
Mandy Simons. 2005 · 2005
Earlier work this paper cites.
A brief history of natural logic
Johan van Benthem. 2008 · 2008
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2020 · 2009
Earlier work this paper cites.
An extended model of natural logic
Bill MacCartney and Christopher D Manning. 2009 · 2009
Earlier work this paper cites.
Modality
Paul Portner. 2009 · 2009
Earlier work this paper cites.
Recognizing textual entailment: Rational, evaluation and approaches–erratum
Ido Dagan, Bill Dolan, Bernardo Magnini, and Dan Roth. 2010 · 2010
Earlier work this paper cites.
Connectives without truth-tables
Nathan Klinedinst and Daniel Rothschild. 2012 · 2012
Earlier work this paper cites.
Modals and Conditionals
Angelika Kratzer. 2012 · 2012
Earlier work this paper cites.
Proofwriter: Generating implications, proofs, and abductive statements over natural language
Oyvind Tafjord, Bhavana Dalvi Mishra, and Peter Clark. 2020 · 2012
Cited alongside, same era.
A counterexample to modus tollens
Seth Yalcin. 2012 · 2012
Cited alongside, same era.
The probabilities of conditionals revisited
Igor Douven and Sara Verbrugge. 2013 · 2013
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning. 2015 · 2015
Cited alongside, same era.
Covert across-the-board movement revisited: Free choice and the scope of modals
Marie-Christine Meyer and Uli Sauerland. 2017 · 2017
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Sparks of Artificial General Intelligence: Early experiments with GPT-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al. 2023 · 2023
Later among the works it cites.
Learning to teach large language models logical reasoning
Meiqi Chen, Yubo Ma, Kaitao Song, Yixin Cao, Yan Zhang, and Dongsheng Li. 2023 · 2023
Later among the works it cites.
Selection-inference: Exploiting large language models for interpretable logical reasoning
Antonia Creswell, Murray Shanahan, and Irina Higgins. 2023 · 2023
Later among the works it cites.
Towards reasoning in large language models: A survey
Jie Huang and Kevin Chen-Chuan Chang. 2023 · 2023
Later among the works it cites.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
Sluicing on free choice
Melissa Fusco. 2019 · 2019
Cited alongside, same era.
Inherent Disagreements in Human Textual Inferences
Ellie Pavlick and Tom Kwiatkowski. 2019 · 2019
Cited alongside, same era.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al. 2023 · 2020
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, et al. 2021 · 2021
Cited alongside, same era.
The Logic of Conditionals
Paul Egré and Hans Rott. 2021 · 2021
Cited alongside, same era.
Logiqa: a challenge dataset for machine reading comprehension with logical reasoning
Jian Liu, Leyang Cui, Hanmeng Liu, Dandan Huang, Yile Wang, and Yue Zhang. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
LAMBADA: Backward chaining for automated reasoning in natural language
Mehran Kazemi, Najoung Kim, Deepti Bhatia, Xin Xu, and Deepak Ramachandran. 2023 · 2023
Later among the works it cites.
LINC: A neurosymbolic approach for logical reasoning by combining language models with first-order logic provers
Theo Olausson, Alex Gu, Ben Lipkin, Cedegao Zhang, Armando Solar-Lezama, Joshua Tenenbaum, and Roger Levy. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
Logic-LM: Empowering large language models with symbolic solvers for faithful logical reasoning
Liangming Pan, Alon Albalak, Xinyi Wang, and William Wang. 2023 · 2023
Later among the works it cites.
Certified deductive reasoning with language models
Gabriel Poesia, Kanishk Gandhi, Eric Zelikman, and Noah D Goodman. 2023 · 2023
Later among the works it cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al. 2023 · 2023
Later among the works it cites.
Language models are greedy reasoners: A systematic formal analysis of chain-of-thought
Abulhair Saparov and He He. 2023 · 2023
Later among the works it cites.
Testing the general deductive reasoning capacity of large language models using ood examples
Abulhair Saparov, Richard Yuanzhe Pang, Vishakh Padmakumar, Nitish Joshi, Mehran Kazemi, Najoung Kim, and He He. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Are language models worse than humans at following prompts? it’s complicated
Albert Webson, Alyssa Marie Loo, Qinan Yu, and Ellie Pavlick. 2023 · 2023
Later among the works it cites.
Fangzhi Xu, Qika Lin, Jiawei Han, Tianzhe Zhao, Jun Liu, and Erik Cambria. 2023 · 2023
Later among the works it cites.
Satlm: Satisfiability-aided language models using declarative prompting
Xi Ye, Qiaochu Chen, Isil Dillig, and Greg Durrett. 2023 · 2023
Later among the works it cites.
Chatbot arena: An open platform for evaluating llms by human preference
Wei-Lin Chiang, Lianmin Zheng, Ying Sheng, Anastasios Nikolas Angelopoulos, Tianle Li, Dacheng Li, Hao Zhang, Banghua Zhu, Michael Jordan, Joseph E Gonzalez, et al. 2024 · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al. 2024 · 2024
Closest in time.
Modal logic
James Garson. 2024 · 2024
Closest in time.
Gemma 2: Improving open language models at a practical size
Gemma Team. 2024 · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al. 2024 · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Machel Reid, Nikolay Savinov, Denis Teplyashin, Dmitry Lepikhin, Timothy Lillicrap, Jean-baptiste Alayrac, Radu Soricut, Angeliki Lazaridou, Orhan Firat, Julian Schrittwieser, et al. 2024 · 2024
Closest in time.
A & b== b & a: Triggering logical reasoning failures in large language models
Yuxuan Wan, Wenxuan Wang, Yiliu Yang, Youliang Yuan, Jen-tse Huang, Pinjia He, Wenxiang Jiao, and Michael R Lyu. 2024 · 2024
Closest in time.