Fetching the paper…
Reading the bibliography…
This survey paper proposes a clearer view of natural language reasoning in the field of Natural Language Processing (NLP), both conceptually and practically.
Dictionary of Philosophy
P. A. Angeles · 1981
Earlier work this paper cites.
Informal logic and the theory of reasoning
M. A. Finocchiaro · 1984
Earlier work this paper cites.
Epistemology and cognition
A. I. Goldman · 1986
Earlier work this paper cites.
Critical thinking as argument analysis
T. Govier · 1989
Earlier work this paper cites.
What is reasoning? what is an argument?
D. N. Walton · 1990
Earlier work this paper cites.
Reasoning and the logic of things: The Cambridge conferences lectures of 1898
C. S. Peirce · 1992
Earlier work this paper cites.
CYC: A large-scale investment in knowledge infrastructure
D. B. Lenat · 1995
Earlier work this paper cites.
The dictionary of philosophy
D. D. Runes · 2001
Earlier work this paper cites.
Conceptnet—a practical commonsense reasoning tool-kit
H. Liu and P. Singh · 2004
Earlier work this paper cites.
The Oxford Dictionary of Philosophy
S. Blackburn · 2008
Earlier work this paper cites.
The icwsm 2009 spinn3r dataset
K. Burton, A. Java, and I. Soboroff · 2009
Earlier work this paper cites.
Robust disambiguation of named entities in text
J. Hoffart, M. A. Yosef, I. Bordino, H. Fürstenau, M. Pinkal, M. Spaniol, B. Taneva, S. Thater, and G. Weikum · 2011
Earlier work this paper cites.
Thinking, fast and slow
D. Kahneman · 2011
Earlier work this paper cites.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
M. Roemmele, C. A. Bejan, and A. S. Gordon · 2011
Earlier work this paper cites.
Recognizing Textual Entailment: Models and Applications
I. Dagan, D. Roth, M. Sammons, and F. M. Zanzotto · 2013
Earlier work this paper cites.
A concise introduction to logic
P. J. Hurley · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
S. R. Bowman, G. Angeli, C. Potts, and C. D. Manning · 2015
Earlier work this paper cites.
A corpus and cloze evaluation for deeper understanding of commonsense stories
N. Mostafazadeh, N. Chambers, X. He, D. Parikh, D. Batra, L. Vanderwende, P. Kohli, and J. F. Allen · 2016
Earlier work this paper cites.
Towards ai-complete question answering: A set of prerequisite toy tasks
J. Weston, A. Bordes, S. Chopra, and T. Mikolov · 2016
Earlier work this paper cites.
inference
T. E. o. E. Britannica · 2017
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
R. Speer, J. Chin, and C. Havasi · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
e-snli: Natural language inference with natural language explanations
O. Camburu, T. Rocktäschel, T. Lukasiewicz, and P. Blunsom · 2018
Earlier work this paper cites.
Think you have solved question answering? try arc, the AI2 reasoning challenge
P. Clark, I. Cowhey, O. Etzioni, T. Khot, A. Sabharwal, C. Schoenick, and O. Tafjord · 2018
Earlier work this paper cites.
XNLI: evaluating cross-lingual sentence representations
A. Conneau, R. Rinott, G. Lample, A. Williams, S. R. Bowman, H. Schwenk, and V. Stoyanov · 2018
Earlier work this paper cites.
The argument reasoning comprehension task: Identification and reconstruction of implicit warrants
I. Habernal, H. Wachsmuth, I. Gurevych, and B. Stein · 2018
Earlier work this paper cites.
Scitail: A textual entailment dataset from science question answering
T. Khot, A. Sabharwal, and P. Clark · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? A new dataset for open book question answering
T. Mihaylov, P. Clark, T. Khot, and A. Sabharwal · 2018
Earlier work this paper cites.
Hypothesis only baselines in natural language inference
A. Poliak, J. Naradowsky, A. Haldar, R. Rudinger, and B. V. Durme · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever · 2018
Earlier work this paper cites.
Event2mind: Commonsense inference on events, intents, and reactions
H. Rashkin, M. Sap, E. Allaway, N. A. Smith, and Y. Choi · 2018
Earlier work this paper cites.
Interpretation of natural language rules in conversational machine reading
M. Saeidi, M. Bartolo, P. S. H. Lewis, S. Singh, T. Rocktäschel, M. Sheldon, G. Bouchard, and S. Riedel · 2018
Earlier work this paper cites.
Reasoning about actions and state changes by injecting commonsense knowledge
N. Tandon, B. Dalvi, J. Grus, W. Yih, A. Bosselut, and P. Clark · 2018
Earlier work this paper cites.
Performance impact caused by hidden bias of training data for recognizing textual entailment
M. Tsuchiya · 2018
Earlier work this paper cites.
Constructing datasets for multi-hop reading comprehension across documents
J. Welbl, P. Stenetorp, and S. Riedel · 2018
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
A. Williams, N. Nangia, and S. R. Bowman · 2018
Earlier work this paper cites.
Hotpotqa: A dataset for diverse, explainable multi-hop question answering
Z. Yang, P. Qi, S. Zhang, Y. Bengio, W. W. Cohen, R. Salakhutdinov, and C. D. Manning · 2018
Earlier work this paper cites.
SWAG: A large-scale adversarial dataset for grounded commonsense inference
R. Zellers, Y. Bisk, R. Schwartz, and Y. Choi · 2018
Earlier work this paper cites.
Understanding dataset design choices for multi-hop reasoning
J. Chen and G. Durrett · 2019
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
J. Devlin, M. Chang, K. Lee, and K. Toutanova · 2019
Earlier work this paper cites.
Cosmos QA: machine reading comprehension with contextual commonsense reasoning
L. Huang, R. L. Bras, C. Bhagavatula, and Y. Choi · 2019
Earlier work this paper cites.
Avoiding reasoning shortcuts: Adversarial evaluation, training, and model development for multi-hop QA
Y. Jiang and M. Bansal · 2019
Earlier work this paper cites.
Attention is (not) all you need for commonsense reasoning
T. Klein and M. Nabi · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
T. Kwiatkowski, J. Palomaki, O. Redfield, M. Collins, A. P. Parikh, C. Alberti, D. Epstein, I. Polosukhin, J. Devlin, K. Lee, K. Toutanova, L. Jones, M. Kelcey, M. Chang, A. M. Dai, J. Uszkoreit, Q. Le, and S. Petrov · 2019
Earlier work this paper cites.
Reasoning over paragraph effects in situations
K. Lin, O. Tafjord, P. Clark, and M. Gardner · 2019
Earlier work this paper cites.
Compositional questions do not necessitate multi-hop reasoning
S. Min, E. Wallace, S. Singh, M. Gardner, H. Hajishirzi, and L. Zettlemoyer · 2019
Earlier work this paper cites.
Multi-hop reading comprehension through question decomposition and rescoring
S. Min, V. Zhong, L. Zettlemoyer, and H. Hajishirzi · 2019
Earlier work this paper cites.
Counterfactual story reasoning and generation
L. Qin, A. Bosselut, A. Holtzman, C. Bhagavatula, E. Clark, and Y. Choi · 2019
Earlier work this paper cites.
Dynamically fused graph network for multi-hop reasoning
L. Qiu, Y. Xiao, Y. Qu, H. Zhou, L. Li, W. Zhang, and Y. Yu · 2019
Earlier work this paper cites.
Explain yourself! leveraging language models for commonsense reasoning
N. F. Rajani, B. McCann, C. Xiong, and R. Socher · 2019
Earlier work this paper cites.
ATOMIC: an atlas of machine commonsense for if-then reasoning
M. Sap, R. L. Bras, E. Allaway, C. Bhagavatula, N. Lourie, H. Rashkin, B. Roof, N. A. Smith, and Y. Choi · 2019
Earlier work this paper cites.
Social iqa: Commonsense reasoning about social interactions
M. Sap, H. Rashkin, D. Chen, R. L. Bras, and Y. Choi · 2019
Earlier work this paper cites.
CLUTRR: A diagnostic benchmark for inductive reasoning from text
K. Sinha, S. Sodhani, J. Dong, J. Pineau, and W. L. Hamilton · 2019
Earlier work this paper cites.
A corpus for reasoning about natural language grounded in photographs
A. Suhr, S. Zhou, A. Zhang, I. Zhang, H. Bai, and Y. Artzi · 2019
Earlier work this paper cites.
Commonsenseqa: A question answering challenge targeting commonsense knowledge
A. Talmor, J. Herzig, N. Lourie, and J. Berant · 2019
Earlier work this paper cites.
WIQA: A dataset for "what if…" reasoning over procedural text
N. Tandon, B. Dalvi, K. Sakaguchi, P. Clark, and A. Bosselut · 2019
Earlier work this paper cites.
HEAD-QA: A healthcare dataset for complex reasoning
D. Vilares and C. Gómez-Rodríguez · 2019
Earlier work this paper cites.
Can neural networks understand monotonicity reasoning?
H. Yanaka, K. Mineshima, D. Bekki, K. Inui, S. Sekine, L. Abzianidze, and J. Bos · 2019
Earlier work this paper cites.
HELP: A dataset for identifying shortcomings of neural models in monotonicity reasoning
H. Yanaka, K. Mineshima, D. Bekki, K. Inui, S. Sekine, L. Abzianidze, and J. Bos · 2019
Earlier work this paper cites.
Hellaswag: Can a machine really finish your sentence?
R. Zellers, A. Holtzman, Y. Bisk, A. Farhadi, and Y. Choi · 2019
Earlier work this paper cites.
E3: entailment-driven extracting and editing for conversational machine reading
V. Zhong and L. Zettlemoyer · 2019
Earlier work this paper cites.
Abductive commonsense reasoning
C. Bhagavatula, R. L. Bras, C. Malaviya, K. Sakaguchi, A. Holtzman, H. Rashkin, D. Downey, W. Yih, and Y. Choi · 2020
Earlier work this paper cites.
PIQA: reasoning about physical commonsense in natural language
Y. Bisk, R. Zellers, R. L. Bras, J. Gao, and Y. Choi · 2020
Earlier work this paper cites.
Protoqa: A question answering dataset for prototypical common-sense reasoning
M. Boratko, X. Li, T. O’Gorman, R. Das, D. Le, and A. McCallum · 2020
Cited alongside, same era.
Language models are few-shot learners
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei · 2020
Cited alongside, same era.
Uncertain natural language inference
T. Chen, Z. Jiang, A. Poliak, K. Sakaguchi, and B. V. Durme · 2020
Cited alongside, same era.
Tabfact: A large-scale dataset for table-based fact verification
W. Chen, H. Wang, J. Chen, Y. Zhang, H. Wang, S. Li, X. Zhou, and W. Y. Wang · 2020
Cited alongside, same era.
Transformers as soft reasoners over language
P. Clark, O. Tafjord, and K. Richardson · 2020
Cited alongside, same era.
Scaling instruction-finetuned language models
H. W. Chung, L. Hou, S. Longpre, B. Zoph, Y. Tay, W. Fedus, E. Li, X. Wang, M. Dehghani, S. Brahma, A. Webson, S. S. Gu, Z. Dai, M. Suzgun, X. Chen, A. Chowdhery, S. Narang, G. Mishra, A. Yu, V. Y. Zhao, Y. Huang, A. M. Dai, H. Yu, S. Petrov, E. H. Chi, J. Dean, J. Devlin, A. Roberts, D. Zhou, Q. V. Le, and J. Wei · 2022
Later among the works it cites.
Faithful reasoning using large language models
A. Creswell and M. Shanahan · 2022
Later among the works it cites.
Selection-inference: Exploiting large language models for interpretable logical reasoning
A. Creswell, M. Shanahan, and I. Higgins · 2022
Later among the works it cites.
Why can GPT learn in-context? language models secretly perform gradient descent as meta-optimizers
D. Dai, Y. Sun, L. Dong, Y. Hao, Z. Sui, and F. Wei · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Feng, X. Chen, B. Y. Lin, P. Wang, J. Yan, and X. Ren · 2020
Cited alongside, same era.
Social chemistry 101: Learning to reason about social and moral norms
M. Forbes, J. D. Hwang, V. Shwartz, M. Sap, and Y. Choi · 2020
Cited alongside, same era.
Constructing A multi-hop QA dataset for comprehensive evaluation of reasoning steps
X. Ho, A. D. Nguyen, S. Sugawara, and A. Aizawa · 2020
Cited alongside, same era.
An analysis of natural language inference benchmarks through the lens of negation
M. M. Hossain, V. Kovatchev, P. Dutta, T. Kao, E. Wei, and E. Blanco · 2020
Cited alongside, same era.
R4C: A benchmark for evaluating RC systems to get the right answer for the right reason
N. Inoue, P. Stenetorp, and K. Inui · 2020
Cited alongside, same era.
Learning to explain: Datasets and models for identifying valid reasoning chains in multihop question-answering
H. Jhamtani and P. Clark · 2020
Cited alongside, same era.
QASC: A dataset for question answering via sentence composition
T. Khot, P. Clark, M. Guerquin, P. Jansen, and A. Sabharwal · 2020
Cited alongside, same era.
I. Dasgupta, A. K. Lampinen, S. C. Y. Chan, A. Creswell, D. Kumaran, J. L. McClelland, and F. Hill · 2022
Later among the works it cites.
Premise-based multimodal reasoning: Conditional inference on joint textual and visual clues
Q. Dong, Z. Qin, H. Xia, T. Feng, S. Tong, H. Meng, L. Xu, Z. Wei, W. Zhan, B. Chang, S. Li, T. Liu, and Z. Sui · 2022
Later among the works it cites.
e-care: a new dataset for exploring explainable causal reasoning
L. Du, X. Ding, K. Xiong, T. Liu, and B. Qin · 2022
Later among the works it cites.
CQG: A simple and effective controlled generation framework for multi-hop question generation
Z. Fei, Q. Zhang, T. Gui, D. Liang, S. Wang, W. Wu, and X. Huang · 2022
Later among the works it cites.
Peirce’s logic
A.-V. P. Francesco Bellucci · 2022
Later among the works it cites.
Misinfo reaction frames: Reasoning about readers’ reactions to news headlines
S. Gabriel, S. Hallinan, M. Sap, P. Nguyen, F. Roesner, E. Choi, and Y. Choi · 2022
Later among the works it cites.
Folio: Natural language reasoning with first-order logic
S. Han, H. Schoelkopf, Y. Zhao, Z. Qi, M. Riddell, L. Benson, L. Sun, E. Zubova, Y. Qiao, M. Burtell, D. Peng, J. Fan, Y. Liu, B. Wong, M. Sailor, A. Ni, L. Nan, J. Kasai, T. Yu, R. Zhang, S. Joty, A. R. Fabbri, W. Kryscinski, X. V. Lin, C. Xiong, and D. Radev · 2022
Later among the works it cites.
Wikiwhy: Answering and explaining cause-and-effect questions
M. Ho, A. Sharma, J. Chang, M. Saxon, S. Levy, Y. Lu, and W. Y. Wang · 2022
Later among the works it cites.
Large language models are reasoning teachers
N. Ho, L. Schmid, and S. Yun · 2022
Later among the works it cites.
METGEN: A module-based entailment tree generation framework for answer explanation
R. Hong, H. Zhang, X. Yu, and C. Zhang · 2022
Later among the works it cites.
Towards reasoning in large language models: A survey
J. Huang and K. C. Chang · 2022
Later among the works it cites.
Large language models can self-improve
J. Huang, S. S. Gu, L. Hou, Y. Wu, X. Wang, H. Yu, and J. Han · 2022
Later among the works it cites.
Metalogic: Logical reasoning explanations with fine-grained structure
Y. Huang, H. Zhang, R. Hong, X. Liang, C. Zhang, and D. Yu · 2022
Later among the works it cites.
Merit: Meta-path guided contrastive learning for logical reasoning
F. Jiao, Y. Guo, X. Song, and L. Nie · 2022
Later among the works it cites.
Maieutic prompting: Logically consistent reasoning with recursive explanations
J. Jung, L. Qin, S. Welleck, F. Brahman, C. Bhagavatula, R. L. Bras, and Y. Choi · 2022
Later among the works it cites.
LAMBADA: backward chaining for automated reasoning in natural language
S. M. Kazemi, N. Kim, D. Bhatia, X. Xu, and D. Ramachandran · 2022
Later among the works it cites.
Large language models are zero-shot reasoners
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa · 2022
Later among the works it cites.
Generated knowledge prompting for commonsense reasoning
J. Liu, A. Liu, X. Lu, S. Welleck, P. West, R. L. Bras, Y. Choi, and H. Hajishirzi · 2022
Later among the works it cites.
Teaching small language models to reason
L. C. Magister, J. Mallinson, J. Adámek, E. Malmi, and A. Severyn · 2022
Later among the works it cites.
Logicinference: A new dataset for teaching logical inference to seq2seq models
S. Ontañón, J. Ainslie, V. Cvicek, and Z. Fisher · 2022
Later among the works it cites.
Is a question decomposition unit all we need?
P. Patel, S. Mishra, M. Parmar, and C. Baral · 2022
Later among the works it cites.
Reasoning like program executors
X. Pi, Q. Liu, B. Chen, M. Ziyadi, Z. Lin, Y. Gao, Q. Fu, J. Lou, and W. Chen · 2022
Later among the works it cites.
Measuring and narrowing the compositionality gap in language models
O. Press, M. Zhang, S. Min, L. Schmidt, N. A. Smith, and M. Lewis · 2022
Later among the works it cites.
Reasoning with language model prompting: A survey
S. Qiao, Y. Ou, N. Zhang, X. Chen, Y. Yao, S. Deng, C. Tan, F. Huang, and H. Chen · 2022
Later among the works it cites.
Interpretable proof generation via iterative backward reasoning
H. Qu, Y. Cao, J. Gao, L. Ding, and R. Xu · 2022
Later among the works it cites.
CONDAQA: A contrastive reading comprehension dataset for reasoning about negation
A. Ravichander, M. Gardner, and A. Marasovic · 2022
Later among the works it cites.
Entailment tree explanations via iterative retrieval-generation reasoner
D. N. Ribeiro, S. Wang, X. Ma, R. Dong, X. Wei, H. Zhu, X. Chen, P. Xu, Z. Huang, A. O. Arnold, and D. Roth · 2022
Later among the works it cites.
Scinli: A corpus for natural language inference on scientific text
M. Sadat and C. Caragea · 2022
Later among the works it cites.
Robustlr: Evaluating robustness to logical perturbation in deductive reasoning
S. Sanyal, Z. Liao, and X. Ren · 2022
Later among the works it cites.
Fairr: Faithful and robust deductive reasoning over natural language
S. Sanyal, H. Singh, and X. Ren · 2022
Later among the works it cites.
APOLLO: A simple approach for adaptive pretraining of language models for logical reasoning
S. Sanyal, Y. Xu, S. Wang, Z. Yang, R. Pryzant, W. Yu, C. Zhu, and X. Ren · 2022
Later among the works it cites.
Language models are greedy reasoners: A systematic formal analysis of chain-of-thought
A. Saparov and H. He · 2022
Later among the works it cites.
K. Shridhar, A. Stolfo, and M. Sachan · 2022
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
A. Srivastava, A. Rastogi, A. Rao, A. A. M. Shoeb, A. Abid, A. Fisch, A. R. Brown, A. Santoro, A. Gupta, A. Garriga-Alonso, A. Kluska, A. Lewkowycz, A. Agarwal, A. Power, A. Ray, A. Warstadt, A. W. Kocurek, A. Safaya, A. Tazarv, A. Xiang, A. Parrish, A. Nie, A. Hussain, A. Askell, A. Dsouza, A. Rahane, A. S. Iyer, A. Andreassen, A. Santilli, A. Stuhlmüller, A. M. Dai, A. La, A. K. Lampinen, A. Zou, A. Jiang, A. Chen, A. Vuong, A. Gupta, A. Gottardi, A. Norelli, A. Venkatesh, A. Gholamidavoodi, A. Tabassum, A. Menezes, A. Kirubarajan, A. Mullokandov, A. Sabharwal, A. Herrick, A. Efrat, A. Erdem, A. Karakas, and et al · 2022
Later among the works it cites.
Challenging big-bench tasks and whether chain-of-thought can solve them
M. Suzgun, N. Scales, N. Schärli, S. Gehrmann, Y. Tay, H. W. Chung, A. Chowdhery, Q. V. Le, E. H. Chi, D. Zhou, and J. Wei · 2022
Later among the works it cites.
Entailer: Answering questions with faithful and truthful chains of reasoning
O. Tafjord, B. D. Mishra, and P. Clark · 2022
Later among the works it cites.
Musique: Multihop questions via single-hop question composition
H. Trivedi, N. Balasubramanian, T. Khot, and A. Sabharwal · 2022
Later among the works it cites.
Self-consistency improves chain of thought reasoning in language models
X. Wang, J. Wei, D. Schuurmans, Q. V. Le, E. H. Chi, and D. Zhou · 2022
Later among the works it cites.
Emergent abilities of large language models
J. Wei, Y. Tay, R. Bommasani, C. Raffel, B. Zoph, S. Borgeaud, D. Yogatama, M. Bosma, D. Zhou, D. Metzler, E. H. Chi, T. Hashimoto, O. Vinyals, P. Liang, J. Dean, and W. Fedus · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, E. H. Chi, Q. Le, and D. Zhou · 2022
Later among the works it cites.
Generating data to mitigate spurious correlations in natural language inference datasets
Y. Wu, M. Gardner, P. Stenetorp, and P. Dasigi · 2022
Later among the works it cites.
An explanation of in-context learning as implicit bayesian inference
S. M. Xie, A. Raghunathan, P. Liang, and T. Ma · 2022
Later among the works it cites.
Generating natural language proofs with verifier-guided search
K. Yang, J. Deng, and D. Chen · 2022
Later among the works it cites.
Language models as inductive reasoners
Z. Yang, L. Dong, X. Du, H. Cheng, E. Cambria, X. Liu, J. Gao, and F. Wei · 2022
Later among the works it cites.
Complementary explanations for effective in-context learning
X. Ye, S. Iyer, A. Celikyilmaz, V. Stoyanov, G. Durrett, and R. Pasunuru · 2022
Later among the works it cites.
Abductionrules: Training transformers to explain unexpected inputs
N. Young, Q. Bao, J. Bensemann, and M. Witbrock · 2022
Later among the works it cites.
ALERT: adapting language models to reasoning tasks
P. Yu, T. Wang, O. Golovneva, B. AlKhamissy, G. Ghosh, M. T. Diab, and A. Celikyilmaz · 2022
Later among the works it cites.
Star: Bootstrapping reasoning with reasoning
E. Zelikman, Y. Wu, J. Mu, and N. D. Goodman · 2022
Later among the works it cites.
On the paradox of learning to reason from data
H. Zhang, L. H. Li, T. Meng, K. Chang, and G. V. den Broeck · 2022
Later among the works it cites.
Greaselm: Graph reasoning enhanced language models
X. Zhang, A. Bosselut, M. Yasunaga, H. Ren, P. Liang, C. D. Manning, and J. Leskovec · 2022
Later among the works it cites.
Automatic chain of thought prompting in large language models
Z. Zhang, A. Zhang, M. Li, and A. Smola · 2022
Later among the works it cites.
Disentangling reasoning capabilities from language models with compositional reasoning transformers
W. Zhong, T. Ma, J. Wang, J. Yin, T. Zhao, C.-Y. Lin, and N. Duan · 2022
Later among the works it cites.
Analytical reasoning of text
W. Zhong, S. Wang, D. Tang, Z. Xu, D. Guo, Y. Chen, J. Wang, J. Yin, M. Zhou, and N. Duan · 2022
Later among the works it cites.
Learning to decompose: Hypothetical question decomposition based on comparable texts
B. Zhou, K. Richardson, X. Yu, and D. Roth · 2022
Later among the works it cites.
Least-to-most prompting enables complex reasoning in large language models
D. Zhou, N. Schärli, L. Hou, J. Wei, N. Scales, X. Wang, D. Schuurmans, C. Cui, O. Bousquet, Q. Le, and E. Chi · 2022
Later among the works it cites.
Y. Bang, S. Cahyawijaya, N. Lee, W. Dai, D. Su, B. Wilie, H. Lovenia, Z. Ji, T. Yu, W. Chung, Q. V. Do, Y. Xu, and P. Fung · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, H. Nori, H. Palangi, M. T. Ribeiro, and Y. Zhang · 2023
Closest in time.
Why think step-by-step? reasoning emerges from the locality of experience
B. Prystawski and N. D. Goodman · 2023
Closest in time.
Logical reasoning over natural language as knowledge representation: A survey
Z. Yang, X. Du, R. Mao, J. Ni, and E. Cambria · 2023
Closest in time.