Fetching the paper…
Reading the bibliography…
We assess the ability of large language models (LLMs) to answer causal questions by analyzing their strengths and weaknesses against three types of causal question.
Causation, prediction, and search
P. Spirtes, C. N. Glymour, R. Scheines, and D. Heckerman · 2000
Earlier work this paper cites.
Methods for causal inference from gene perturbation experiments and validation
N. Meinshausen, A. Hauser, J. M. Mooij, J. Peters, P. Versteeg, and P. Bühlmann · 2016
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Causal structure learning
C. Heinze-Deml, M. H. Maathuis, and N. Meinshausen · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever, et al · 2018
Earlier work this paper cites.
Causal discovery of feedback networks with functional magnetic resonance imaging
R. Sanchez-Romero, J. D. Ramsey, K. Zhang, M. K. Glymour, B. Huang, and C. Glymour · 2018
Earlier work this paper cites.
Review of causal discovery methods based on graphical models
C. Glymour, K. Zhang, and P. Spirtes · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever, et al · 2019
Earlier work this paper cites.
Inferring causation from time series in earth system sciences
J. Runge, S. Bathiany, E. Bollt, G. Camps-Valls, D. Coumou, E. Deyle, C. Glymour, M. Kretschmer, M. D. Mahecha, J. Muñoz-Marí, et al · 2019
Earlier work this paper cites.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Cited alongside, same era.
Joint causal inference from multiple contexts
J. M. Mooij, S. Magliacane, and T. Claassen · 2020
Cited alongside, same era.
Shaking the foundations: delusions in sequence models for interaction and control
P. A. Ortega, M. Kunesch, G. Delétang, T. Genewein, J. Grau-Moya, J. Veness, J. Buchli, J. Degrave, B. Piot, J. Perolat, et al · 2021
Cited alongside, same era.
Reprogramming the American Dream: From Rural America to Silicon Valley—Making AI Serve Us All
K. Scott and G. Shaw · 2021
Cited alongside, same era.
Grey-box extraction of natural language models
S. Zanella-Beguelin, S. Tople, A. Paverd, and B. Köpf · 2021
Cited alongside, same era.
Lamda: Language models for dialog applications
R. Thoppilan, D. De Freitas, J. Hall, N. Shazeer, A. Kulshreshtha, H.-T. Cheng, A. Jin, T. Bos, L. Baker, Y. Du, et al · 2022
Later among the works it cites.
Toward self-learning end-to-end task-oriented dialog systems
X. Zhang, B. Peng, J. Gao, and H. Meng · 2022
Later among the works it cites.
Sparks of artificial general intelligence: Early experiments with gpt-4, 2023
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, H. Nori, H. Palangi, M. T. Ribeiro, and Y. Zhang · 2023
Closest in time.
How close is chatgpt to human experts? comparison corpus, evaluation, and detection
B. Guo, X. Zhang, Z. Wang, M. Jiang, J. Nie, Y. Ding, J. Yue, and Y. Wu · 2023
Closest in time.
Dissociating language and thought in large language models: a cognitive perspective
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Fu, H. Peng, A. Sabharwal, P. Clark, and T. Khot · 2022
Cited alongside, same era.
Deep end-to-end causal inference
T. Geffner, J. Antoran, A. Foster, W. Gong, C. Ma, E. Kiciman, A. Sharma, A. Lamb, M. Kukla, N. Pawlowski, et al · 2022
Cited alongside, same era.
Causalnlp tutorial: An introduction to causality for natural language processing
Z. Jin, A. Feder, and K. Zhang · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Cited alongside, same era.
K. Mahowald, A. A. Ivanova, I. A. Blank, N. Kanwisher, J. B. Tenenbaum, and E. Fedorenko · 2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.
Wolfram|alpha as the way to bring computational knowledge superpowers to chatgpt
S. Wolfram · 2023
Closest in time.
On the opportunity of causal deep generative models: A survey and future directions
G. Zhou, L. Yao, X. Xu, C. Wang, L. Zhu, and K. Zhang · 2023
Closest in time.