Fetching the paper…
Reading the bibliography…
Understanding narratives requires reasoning about the cause-and-effect relationships between events mentioned in the text.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Causal discovery toolbox: Uncover causal relationships in python
Diviyan Kalainathan and Olivier Goudet. 2019 · 1903
Earlier work this paper cites.
Simple bert models for relation extraction and semantic role labeling
Peng Shi and Jimmy Lin. 2019 · 1904
Earlier work this paper cites.
Multi-hop reading comprehension across multiple documents by reasoning over heterogeneous graphs
Ming Tu, Guangtao Wang, Jing Huang, Yun Tang, Xiaodong He, and Bowen Zhou. 2019 · 1905
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 1908
Earlier work this paper cites.
Causality and the science of human behavior
Adolf Grunbaum. 1952 · 1952
Earlier work this paper cites.
Building a corpus of temporal-causal structure
Steven Bethard, William J Corvey, Sara Klingenstein, and James H Martin. 2008 · 2008
Earlier work this paper cites.
Scilens news platform: a system for real-time evaluation of news articles
Angelika Romanou, Panayiotis Smeros, Carlos Castillo, and Karl Aberer. 2020 · 2008
Earlier work this paper cites.
Causal inference in statistics: An overview
Judea Pearl. 2009 · 2009
Earlier work this paper cites.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Melissa Roemmele, Cosmin Adrian Bejan, and Andrew S Gordon. 2011 · 2011
Earlier work this paper cites.
Causation, touch, and the perception of force
Phillip Wolff and Jason Shepard. 2013 · 2013
Earlier work this paper cites.
Illusions of causality: how they bias our everyday thinking and how they could be reduced
Helena Matute, Fernando Blanco, Ion Yarritu, Marcos Díaz-Lago, Miguel A Vadillo, and Itxaso Barberia. 2015 · 2015
Earlier work this paper cites.
Actual Causality
Joseph Y. Halpern. 2016 · 2016
Earlier work this paper cites.
Identifying causal relations using parallel wikipedia articles
Christopher Hidey and Kathleen McKeown. 2016 · 2016
Earlier work this paper cites.
Distinguishing cause from effect using observational data: methods and benchmarks
Joris M Mooij, Jonas Peters, Dominik Janzing, Jakob Zscheischler, and Bernhard Schölkopf. 2016 · 2016
Earlier work this paper cites.
Normality and actual causal strength
Thomas F Icard, Jonathan F Kominsky, and Joshua Knobe. 2017 · 2017
Earlier work this paper cites.
The book of why: the new science of cause and effect
Judea Pearl and Dana Mackenzie. 2018 · 2018
Cited alongside, same era.
Constructing datasets for multi-hop reading comprehension across documents
Johannes Welbl, Pontus Stenetorp, and Sebastian Riedel. 2018 · 2018
Cited alongside, same era.
Causal discovery with attention-based convolutional neural networks
Meike Nauta, Doina Bucur, and Christin Seifert. 2019 · 2019
Cited alongside, same era.
Scilens: Evaluating the quality of scientific news articles using social media and scientific literature indicators
Panayiotis Smeros, Carlos Castillo, and Karl Aberer. 2019 · 2019
Cited alongside, same era.
Detecting causal language use in science findings
Bei Yu, Yingya Li, and Jun Wang. 2019 · 2019
Cited alongside, same era.
Multi-sentence argument linking
Seth Ebner, Patrick Xia, Ryan Culkin, Kyle Rawlins, and Benjamin Van Durme. 2020 · 2020
A counterfactual model of causal judgment in double prevention
Kevin O’Neill, Tadeg Quillien, and Paul Henne. 2022 · 2022
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, et al. 2022 · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Instructeval: Towards holistic evaluation of instruction-tuned large language models
Yew Ken Chia, Pengfei Hong, Lidong Bing, and Soujanya Poria. 2023 · 2023
Closest in time.
Chatgpt goes to law school
Jonathan H Choi, Kristin E Hickman, Amy Monahan, and Daniel Schwarcz. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Crass: A novel data set and benchmark to test counterfactual reasoning of large language models
Jörg Frohberg and Frank Binder. 2021 · 2021
Cited alongside, same era.
A counterfactual simulation model of causal judgments for physical events
Tobias Gerstenberg, Noah D Goodman, David A Lagnado, and Joshua B Tenenbaum. 2021 · 2021
Cited alongside, same era.
Pengcheng He, Jianfeng Gao, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Norms affect prospective causal judgments
Paul Henne, Kevin O’Neill, Paul Bello, Sangeet Khemlani, and Felipe De Brigard. 2021 · 2021
Cited alongside, same era.
Benchmarking of data-driven causality discovery approaches in the interactions of arctic sea ice and atmosphere
Yiyi Huang, Matthäus Kleindessner, Alexey Munishkin, Debvrat Varshney, Pei Guo, and Jianwu Wang. 2021 · 2021
Cited alongside, same era.
Unleash gpt-2 power for event detection
Amir Pouran Ben Veyseh, Viet Dac Lai, Franck Dernoncourt, and Thien Huu Nguyen. 2021 · 2021
Cited alongside, same era.
Closest in time.
Calm-bench: A multi-task benchmark for evaluating causality-aware language models
Dhairya Dalal, Paul Buitelaar, and Mihael Arcan. 2023 · 2023
Closest in time.
Alon Jacovi, Avi Caciularu, Omer Goldman, and Yoav Goldberg. 2023 · 2023
Closest in time.
Causal reasoning and large language models: Opening a new frontier for causality
Emre Kıcıman, Robert Ness, Amit Sharma, and Chenhao Tan. 2023 · 2023
Closest in time.
The flan collection: Designing data and methods for effective instruction tuning
Shayne Longpre, Le Hou, Tu Vu, Albert Webson, Hyung Won Chung, Yi Tay, Denny Zhou, Quoc V Le, Barret Zoph, Jason Wei, et al. 2023 · 2023
Closest in time.
Capabilities of gpt-4 on medical challenge problems
Harsha Nori, Nicholas King, Scott Mayer McKinney, Dean Carignan, and Eric Horvitz. 2023 · 2023
Closest in time.
Gpt-4 technical report
R OpenAI. 2023 · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Closest in time.
Understanding causality with large language models: Feasibility and opportunities
Cheng Zhang, Stefan Bauer, Paul Bennett, Jiangfeng Gao, Wenbo Gong, Agrin Hilmkil, Joel Jennings, Chao Ma, Tom Minka, Nick Pawlowski, et al. 2023 · 2023
Closest in time.
Extracting victim counts from text
Mian Zhong, Shehzaad Dhuliawala, and Niklas Stoehr. 2023 · 2023
Closest in time.