Fetching the paper…
Reading the bibliography…
The ability to perform causal reasoning is widely considered a core feature of intelligence.
Fine-tuning language models from human preferences
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul F. Christiano, and Geoffrey Irving. 2019 · 1909
Earlier work this paper cites.
Causal inference and data fusion in econometrics
Paul Hünermund and Elias Bareinboim. 2019 · 1912
Earlier work this paper cites.
Fair inference on outcomes
Razieh Nabi and Ilya Shpitser. 2018 · 1940
Earlier work this paper cites.
Smoking and carcinoma of the lung
Richard Doll and A Bradford Hill. 1950 · 1950
Earlier work this paper cites.
The mortality of doctors in relation to their smoking habits
Richard Doll and A Bradford Hill. 1954 · 1954
Earlier work this paper cites.
Probabilistic reasoning in intelligent systems: Networks of plausible inference
Judea Pearl. 1988 · 1988
Earlier work this paper cites.
Rank-based systems: A simple approach to belief revision, belief update, and reasoning about evidence and actions
Moisés Goldszmidt and Judea Pearl. 1992 · 1992
Earlier work this paper cites.
Causal diagrams for empirical research
Judea Pearl. 1995 · 1995
Earlier work this paper cites.
The causal effect of education on earnings
David Card. 1999 · 1999
Earlier work this paper cites.
Causality: Models, reasoning and inference
Judea Pearl et al. 2000 · 2000
Earlier work this paper cites.
Causation, Prediction, and Search, Second Edition
Peter Spirtes, Clark Glymour, and Richard Scheines. 2000 · 2000
Earlier work this paper cites.
A rule-based style and grammar checker
Daniel Naber et al. 2003 · 2003
Earlier work this paper cites.
Returns to investment in education: A further update
George Psacharopoulos and Harry Anthony Patrinos. 2004 · 2004
Earlier work this paper cites.
Vaccines: Past, present and future
Stanley A Plotkin. 2005 · 2005
Earlier work this paper cites.
Causation and causal inference in epidemiology
Kenneth J Rothman and Sander Greenland. 2005 · 2005
Earlier work this paper cites.
Earnings functions, rates of return and treatment effects: The mincer equation and beyond
James J Heckman, Lance J Lochner, and Petra E Todd. 2006 · 2006
Earlier work this paper cites.
Pearl’s calculus of intervention is complete
Yimin Huang and Marco Valtorta. 2006 · 2006
Earlier work this paper cites.
Identification of conditional interventional distributions
Ilya Shpitser and Judea Pearl. 2006a · 2006
Earlier work this paper cites.
Identification of joint interventional distributions in recursive semi-markovian causal models
Ilya Shpitser and Judea Pearl. 2006b · 2006
Earlier work this paper cites.
Probabilistic networks and expert systems: Exact computational methods for Bayesian networks
Robert G Cowell, Philip Dawid, Steffen L Lauritzen, and David J Spiegelhalter. 2007 · 2007
Earlier work this paper cites.
Causality and counterfactuals in the situation calculus
Mark Hopkins and Judea Pearl. 2007 · 2007
Earlier work this paper cites.
Causal cognition in human and nonhuman animals: A comparative, critical review
Derek C Penn and Daniel J Povinelli. 2007 · 2007
Earlier work this paper cites.
Building a corpus of temporal-causal structure
Steven Bethard, William Corvey, Sara Klingenstein, and James H. Martin. 2008 · 2008
Earlier work this paper cites.
Connective-based local coherence analysis: A lexicon for recognizing causal relationships
Manfred Stede. 2008 · 2008
Earlier work this paper cites.
SemEval-2010 task 8: Multi-way classification of semantic relations between pairs of nominals
Iris Hendrickx, Su Nam Kim, Zornitsa Kozareva, Preslav Nakov, Diarmuid Ó Séaghdha, Sebastian Padó, Marco Pennacchiotti, Lorenza Romano, and Stan Szpakowicz. 2010 · 2010
Earlier work this paper cites.
How does your kindergarten classroom affect your earnings? Evidence from project star
Raj Chetty, John N Friedman, Nathaniel Hilger, Emmanuel Saez, Diane Whitmore Schanzenbach, and Danny Yagan. 2011 · 2011
Earlier work this paper cites.
Largest measles epidemic in North America in a decade—Quebec, Canada, 2011: Contribution of susceptibility, serendipity, and superspreading events
Gaston De Serres, France Markowski, Eveline Toth, Monique Landry, Danielle Auger, Marlène Mercier, Philippe Bélanger, Bruno Turmel, Horacio Arruda, Nicole Boulianne, et al. 2013 · 2011
Earlier work this paper cites.
Minimally supervised event causality identification
Quang Do, Yee Seng Chan, and Dan Roth. 2011 · 2011
Earlier work this paper cites.
The algorithmization of counterfactuals
Judea Pearl. 2011 · 2011
Earlier work this paper cites.
SemEval-2012 task 7: Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Andrew Gordon, Zornitsa Kozareva, and Melissa Roemmele. 2012 · 2012
Earlier work this paper cites.
Causal inference in public health
Thomas A Glass, Steven N Goodman, Miguel A Hernán, and Jonathan M Samet. 2013 · 2013
Earlier work this paper cites.
Benefits from immunization during the vaccines for children program era—united states, 1994–2013
Cynthia G Whitney, Fangjun Zhou, James Singleton, and Anne Schuchat. 2014 · 2013
Earlier work this paper cites.
Sapiens: A brief history of humankind
Yuval Noah Harari. 2014 · 2014
Earlier work this paper cites.
Annotating causality in the TempEval-3 corpus
Paramita Mirza, Rachele Sprugnoli, Sara Tonelli, and Manuela Speranza. 2014 · 2014
Earlier work this paper cites.
Causal inference in statistics: A primer
Madelyn Glymour, Judea Pearl, and Nicholas P Jewell. 2016 · 2016
Cited alongside, same era.
Identifying causal relations using parallel Wikipedia articles
Christopher Hidey and Kathy McKeown. 2016 · 2016
Cited alongside, same era.
CaTeRS: Causal and temporal relation scheme for semantic annotation of event structures
Nasrin Mostafazadeh, Alyson Grealish, Nathanael Chambers, James Allen, and Lucy Vanderwende. 2016 · 2016
Cited alongside, same era.
Causal inference in statistics: A primer
Judea Pearl, Madelyn Glymour, and Nicholas P Jewell. 2016 · 2016
Cited alongside, same era.
The BECauSE corpus 2.0: Annotating causality and overlapping relations
Jesse Dunietz, Lori Levin, and Jaime Carbonell. 2017 · 2017
Cited alongside, same era.
Elements of causal inference: Foundations and learning algorithms
Jonas Peters, Dominik Janzing, and Bernhard Schölkopf. 2017 · 2017
Roscoe: A suite of metrics for scoring step-by-step reasoning
Olga Golovneva, Moya Chen, Spencer Poff, Martin Corredor, Luke Zettlemoyer, Maryam Fazel-Zarandi, and Asli Celikyilmaz. 2022 · 2022
Later among the works it cites.
Wikiwhy: Answering and explaining cause-and-effect questions
Matthew Ho, Aditya Sharma, Justin Chang, Michael Saxon, Sharon Levy, Yujie Lu, and William Yang Wang. 2022 · 2022
Later among the works it cites.
Zhijing Jin, Abhinav Lalwani, Tejas Vaidhya, Xiaoyu Shen, Yiwen Ding, Zhiheng Lyu, Mrinmaya Sachan, Rada Mihalcea, and Bernhard Schölkopf. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F. Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Causal reasoning for algorithmic fairness
Joshua R Loftus, Chris Russell, Matt J Kusner, and Ricardo Silva. 2018 · 2018
Cited alongside, same era.
The book of why: The new science of cause and effect
Judea Pearl and Dana Mackenzie. 2018 · 2018
Cited alongside, same era.
Event2Mind: Commonsense inference on events, intents, and reactions
Hannah Rashkin, Maarten Sap, Emily Allaway, Noah A. Smith, and Yejin Choi. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Cited alongside, same era.
Counterfactual story reasoning and generation
Lianhui Qin, Antoine Bosselut, Ari Holtzman, Chandra Bhagavatula, Elizabeth Clark, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Comment: understanding simpson’s paradox
Judea Pearl. 2022 · 2022
Later among the works it cites.
External validity: From do-calculus to transportability across populations
Judea Pearl and Elias Bareinboim. 2022 · 2022
Later among the works it cites.
Drago Plecko and Elias Bareinboim. 2022 · 2022
Later among the works it cites.
Impact of pretraining term frequencies on few-shot numerical reasoning
Yasaman Razeghi, Robert L Logan IV, Matt Gardner, and Sameer Singh. 2022 · 2022
Later among the works it cites.
Large language models encode clinical knowledge
Karan Singhal, Shekoofeh Azizi, Tao Tu, S. Sara Mahdavi, Jason Wei, Hyung Won Chung, Nathan Scales, Ajay Kumar Tanwani, Heather Cole-Lewis, Stephen Pfohl, Perry Payne, Martin Seneviratne, Paul Gamble, Chris Kelly, Nathaneal Schärli, Aakanksha Chowdhery, Philip Andrew Mansfield, Blaise Agüera y Arcas, Dale R. Webster, Gregory S. Corrado, Yossi Matias, Katherine Chou, Juraj Gottweis, Nenad Tomasev, Yun Liu, Alvin Rajkomar, Joelle K. Barral, Christopher Semturs, Alan Karthikesalingam, and Vivek Natarajan. 2022 · 2022
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R. Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, Agnieszka Kluska, Aitor Lewkowycz, Akshat Agarwal, Alethea Power, Alex Ray, Alex Warstadt, Alexander W. Kocurek, Ali Safaya, Ali Tazarv, Alice Xiang, Alicia Parrish, Allen Nie, Aman Hussain, Amanda Askell, Amanda Dsouza, Ameet Rahane, Anantharaman S. Iyer, Anders Andreassen, Andrea Santilli, Andreas Stuhlmüller, Andrew M. Dai, Andrew La, Andrew K. Lampinen, Andy Zou, Angela Jiang, Angelica Chen, Anh Vuong, Animesh Gupta, Anna Gottardi, Antonio Norelli, Anu Venkatesh, Arash Gholamidavoodi, Arfa Tabassum, Arul Menezes, Arun Kirubarajan, Asher Mullokandov, Ashish Sabharwal, Austin Herrick, Avia Efrat, Aykut Erdem, Ayla Karakas, and et al. 2022 · 2022
Later among the works it cites.
Super-naturalinstructions: Generalization via declarative instructions on 1600+ NLP tasks
Yizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi, Amirreza Mirzaei, Atharva Naik, Arjun Ashok, Arut Selvan Dhanasekaran, Anjana Arunkumar, David Stap, Eshaan Pathak, Giannis Karamanolakis, Haizhi Gary Lai, Ishan Purohit, Ishani Mondal, Jacob Anderson, Kirby Kuznia, Krima Doshi, Kuntal Kumar Pal, Maitreya Patel, Mehrad Moradshahi, Mihir Parmar, Mirali Purohit, Neeraj Varshney, Phani Rohitha Kaza, Pulkit Verma, Ravsehaj Singh Puri, Rushang Karia, Savan Doshi, Shailaja Keyur Sampat, Siddhartha Mishra, Sujan Reddy A, Sumanta Patro, Tanay Dixit, and Xudong Shen. 2022 · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed H. Chi, Quoc V. Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
OPT: open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona T. Diab, Xian Li, Xi Victoria Lin, Todor Mihaylov, Myle Ott, Sam Shleifer, Kurt Shuster, Daniel Simig, Punit Singh Koura, Anjali Sridhar, Tianlu Wang, and Luke Zettlemoyer. 2022 · 2022
Later among the works it cites.
Education in the era of generative artificial intelligence (AI): Understanding the potential benefits of ChatGPT in promoting teaching and learning
David Baidoo-Anu and Leticia Owusu Ansah. 2023 · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with GPT-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott M. Lundberg, Harsha Nori, Hamid Palangi, Marco Túlio Ribeiro, and Yi Zhang. 2023 · 2023
Closest in time.
Testing GPT-4 with Wolfram Alpha and Code Interpreter plug-ins on math and science problems
Ernest Davis and Scott Aaronson. 2023 · 2023
Closest in time.
We may be surprised again: Why i take llms seriously
Ferenc Huszár. 2023 · 2023
Closest in time.
A PhD student’s perspective on research in NLP in the era of very large language models
Oana Ignat, Zhijing Jin, Artem Abzaliev, Laura Biester, Santiago Castro, Naihao Deng, Xinyi Gao, Aylin Gunal, Jacky He, Ashkan Kazemi, Muhammad Khalifa, Namho Koh, Andrew Lee, Siyang Liu, Do June Min, Shinka Mori, Joan Nwatu, Verónica Pérez-Rosas, Siqi Shen, Zekun Wang, Winston Wu, and Rada Mihalcea. 2023 · 2023
Closest in time.
Can large language models infer causation from correlation?
Zhijing Jin, Jiarui Liu, Zhiheng Lyu, Spencer Poff, Mrinmaya Sachan, Rada Mihalcea, Mona T. Diab, and Bernhard Schölkopf. 2023 · 2023
Closest in time.
Gpt-4 passes the bar exam
Daniel Martin Katz, Michael James Bommarito, Shang Gao, and Pablo Arredondo. 2023 · 2023
Closest in time.
Evaluating vaccine allocation strategies using simulation-assisted causal modeling
Armin Kekić, Jonas Dehning, Luigi Gresele, Julius von Kügelgen, Viola Priesemann, and Bernhard Schölkopf. 2023 · 2023
Closest in time.
Causal reasoning and large language models: Opening a new frontier for causality
Emre Kıcıman, Robert Ness, Amit Sharma, and Chenhao Tan. 2023 · 2023
Closest in time.
Israeli data: How can efficacy vs. severe disease be strong when 60% of hospitalized are vaccinated?
Jeffrey Morris. 2021 · 2023
Closest in time.
Capabilities of GPT-4 on medical challenge problems
Harsha Nori, Nicholas King, Scott Mayer McKinney, Dean Carignan, and Eric Horvitz. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Is chatgpt a general-purpose natural language processing task solver?
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, and Diyi Yang. 2023 · 2023
Closest in time.
Chatgpt: Bullshit spewer or the end of traditional assessments in higher education?
Jürgen Rudolph, Samson Tan, and Shannon Tan. 2023 · 2023
Closest in time.
A causal framework to quantify the robustness of mathematical reasoning with language models
Alessandro Stolfo, Zhijing Jin, Kumar Shridhar, Bernhard Schölkopf, and Mrinmaya Sachan. 2023 · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurélien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.
Causal parrots: Large language models may talk causality but are not causal
Matej Zečević, Moritz Willig, Devendra Singh Dhami, and Kristian Kersting. 2023 · 2023
Closest in time.
Understanding causality with large language models: Feasibility and opportunities
Cheng Zhang, Stefan Bauer, Paul Bennett, Jiangfeng Gao, Wenbo Gong, Agrin Hilmkil, Joel Jennings, Chao Ma, Tom Minka, Nick Pawlowski, et al. 2023 · 2023
Closest in time.
Can large language models transform computational social science?
Caleb Ziems, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2023 · 2023
Closest in time.