Fetching the paper…
Reading the bibliography…
Human commonsense understanding of the physical and social world is organized around intuitive theories.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
The problem of abortion and the doctrine of the double effect
Philippa Foot · 1967
Earlier work this paper cites.
Killing, letting die, and the trolley problem
Judith Jarvis Thomson · 1976
Earlier work this paper cites.
Causation in the law
H. L. A. Hart and T. Honoré · 1985
Earlier work this paper cites.
Knowledge-based causal attribution: The abnormal conditions focus model
D. J. Hilton and B. R. Slugoski · 1986
Earlier work this paper cites.
Mental simulation of causality
Gary L Wells and Igor Gavanski · 1989
Earlier work this paper cites.
Status-quo and omission biases
Ilana Ritov and Jonathan Baron · 1992
Earlier work this paper cites.
Cognitive development: Foundational theories of core domains
H. M. Wellman and S. A. Gelman · 1992
Earlier work this paper cites.
Mental simulation and causal attribution: When simulating an event does not affect fault assignment
Ahogni N’gbala and Nyla R Branscombe · 1995
Earlier work this paper cites.
Crediting causality
B. A. Spellman · 1997
Earlier work this paper cites.
Judgment dissociation theory: An analysis of differences in causal, counterfactual and covariational reasoning
David R Mandel · 2003
Earlier work this paper cites.
Omission bias, individual differences, and normality
Jonathan Baron and Ilana Ritov · 2004
Earlier work this paper cites.
Moral minds: How nature designed our universal sense of right and wrong
Marc Hauser · 2006
Earlier work this paper cites.
Morality and Self-interest
Paul Bloomfield · 2007
Earlier work this paper cites.
Universal moral grammar: Theory, evidence and the future
John Mikhail · 2007
Earlier work this paper cites.
Throwing a bomb on a person versus throwing a person on a bomb: Intervention myopia in moral intuitions
Michael R Waldmann and Jörn H Dieterich · 2007
Earlier work this paper cites.
Representing causation
Phillip Wolff · 2007
Earlier work this paper cites.
Causal judgment and moral judgment: Two experiments
Joshua Knobe and Ben Fraser · 2008
Earlier work this paper cites.
Who shalt not kill? individual differences in working memory capacity, executive control, and moral judgment
Adam B Moore, Brian A Clark, and Michael J Kane · 2008
Earlier work this paper cites.
Pushing moral buttons: The interaction between personal force and intention in moral judgment
Joshua D Greene, Fiery A Cushman, Lisa E Stewart, Kelly Lowenberg, Leigh E Nystrom, and Jonathan D Cohen · 2009
Earlier work this paper cites.
Cause and norm
Christopher Hitchcock and Joshua Knobe · 2009
Earlier work this paper cites.
SemEval-2010 task 8: Multi-way classification of semantic relations between pairs of nominals
Iris Hendrickx, Su Nam Kim, Zornitsa Kozareva, Preslav Nakov, Diarmuid Ó Séaghdha, Sebastian Padó, Marco Pennacchiotti, Lorenza Romano, and Stan Szpakowicz · 2010
Earlier work this paper cites.
Selecting explanations from causal chains: Do statistical principles explain preferences for voluntary causes
D. J. Hilton, J. McClure, and R. M. Sutton · 2010
Earlier work this paper cites.
Causation, norm violation, and culpable control
Mark D Alicke, David Rose, and Dori Bloom · 2011
Earlier work this paper cites.
The omission effect in moral cognition: Toward a functional explanation
Peter DeScioli, Rebecca Bruening, and Robert Kurzban · 2011
Earlier work this paper cites.
How the source, inevitability and means of bringing about harm interact in folk-moral judgments
Bryce Huebner, Marc D Hauser, and Phillip Pettit · 2011
Earlier work this paper cites.
Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Melissa Roemmele, Cosmin Adrian Bejan, and Andrew S Gordon · 2011
Earlier work this paper cites.
When contributions make a difference: Explaining order effects in responsibility attributions
Tobias Gerstenberg and David A. Lagnado · 2012
Earlier work this paper cites.
Putting the trolley in order: Experimental philosophy and the loop case
S Matthew Liao, Alex Wiegmann, Joshua Alexander, and Gerard Vong · 2012
Earlier work this paper cites.
Order effects in moral judgment
Alex Wiegmann, Yasmina Okan, and Jonas Nagel · 2012
Earlier work this paper cites.
Simulation as an engine of physical scene understanding
Peter W Battaglia, Jessica B Hamrick, and Joshua B Tenenbaum · 2013
Earlier work this paper cites.
Causal responsibility and counterfactuals
D. A. Lagnado, T. Gerstenberg, and R. Zultan · 2013
Earlier work this paper cites.
Moral judgment reloaded: a moral dilemma validation study
Julia F Christensen, Albert Flexas, Margareta Calabrese, Nadine K Gut, and Antoni Gomila · 2014
Earlier work this paper cites.
Causal inference in conjoint analysis: Understanding multidimensional choices via stated preference experiments
Jens Hainmueller, Daniel J Hopkins, and Teppei Yamamoto · 2014
Earlier work this paper cites.
The good, the bad, and the timely: how temporal order and moral judgment influence causal selection
Kevin Reuter, Lara Kirfel, Raphael Van Riel, and Luca Barlassina · 2014
Earlier work this paper cites.
Causation, norms, and omissions: A study of causal judgments
Randolph Clarke, Joshua Shepherd, John Stigall, Robyn Repko Waller, and Chris Zarpentine · 2015
Cited alongside, same era.
Commonsense reasoning and commonsense knowledge in artificial intelligence
Ernest Davis and Gary Marcus · 2015
Cited alongside, same era.
Burning Conscience: The Case Of The Hiroshima Pilot Claude Eatherly
Claude Eatherly and Günther Anders · 2015
Cited alongside, same era.
Cause, responsibility and blame: a structural-model approach
Joseph Y Halpern · 2015
Cited alongside, same era.
Inference of intention and permissibility in moral decision making
Max Kleiman-Weiner, Tobias Gerstenberg, Sydney Levine, and Joshua B Tenenbaum · 2015
Cited alongside, same era.
Causal superseding
Jonathan F Kominsky, Jonathan Phillips, Tobias Gerstenberg, David Lagnado, and Joshua Knobe · 2015
Aligning ai with shared human values
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song, and Jacob Steinhardt · 2020
Later among the works it cites.
Learning to summarize with human feedback
Nisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F Christiano · 2020
Later among the works it cites.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al · 2021
Later among the works it cites.
Moral stories: Situated reasoning about norms, intents, actions, and their consequences
Denis Emelin, Ronan Le Bras, Jena D. Hwang, Maxwell Forbes, and Yejin Choi · 2021
Later among the works it cites.
A counterfactual simulation model of causation by omission
Tobias Gerstenberg and Simon Stephan · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Causality in thought
Steven A Sloman and David Lagnado · 2015
Cited alongside, same era.
A corpus and cloze evaluation for deeper understanding of commonsense stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James Allen · 2016
Cited alongside, same era.
The role of prescriptive norms and knowledge in children’s and adults’ causal selection
Jana Samland, Marina Josephs, Michael R Waldmann, and Hannes Rakoczy · 2016
Cited alongside, same era.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Chris L. Baker, Julian Jara-Ettinger, Rebecca Saxe, and Joshua B. Tenenbaum · 2017
Cited alongside, same era.
Intuitive theories
Tobias Gerstenberg and Joshua B. Tenenbaum · 2017
Cited alongside, same era.
Cause by omission and norm: Not watering plants
Paul Henne, Ángel Pinillos, and Felipe De Brigard · 2017
Cited alongside, same era.
Later among the works it cites.
A counterfactual simulation model of causal judgments for physical events
Tobias Gerstenberg, Noah D. Goodman, David A. Lagnado, and Joshua B. Tenenbaum · 2021
Later among the works it cites.
Data science as political action: Grounding data science in a politics of justice
Ben Green · 2021
Later among the works it cites.
What would jiminy cricket do? towards agents that behave morally
Dan Hendrycks, Mantas Mazeika, Andy Zou, Sahil Patel, Christine Zhu, Jesus Navarro, Dawn Song, Bo Li, and Jacob Steinhardt · 2021
Later among the works it cites.
Counterfactual thinking and recency effects in causal judgment
Paul Henne, Aleksandra Kulesza, Karla Perez, and Augustana Houcek · 2021
Later among the works it cites.
Delphi: Towards machine ethics and norms
Liwei Jiang, Jena D Hwang, Chandra Bhagavatula, Ronan Le Bras, Maxwell Forbes, Jon Borchardt, Jenny Liang, Oren Etzioni, Maarten Sap, and Yejin Choi · 2021
Later among the works it cites.
Zachary Kenton, Tom Everitt, Laura Weidinger, Iason Gabriel, Vladimir Mikulik, and Geoffrey Irving · 2021
Later among the works it cites.
Scruples: A corpus of community ethical judgments on 32,000 real-life anecdotes
Nicholas Lourie, Ronan Le Bras, and Yejin Choi · 2021
Later among the works it cites.
Degrading causation
Kevin O’Neill, Paul Henne, Paul Bello, John Pearson, and Felipe De Brigard · 2021
Later among the works it cites.
Crossed wires: Blaming artifacts for bad outcomes
Justin Sytsma · 2021
Later among the works it cites.
A word on machine ethics: A response to jiang et al. (2021), 2021
Zeerak Talat, Hagen Blix, Josef Valvoda, Maya Indira Ganesh, Ryan Cotterell, and Adina Williams · 2021
Later among the works it cites.
Human-level reinforcement learning through theory-based modeling, exploration, and planning
Pedro A Tsividis, Joao Loula, Jake Burga, Nathan Foss, Andres Campero, Thomas Pouncy, Samuel J Gershman, and Joshua B Tenenbaum · 2021
Later among the works it cites.
Revisiting the" video" in video-language understanding
Shyamal Buch, Cristóbal Eyzaguirre, Adrien Gaidon, Jiajun Wu, Li Fei-Fei, and Juan Carlos Niebles · 2022
Later among the works it cites.
Dealing with disagreements: Looking beyond the majority vote in subjective annotations
Aida Mostafazadeh Davani, Mark Díaz, and Vinodkumar Prabhakaran · 2022
Later among the works it cites.
When to make exceptions: Exploring language models as accounts of human moral judgment
Zhijing Jin, Sydney Levine, Fernando Gonzalez, Ojasv Kamal, Maarten Sap, Mrinmaya Sachan, Rada Mihalcea, Josh Tenenbaum, and Bernhard Schölkopf · 2022
Later among the works it cites.
Towards understanding how machines can learn causal overhypotheses
Eliza Kosoy, David M Chan, Adrian Liu, Jasmine Collins, Bryanna Kaufmann, Sandy Han Huang, Jessica B Hamrick, John Canny, Nan Rosemary Ke, and Alison Gopnik · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Social simulacra: Creating populated prototypes for social computing systems
Joon Sung Park, Lindsay Popowski, Carrie Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein · 2022
Later among the works it cites.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, and Denny Zhou · 2022
Later among the works it cites.
Large language models are human-level prompt engineers
Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba · 2022
Later among the works it cites.
The moral integrity corpus: A benchmark for ethical dialogue systems
Caleb Ziems, Jane A Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang · 2022
Later among the works it cites.
Exploring the psychology of gpt-4’s moral and legal reasoning
Guilherme FCF Almeida, José Luiz Nunes, Neele Engelmann, Alex Wiegmann, and Marcelo de Araújo · 2023
Closest in time.
Can ai language models replace human participants?
Danica Dillion, Niket Tandon, Yuling Gu, and Kurt Gray · 2023
Closest in time.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli · 2023
Closest in time.
Causal reasoning and large language models: Opening a new frontier for causality
Emre Kıcıman, Robert Ness, Amit Sharma, and Chenhao Tan · 2023
Closest in time.
Inverse scaling: When bigger isn’t better
Ian R McKenzie, Alexander Lyzhov, Michael Pieler, Alicia Parrish, Aaron Mueller, Ameya Prabhu, Euan McLean, Aaron Kirtland, Alexis Ross, Alisa Liu, et al · 2023
Closest in time.
Gpt-4 technical report
OpenAI · 2023
Closest in time.
Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark
Alexander Pan, Jun Shern Chan, Andy Zou, Nathaniel Li, Steven Basart, Thomas Woodside, Hanlin Zhang, Scott Emmons, and Dan Hendrycks · 2023
Closest in time.
Counterfactuals and the logic of causal selection
Tadeg Quillien and Christopher G Lucas · 2023
Closest in time.
The puzzle of evaluating moral cognition in artificial agents
Madeline G Reinecke, Yiran Mao, Markus Kunesch, Edgar A Duéñez-Guzmán, Julia Haas, and Joel Z Leibo · 2023
Closest in time.
Evaluating the moral beliefs encoded in llms
Nino Scherrer, Claudia Shi, Amir Feder, and David M Blei · 2023
Closest in time.