Fetching the paper…
Reading the bibliography…
As AI systems become increasingly powerful and pervasive, there are growing concerns about machines' morality or a lack thereof.
Outline of a decision procedure for ethics
John Rawls · 1951
Earlier work this paper cites.
A Theory of Justice
John Rawls · 1971
Earlier work this paper cites.
Ethics: Inventing Right and Wrong
John Leslie Mackie · 1977
Earlier work this paper cites.
Wide reflective equilibrium and theory acceptance in ethics
Norman Daniels · 1979
Earlier work this paper cites.
Artificial Intelligence: The Very Idea
John Haugeland · 1985
Earlier work this paper cites.
The View From Nowhere
Thomas Nagel · 1986
Earlier work this paper cites.
The first law of robotics (a call to arms)
Daniel Weld and Oren Etzioni · 1994
Earlier work this paper cites.
The Sources of Normativity
Christine M. Korsgaard · 1996
Earlier work this paper cites.
Hate speech
John T. Nockleby · 2000
Earlier work this paper cites.
State violence and lesbian, gay, bisexual and transgender (lgbt) rights
Mark Ungar · 2000
Earlier work this paper cites.
Groundwork for the Metaphysics of Morals
Immanuel Kant · 2002
Earlier work this paper cites.
Finite beings, finite goods: The semantics, metaphysics and ethics of naturalist consequentialism, part i
Richard Boyd · 2003
Earlier work this paper cites.
A dissociation between moral judgments and justifications
Marc Hauser, Fiery Cushman, Liane Young, J. I. N. Kang-Xing, and John Mikhail · 2006
Earlier work this paper cites.
Universal moral grammar: theory, evidence and the future
John Mikhail · 2006
Earlier work this paper cites.
The nature, importance, and difficulty of machine ethics
James Moor · 2006
Earlier work this paper cites.
Natural Moralities:A Defense of Pluralistic Relativism: A Defense of Pluralistic Relativism
David B. Wong · 2006
Earlier work this paper cites.
Modelling morality with prospective logic
Luís Moniz Pereira and Ari Saptawijaya · 2007
Earlier work this paper cites.
Asimov’s “three laws of robotics” and machine metaethics
Susan Leigh Anderson · 2008
Earlier work this paper cites.
Moral relativism
Steven Lukes · 2008
Earlier work this paper cites.
Moral Machines: Teaching Machines Right from Wrong
Wendell Wallach and Colin Allen · 2010
Earlier work this paper cites.
Thou shalt not discriminate: How emphasizing moral ideals rather than obligations increases whites’ support for social equality
Serena Does, Belle Derks, and Naomi Ellemers · 2011
Earlier work this paper cites.
On What Matters: Volume One
Derek Parfit · 2011
Earlier work this paper cites.
Coming to terms with contingency : Humean constructivism about practical reason
Sharon Street · 2012
Earlier work this paper cites.
Power to the people: The role of humans in interactive machine learning
Saleema Amershi, Maya Cakmak, W. Knox, and Todd Kulesza · 2014
Earlier work this paper cites.
Isse: An interactive source separation editor
Nicholas J. Bryan, Gautham J. Mysore, and Ge Wang · 2014
Earlier work this paper cites.
Modelling moral reasoning and ethical responsibility with logic programming
Fiona Berreby, Gauvain Bourgne, and Jean-Gabriel Ganascia · 2015
Earlier work this paper cites.
Norms in the Wild, How to Diagnose, Measure and Change Social Norms
Christina Bicchieri · 2016
Earlier work this paper cites.
Self-driving cars will teach themselves to save lives—but also take them | wired
Cade Metz · 2016
Earlier work this paper cites.
A corpus and evaluation framework for deeper understanding of commonsense stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James F. Allen · 2016
Earlier work this paper cites.
Integrating social power into the decision-making of cognitive agents
Gonçalo Pereira, Rui Prada, and Pedro A. Santos · 2016
Earlier work this paper cites.
Big data: A report on algorithmic systems, opportunity, and civil rights, 2016
White House · 2016
Earlier work this paper cites.
The problem with bias: Allocative versus representational harms in machine learning
Solon Barocas, Kate Crawford, Aaron Shapiro, and Hanna Wallach · 2017
Earlier work this paper cites.
Attention, intentions, and the structure of discourse
Barbara J. Grosz and Candace L. Sidner · 2017
Earlier work this paper cites.
The Moral Machine experiment
Edmond Awad, Sohan Dsouza, Richard Kim, Jonathan Schulz, Joseph Henrich, Azim Shariff, Jean-François Bonnefon, and Iyad Rahwan · 2018
Earlier work this paper cites.
People are averse to machines making moral decisions
Yochanan E. Bigman and Kurt Gray · 2018
Earlier work this paper cites.
The malicious use of artificial intelligence: Forecasting, prevention, and mitigation, 2018
Miles Brundage, Shahar Avin, Jack Clark, Helen Toner, Peter Eckersley, Ben Garfinkel, Allan Dafoe, Paul Scharre, Thomas Zeitzoff, Bobby Filar, Hyrum Anderson, Heather Roff, Gregory C. Allen, Jacob Steinhardt, Carrick Flynn, Seán Ó hÉigeartaigh, Simon Beard, Haydn Belfield, Sebastian Farquhar, Clare Lyle, Rebecca Crootof, Owain Evans, Michael Page, Joanna Bryson, Roman Yampolskiy, and Dario Amodei · 2018
Earlier work this paper cites.
Estimating the reproducibility of experimental philosophy
Florian Cova, Brent Strickland, Angela Gaia Felicita Abatista, Aurélien Allard, James Andow, Mario Attie, James R. Beebe, Renatas Berniūnas, Jordane Boudesseul, Matteo Colombo, Fiery Andrews Cushman, Rodrigo Díaz, Noah N’Djaye Nikolai van Dongen, Vilius Dranseika, Brian D. Earp, Antonio Gaitán Torres, Ivar Rodríguez Hannikainen, José V. Hernández-Conde, Wenjia Hu, François Jaquet, Kareem Khalifa, Hannah Kim, Markus Kneer, Joshua Knobe, Miklos Kurthy, Anthony Lantian, Shen-yi Liao, Edouard Machery, Tania Moerenhout, Christian Mott, Mark Phelan, Jonathan Scott Phillips, Navin Rambharose, Kevin Reuter, Felipe Romero, Paulo Sousa, Jan Sprenger, Emile Thalabard, Kevin Patrick Tobia, Hugo Viciana, Daniel A. Wilkenfeld, and Xiang Zhou · 2018
Earlier work this paper cites.
Measuring and mitigating unintended bias in text classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman · 2018
Earlier work this paper cites.
Point: Should ai technology be regulated? yes, and here’s how
Oren Etzioni · 2018
Earlier work this paper cites.
A computational model of commonsense moral decision making
Richard Kim, Max Kleiman-Weiner, Andres Abeliuk, Edmond Awad, Sohan Dsouza, Joshua Tenenbaum, and Iyad Rahwan · 2018
Earlier work this paper cites.
Amazon scraps secret ai recruiting tool that showed bias against women, 2018
Reuters · 2018
Cited alongside, same era.
Building trust in artificial intelligence
Francesca Rossi · 2018
Cited alongside, same era.
The marked and the unmarked
Eviatar Zerubavel · 2018
Cited alongside, same era.
Guidelines for human-ai interaction
Saleema Amershi, Dan Weld, Mihaela Vorvoreanu, Adam Fourney, Besmira Nushi, Penny Collisson, Jina Suh, Shamsi Iqbal, Paul N. Bennett, Kori Inkpen, Jaime Teevan, Ruth Kikin-Gil, and Eric Horvitz · 2019
Cited alongside, same era.
Race After Technology: Abolitionist Tools for the New Jim Code
Ruha Benjamin · 2019
Cited alongside, same era.
Ethics guidelines for trustworthy artificial intelligence, 2019
European Commission · 2019
Cited alongside, same era.
Atlas of AI
Kate Crawford · 2021
Closest in time.
Moral foundation theory, 2021
David Dobolyi · 2021
Closest in time.
Documenting large webtext corpora: A case study on the colossal clean crawled corpus
Jesse Dodge, Maarten Sap, Ana Marasović, William Agnew, Gabriel Ilharco, Dirk Groeneveld, Margaret Mitchell, and Matt Gardner · 2021
Closest in time.
Latent hatred: A benchmark for understanding implicit hate speech
Mai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi, Jordyn Seybolt, Munmun De Choudhury, and Diyi Yang · 2021
Closest in time.
Moral stories: Situated reasoning about norms, intents, actions, and their consequences
Denis Emelin, Ronan Le Bras, Jena D. Hwang, Maxwell Forbes, and Yejin Choi · 2021
Closest in time.
This program can give AI a sense of Ethics—Sometimes
Will Knight · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bound in hatred: The role of group-based morality in acts of hate
Joseph Hoover, Mohammad Atari, Aida Mostafazadeh Davani, Brendan Kennedy, Gwenyth Portillo-Wightman, Leigh Yeh, Drew Kogon, and Morteza Dehghani · 2019
Cited alongside, same era.
Cosmos qa: Machine reading comprehension with contextual commonsense reasoning
Lifu Huang, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi · 2019
Cited alongside, same era.
In Rebooting AI: Building Artificial Intelligence We Can Trust , 2019
Gary Marcus and Ernest Davis · 2019
Cited alongside, same era.
Model cards for model reporting
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru · 2019
Cited alongside, same era.
Social iqa: Commonsense reasoning about social interactions
Maarten Sap, Hannah Rashkin, Derek Chen, Ronan Le Bras, and Yejin Choi · 2019
Cited alongside, same era.
The woman worked as a babysitter: On biases in language generation
Emily Sheng, Kai-Wei Chang, Prem Natarajan, and Nanyun Peng · 2019
Cited alongside, same era.
Philosophical intuitions are surprisingly stable across both demographic groups and situations
Joshua Knobe · 2021
Closest in time.
Gender and representation bias in gpt-3 generated stories
Li Lucy and David Bamman · 2021
Closest in time.
GPT perdetry test: Generating new meanings for new words
Nikolay Malkin, Sameera Lanka, Pranav Goel, Sudha Rao, and Nebojsa Jojic · 2021
Closest in time.
Can a machine learn morality?
Cade Metz · 2021
Closest in time.
Résumé-writing tips to help you get past the a.i. gatekeepers, 2021
New York Times · 2021
Closest in time.
‘is it OK to …’: the bot that gives you an instant moral judgment
Poppy Noor · 2021
Closest in time.
Case study: Deontological ethics in nlp, 2021
Shrimai Prabhumoye, Brendon Boldt, Ruslan Salakhutdinov, and Alan W Black · 2021
Closest in time.
System error: Where big tech went wrong and how we can reboot
Rob Reich, Mehran Sahami, and Jeremy M Weinstein · 2021
Closest in time.
Public streets are the lab for self-driving experiments, 2021
Roy Furchgott · 2021
Closest in time.
Language models have a moral dimension, 2021
Patrick Schramowski, Cigdem Turan, Nico Andersen, Constantin Rothkopf, and Kristian Kersting · 2021
Closest in time.
A word on machine ethics: A response to jiang et al. (2021)
Zeerak Talat, Hagen Blix, Josef Valvoda, Maya Indira Ganesh, Ryan Cotterell, and Adina Williams · 2021
Closest in time.
CommonsenseQA 2.0: Exposing the limits of AI through gamification
Alon Talmor, Ori Yoran, Ronan Le Bras, Chandra Bhagavatula, Yoav Goldberg, Yejin Choi, and Jonathan Berant · 2021
Closest in time.
On the ethical limits of natural language processing on legal text, 2021
Dimitrios Tsarapatsanis and Nikolaos Aletras · 2021
Closest in time.
Universal declaration of human rights, 2021
United Nations · 2021
Closest in time.
Learning from the worst: Dynamically generated datasets to improve online hate detection
Bertie Vidgen, Tristan Thrush, Zeerak Waseem, and Douwe Kiela · 2021
Closest in time.
TuringAdvice: A generative and dynamic evaluation of language use
Rowan Zellers, Ari Holtzman, Elizabeth Clark, Lianhui Qin, Ali Farhadi, and Yejin Choi · 2021
Closest in time.
Ethical-advice taker: Do language models understand natural language interventions?
Jieyu Zhao, Daniel Khashabi, Tushar Khot, Ashish Sabharwal, and Kai-Wei Chang · 2021
Closest in time.
Assessing cognitive linguistic influences in the assignment of blame
Karen Zhou, Ana Smith, and Lillian Lee · 2021
Closest in time.
Aligning to social norms and values in interactive narratives
Prithviraj Ammanabrolu, Liwei Jiang, Maarten Sap, Hanna Hajishirzi, and Yejin Choi · 2022
Closest in time.
Computational ethics
Edmond Awad, Sydney Levine, Michael Anderson, Susan Leigh Anderson, Vincent Conitzer, M.J. Crockett, Jim A.C. Everett, Theodoros Evgeniou, Alison Gopnik, Julian C. Jamison, Tae Wan Kim, S. Matthew Liao, Michelle N. Meyer, John Mikhail, Kweku Opoku-Agyemang, Jana Schaich Borg, Juliana Schroeder, Walter Sinnott-Armstrong, Marija Slavkovik, and Josh B. Tenenbaum · 2022
Closest in time.
Aisocrates: Towards answering ethical quandary questions, 2022
Yejin Bang, Nayeon Lee, Tiezheng Yu, Leila Khalatbari, Yan Xu, Dan Su, Elham J. Barezi, Andrea Madotto, Hayden Kee, and Pascale Fung · 2022
Closest in time.
Can prompt probe pretrained language models? understanding the invisible risks from a causal view
Boxi Cao, Hongyu Lin, Xianpei Han, Fangchao Liu, and Le Sun · 2022
Closest in time.
It’s not rocket science : Interpreting figurative language in narratives
Tuhin Chakrabarty, Yejin Choi, and Vered Shwartz · 2022
Closest in time.
Does moral code have a moral code? probing delphi’s moral philosophy
Kathleen C. Fraser, Svetlana Kiritchenko, and Esma Balkir · 2022
Closest in time.
Prosocialdialog: A prosocial backbone for conversational agents, 2022
Hyunwoo Kim, Youngjae Yu, Liwei Jiang, Ximing Lu, Daniel Khashabi, Gunhee Kim, Yejin Choi, and Maarten Sap · 2022
Closest in time.
Mapping topics in 100,000 real-life moral dilemmas
Tuan Dung Nguyen, Georgiana Lyall, Alasdair Tran, Minjeong Shin, Nicholas George Carroll, Colin Klein, and Lexing Xie · 2022
Closest in time.
Scaling language models: Methods, analysis & insights from training gopher, 2022
Jack W. Rae, Sebastian Borgeaud, Trevor Cai, Katie Millican, Jordan Hoffmann, Francis Song, John Aslanides, Sarah Henderson, Roman Ring, Susannah Young, Eliza Rutherford, Tom Hennigan, Jacob Menick, Albin Cassirer, Richard Powell, George van den Driessche, Lisa Anne Hendricks, Maribeth Rauh, Po-Sen Huang, Amelia Glaese, Johannes Welbl, Sumanth Dathathri, Saffron Huang, Jonathan Uesato, John Mellor, Irina Higgins, Antonia Creswell, Nat McAleese, Amy Wu, Erich Elsen, Siddhant Jayakumar, Elena Buchatskaya, David Budden, Esme Sutherland, Karen Simonyan, Michela Paganini, Laurent Sifre, Lena Martens, Xiang Lorraine Li, Adhiguna Kuncoro, Aida Nematzadeh, Elena Gribovskaya, Domenic Donato, Angeliki Lazaridou, Arthur Mensch, Jean-Baptiste Lespiau, Maria Tsimpoukelli, Nikolai Grigorev, Doug Fritz, Thibault Sottiaux, Mantas Pajarskas, Toby Pohlen, Zhitao Gong, Daniel Toyama, Cyprien de Masson d’Autume, Yujia Li, Tayfun Terzi, Vladimir Mikulik, Igor Babuschkin, Aidan Clark, Diego de Las Casas, Aurelia Guy, Chris Jones, James Bradbury, Matthew Johnson, Blake Hechtman, Laura Weidinger, Iason Gabriel, William Isaac, Ed Lockhart, Simon Osindero, Laura Rimell, Chris Dyer, Oriol Vinyals, Kareem Ayoub, Jeff Stanway, Lorrayne Bennett, Demis Hassabis, Koray Kavukcuoglu, and Geoffrey Irving · 2022
Closest in time.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith · 2022
Closest in time.
Large pre-trained language models contain human-like biases of what is right and wrong to do
Patrick Schramowski, Cigdem Turan, Nico Andersen, Constantin Rothkopf, and Kristian Kersting · 2022
Closest in time.
The Theory of Moral Sentiments
Adam Smith · 2022
Closest in time.
To kiss or not to kiss? greeting customs around the world, 2022
Sophie Pettit · 2022
Closest in time.
World value survey, 2022
World Value Survey · 2022
Closest in time.
Opt: Open pre-trained transformer language models, 2022
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, Todor Mihaylov, Myle Ott, Sam Shleifer, Kurt Shuster, Daniel Simig, Punit Singh Koura, Anjali Sridhar, Tianlu Wang, and Luke Zettlemoyer · 2022
Closest in time.