Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated impressive capabilities in generating fluent text, as well as tendencies to reproduce undesirable social biases.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
Culture’s Recent Consequences: Using Dimension Scores in Theory and Research
Geert Hofstede. 2001 · 2001
Earlier work this paper cites.
Do liberals and conservatives use different moral languages? Two replications and six extensions of Graham, Haidt, and Nosek’s (2009) moral text analysis
Jeremy A. Frimer. 2020 · 2009
Earlier work this paper cites.
Liberals and conservatives rely on different sets of moral foundations
Jesse Graham, Jonathan Haidt, and Brian A. Nosek. 2009 · 2009
Earlier work this paper cites.
Mapping the Moral Domain
Jesse Graham, Brian A. Nosek, Jonathan Haidt, Ravi Iyer, Spassena Koleva, and Peter H. Ditto. 2011 · 2011
Earlier work this paper cites.
Can Innate, Modular “Foundations” Explain Morality? Challenges for Haidt’s Moral Foundations Theory
Christopher Suhler and Pat Churchland. 2011 · 2011
Earlier work this paper cites.
The Righteous Mind: Why Good People Are Divided by Politics and Religion
Jonathan Haidt. 2013 · 2013
Earlier work this paper cites.
From Gulf to Bridge: When Do Moral Arguments Facilitate Political Influence?
Matthew Feinberg and Robb Willer. 2015 · 2015
Earlier work this paper cites.
The moral foundations hypothesis does not replicate well in Black samples
Don E. Davis, Kenneth Rice, Daryl R. Van Tongeren, Joshua N. Hook, Cirleen DeBlaere, Everett L. Worthington Jr., and Elise Choe. 2016 · 2016
Earlier work this paper cites.
Critiques | Moral Foundations Theory
David Dobolyi. 2016 · 2016
Earlier work this paper cites.
Morality Between the Lines : Detecting Moral Sentiment In Text
Justin Garten, Reihane Boghrati, J. Hoover, Kate M. Johnson, and Morteza Dehghani. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Classification of Moral Foundations in Microblog Political Discourse
Kristen Johnson and Dan Goldwasser. 2018 · 2018
Earlier work this paper cites.
The five-factor model of the moral foundations theory is stable across WEIRD and non-WEIRD cultures
Burak Doğruyol, Sinan Alper, and Onurcan Yilmaz. 2019 · 2019
Earlier work this paper cites.
Moral Foundations Dictionary 2.0
Jeremy Frimer. 2019 · 2019
Earlier work this paper cites.
Challenging Moral Attitudes With Moral Messages
Andrew Luttrell, Aviva Philipp-Muller, and Richard E. Petty. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Earlier work this paper cites.
An Investigation of Moral Foundations Theory in Turkey Using Different Measures
Bilge Yalçındağ, Türker Özkan, Sevim Cesur, Onurcan Yilmaz, Beyza Tepe, Zeynep Ecem Piyale, Ali Furkan Biten, and Diane Sunar. 2019 · 2019
Earlier work this paper cites.
Climbing towards NLU: On Meaning, Form, and Understanding in the Age of Data
Emily M. Bender and Alexander Koller. 2020 · 2020
Earlier work this paper cites.
It’s a Match: Moralization and the Effects of Moral Foundations Congruence on Ethical and Unethical Leadership Perception
Maxim Egorov, Karianne Kalshoven, Armin Pircher Verdorfer, and Claudia Peus. 2020 · 2020
Cited alongside, same era.
Social chemistry 101: Learning to reason about social and moral norms
Maxwell Forbes, Jena D. Hwang, Vered Shwartz, Maarten Sap, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Scaling Laws for Neural Language Models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
Do Bots Have Moral Judgement? The Difference Between Bots and Humans in Moral Rhetoric
Ece Çiğdem Mutlu, Toktam Oghaz, Ege Tütüncüler, and Ivan Garibay. 2020 · 2020
Cited alongside, same era.
Morality Classification in Natural Language Text
Matheus C. Pavan, Vitor G. Dos Santos, Alex G. J. Lan, Joao Martins, Wesley R. Santos, Caio Deutsch, Pablo B. Costa, Fernando C. Hsieh, and Ivandre Paraboni. 2020 · 2020
Cited alongside, same era.
Does Moral Code have a Moral Code? Probing Delphi’s Moral Philosophy
Kathleen C. Fraser, Svetlana Kiritchenko, and Esma Balkir. 2022 · 2022
Closest in time.
CommunityLM: Probing partisan worldviews from language models
Hang Jiang, Doug Beeferman, Brandon Roy, and Deb Roy. 2022 · 2022
Closest in time.
When to make exceptions: Exploring language models as accounts of human moral judgment
Zhijing Jin, Sydney Levine, Fernando Gonzalez Adauto, Ojasv Kamal, Maarten Sap, Mrinmaya Sachan, Rada Mihalcea, Josh Tenenbaum, and Bernhard Schölkopf. 2022 · 2022
Closest in time.
Moral Frames Are Persuasive and Moralize Attitudes; Nonmoral Frames Are Persuasive and De-Moralize Attitudes
Rabia I. Kodapanakkal, Mark J. Brandt, Christoph Kogler, and Ilja van Beest. 2022 · 2022
Closest in time.
Model Index for Researchers
OpenAI. 2022 · 2022
Closest in time.
Discovering Language Model Behaviors with Model-Written Evaluations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Cited alongside, same era.
Belief-based Generation of Argumentative Claims
Milad Alshomary, Wei-Fan Chen, Timon Gurcke, and Henning Wachsmuth. 2021 · 2021
Cited alongside, same era.
Toward audience-aware argument generation
Milad Alshomary and Henning Wachsmuth. 2021 · 2021
Cited alongside, same era.
Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences
Denis Emelin, Ronan Le Bras, Jena D. Hwang, Maxwell Forbes, and Yejin Choi. 2021 · 2021
Cited alongside, same era.
The dark side of the ‘Moral Machine’ and the fallacy of computational ethical decision-making for autonomous vehicles
Hubert Etienne. 2021 · 2021
Cited alongside, same era.
On the Sizes of OpenAI API Models
Leo Gao. 2021 · 2021
Cited alongside, same era.
Reanalysing the factor structure of the moral foundations questionnaire
Craig A. Harper and Darren Rhodes. 2021 · 2021
Cited alongside, same era.
Ethan Perez, Sam Ringer, Kamilė Lukošiūtė, Karina Nguyen, Edwin Chen, Scott Heiner, Craig Pettit, Catherine Olsson, Sandipan Kundu, Saurav Kadavath, Andy Jones, Anna Chen, Ben Mann, Brian Israel, Bryan Seethor, Cameron McKinnon, Christopher Olah, Da Yan, Daniela Amodei, Dario Amodei, Dawn Drain, Dustin Li, Eli Tran-Johnson, Guro Khundadze, Jackson Kernion, James Landis, Jamie Kerr, Jared Mueller, Jeeyoon Hyun, Joshua Landau, Kamal Ndousse, Landon Goldberg, Liane Lovitt, Martin Lucas, Michael Sellitto, Miranda Zhang, Neerav Kingsland, Nelson Elhage, Nicholas Joseph, Noemí Mercado, Nova DasSarma, Oliver Rausch, Robin Larson, Sam McCandlish, Scott Johnston, Shauna Kravec, Sheer El Showk, Tamera Lanham, Timothy Telleen-Lawton, Tom Brown, Tom Henighan, Tristan Hume, Yuntao Bai, Zac Hatfield-Dodds, Jack Clark, Samuel R. Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan. 2022 · 2022
Closest in time.
Towards Few-Shot Identification of Morality Frames using In-Context Learning
Shamik Roy, Nishanth Sridhar Nakshatri, and Dan Goldwasser. 2022 · 2022
Closest in time.
WVS Database
World Values Survey. 2022 · 2022
Closest in time.
On the Machine Learning of Ethical Judgments from Natural Language
Zeerak Talat, Hagen Blix, Josef Valvoda, Maya Indira Ganesh, Ryan Cotterell, and Adina Williams. 2022 · 2022
Closest in time.
The Google engineer who thinks the company’s AI has come to life
Nitasha Tiku. 2022 · 2022
Closest in time.
Taxonomy of Risks posed by Language Models
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, John Mellor, Amelia Glaese, Myra Cheng, Borja Balle, Atoosa Kasirzadeh, Courtney Biles, Sasha Brown, Zac Kenton, Will Hawkins, Tom Stepleton, Abeba Birhane, Lisa Anne Hendricks, Laura Rimell, William Isaac, Julia Haas, Sean Legassick, Geoffrey Irving, and Iason Gabriel. 2022 · 2022
Closest in time.
OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, Todor Mihaylov, Myle Ott, Sam Shleifer, Kurt Shuster, Daniel Simig, Punit Singh Koura, Anjali Sridhar, Tianlu Wang, and Luke Zettlemoyer. 2022 · 2022
Closest in time.
The moral integrity corpus: A benchmark for ethical dialogue systems
Caleb Ziems, Jane Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang. 2022 · 2022
Closest in time.
Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies
Gati Aher, Rosa I. Arriaga, and Adam Tauman Kalai. 2023 · 2023
Closest in time.
Out of One, Many: Using Language Models to Simulate Human Samples
Lisa P. Argyle, Ethan C. Busby, Nancy Fulda, Joshua R. Gubler, Christopher Rytting, and David Wingate. 2023 · 2023
Closest in time.
Probing pre-trained language models for cross-cultural differences in values
Arnav Arora, Lucie-aimee Kaffee, and Isabelle Augenstein. 2023 · 2023
Closest in time.
OPT-IML: Scaling Language Model Instruction Meta Learning through the Lens of Generalization
Srinivasan Iyer, Xi Victoria Lin, Ramakanth Pasunuru, Todor Mihaylov, Daniel Simig, Ping Yu, Kurt Shuster, Tianlu Wang, Qing Liu, Punit Singh Koura, Xian Li, Brian O’Horo, Gabriel Pereyra, Jeff Wang, Christopher Dewan, Asli Celikyilmaz, Luke Zettlemoyer, and Ves Stoyanov. 2023 · 2023
Closest in time.
LLaMA: Open and Efficient Foundation Language Models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.