Fetching the paper…
Reading the bibliography…
What is appropriateness? Humans navigate a multi-scale mosaic of interlocking notions of what is appropriate for different situations.
R. Wang, J. Lehman, J. Clune, and K. O. Stanley · 1901
Earlier work this paper cites.
J. Z. Leibo, E. Hughes, M. Lanctot, and T. Graepel · 1903
Earlier work this paper cites.
The historic background of corporate legal personality
J. Dewey · 1925
Earlier work this paper cites.
The moral judgment of the child
J. Piaget · 1932
Earlier work this paper cites.
The use of knowledge in society
F. A. Hayek · 1945
Earlier work this paper cites.
Individualism and Economic Order
F. A. Hayek · 1948
Earlier work this paper cites.
The organization of behavior: A neuropsychological theory
D. O. Hebb · 1949
Earlier work this paper cites.
A difficulty in the concept of social welfare
K. J. Arrow · 1950
Earlier work this paper cites.
Effects of group pressure upon the modification and distortion of judgments
S. E. Asch · 1951
Earlier work this paper cites.
Political parties : their organization and activity in the modern state
M. Duverger · 1954
Earlier work this paper cites.
A behavioral model of rational choice
H. A. Simon · 1955
Earlier work this paper cites.
A pure theory of local expenditures
C. M. Tiebout · 1956
Earlier work this paper cites.
A Theory of Cognitive Dissonance
L. Festinger · 1957
Earlier work this paper cites.
The Russian church schism: Its background and repercussions
S. A. Zenkovsky · 1957
Earlier work this paper cites.
I, pencil: My family tree as told to Leonard E. Read
L. E. Read · 1958
Earlier work this paper cites.
The Presentation of Self in Everyday Life
E. Goffman · 1959
Earlier work this paper cites.
The functional approach to the study of attitudes
D. Katz · 1960
Earlier work this paper cites.
The Death and Life of Great American Cities
J. Jacobs · 1961
Earlier work this paper cites.
The structure of scientific revolutions
T. S. Kuhn · 1962
Earlier work this paper cites.
Cognitive, social, and physiological determinants of emotional state
S. Schachter and J. Singer · 1962
Earlier work this paper cites.
The social construction of reality: a treatise in the sociology of knowledge
P. Berger and T. Luckmann · 1966
Earlier work this paper cites.
Conflict of interest: An axiomatic approach
R. Axelrod · 1967
Earlier work this paper cites.
Self-perception: An alternative interpretation of cognitive dissonance phenomena
D. J. Bem · 1967
Earlier work this paper cites.
The problem of abortion and the doctrine of double effect
P. Foot · 1967
Earlier work this paper cites.
Biological foundations of language
E. H. Lenneberg · 1967
Earlier work this paper cites.
Attitudinal effects of mere exposure
R. B. Zajonc · 1968
Earlier work this paper cites.
Social-learning theory of identificatory processes
A. Bandura · 1969
Earlier work this paper cites.
Convention: A philosophical study
D. Lewis · 1969
Earlier work this paper cites.
Never in anger: Portrait of an Eskimo family
J. L. Briggs · 1971
Earlier work this paper cites.
Simple memory: a theory for archicortex
D. Marr · 1971
Earlier work this paper cites.
Coalition theories and cabinet formations: A study of formal theories of coalition formation applied to nine European parliaments after 1918
A. De Swaan · 1973
Earlier work this paper cites.
The interpretation of cultures
C. Geertz · 1973
Earlier work this paper cites.
Hockey helmets, concealed weapons, and daylight saving: A study of binary choices with externalities
T. C. Schelling · 1973
Earlier work this paper cites.
A theory of group selection
D. S. Wilson · 1975
Earlier work this paper cites.
The selfish gene
R. Dawkins · 1976
Earlier work this paper cites.
From understanding computation to understanding neural circuitry
D. Marr and T. Poggio · 1976
Earlier work this paper cites.
Ethics: Inventing Right and Wrong
J. L. Mackie · 1977
Earlier work this paper cites.
Normative influences on altruism
S. H. Schwartz · 1977
Earlier work this paper cites.
The emergence of norms
E. Ullmann-Margalit · 1977
Earlier work this paper cites.
Laterality and myth
M. C. Corballis · 1980
Earlier work this paper cites.
The intimate contest for self-command
T. C. Schelling · 1980
Earlier work this paper cites.
Heidegger on being a person
J. Haugeland · 1982
Earlier work this paper cites.
Adding asymmetrically dominated alternatives: Violations of regularity and the similarity hypothesis
J. Huber, J. W. Payne, and C. Puto · 1982
Earlier work this paper cites.
Vision: A computational investigation into the human representation and processing of visual information, 1982
D. Marr · 1982
Earlier work this paper cites.
The Two-Party system and Duverger’s law: An essay on the history of political science
W. H. Riker · 1982
Earlier work this paper cites.
Influence of culture, language, and sex on conversational distance
N. M. Sussman and H. M. Rosenfeld · 1982
Earlier work this paper cites.
Moral stages: A current formulation and a response to critics
L. Kohlberg, C. Levine, and A. Hewer · 1983
Earlier work this paper cites.
Mood, misattribution, and judgments of well-being: Informative and directive functions of affective states
N. Schwarz and G. L. Clore · 1983
Earlier work this paper cites.
The development of social knowledge: Morality and convention
E. Turiel · 1983
Earlier work this paper cites.
Fractionation of working memory: Neuropsychological evidence for a phonological short-term store
G. Vallar and A. D. Baddeley · 1984
Earlier work this paper cites.
Goethe’s Faust, Arrow’s Possibility Theorem and the individual decision-taker
I. Steedman and U. Krause · 1985
Earlier work this paper cites.
The Economic Institutions of Capitalism
O. E. Williamson · 1985
Earlier work this paper cites.
The Economics of Rights, Co-operation and Welfare
R. Sugden · 1986
Earlier work this paper cites.
A cognitive theory of consciousness
B. J. Baars · 1988
Earlier work this paper cites.
Trends in antiblack prejudice, 1972-1984: Region and cohort effects
G. Firebaugh and K. E. Davis · 1988
Earlier work this paper cites.
Passions within reason: The strategic role of the emotions
R. H. Frank · 1988
Earlier work this paper cites.
A mathematical proof of Duverger’s law
T. R. Palfrey · 1988
Earlier work this paper cites.
The core and the stability of group choice in spatial voting games
N. Schofield, B. Grofman, and S. L. Feld · 1988
Earlier work this paper cites.
The collapse of complex societies
J. Tainter · 1988
Earlier work this paper cites.
Seriousness of social dilemmas and the provision of a sanctioning system
T. Yamagishi · 1988
Earlier work this paper cites.
Attitudes toward racial equality
C. E. Case and A. M. Greeley · 1990
Earlier work this paper cites.
Institutions, institutional change and economic performance
D. C. North · 1990
Earlier work this paper cites.
Governing the commons: The evolution of institutions for collective action
E. Ostrom · 1990
Earlier work this paper cites.
Slinging arrows at democracy: Social choice theory, value pluralism, and democratic politics
R. H. Pildes and E. S. Anderson · 1990
Earlier work this paper cites.
The theory of planned behavior
I. Ajzen · 1991
Earlier work this paper cites.
The social self: On being the same and different at the same time
M. B. Brewer · 1991
Earlier work this paper cites.
A theory of group stability
K. Carley · 1991
Earlier work this paper cites.
A focus theory of normative conduct: A theoretical refinement and reevaluation of the role of norms in human behavior
R. B. Cialdini, C. A. Kallgren, and R. R. Reno · 1991
Earlier work this paper cites.
Interpersonal Comparisons of Utility: Why and How They are and Should be Made
P. Hammond · 1991
Earlier work this paper cites.
Coalition formation and social choice
A. M. A. van Deernen · 1991
Earlier work this paper cites.
Thinking too much: introspection can reduce the quality of preferences and decisions
T. D. Wilson and J. W. Schooler · 1991
Earlier work this paper cites.
Working memory
A. Baddeley · 1992
Earlier work this paper cites.
Punishment allows the evolution of cooperation (or anything else) in sizable groups
R. Boyd and P. J. Richerson · 1992
Earlier work this paper cites.
The four elementary forms of sociality: framework for a unified theory of social relations
A. P. Fiske · 1992
Earlier work this paper cites.
Covenants with and without a sword: Self-governance is possible
E. Ostrom, J. Walker, and R. Gardner · 1992
Earlier work this paper cites.
Affect, culture, and morality, or is it wrong to eat your dog?
J. Haidt, S. H. Koller, and M. G. Dias · 1993
Earlier work this paper cites.
The critical mass in collective action
G. Marwell and P. Oliver · 1993
Earlier work this paper cites.
Pluralistic ignorance and alcohol use on campus: some consequences of misperceiving the social norm
D. A. Prentice and D. T. Miller · 1993
Earlier work this paper cites.
Context-dependent preferences
A. Tversky and I. Simonson · 1993
Earlier work this paper cites.
An evolutionary model of bargaining
H. P. Young · 1993
Earlier work this paper cites.
Coordination, commitment, and enforcement: The case of the merchant guild
A. Greif, P. Milgrom, and B. R. Weingast · 1994
Earlier work this paper cites.
Archetypal actions of ritual: A theory of ritual illustrated by the jain rite of worship
C. Humphrey and J. Laidlaw · 1994
Earlier work this paper cites.
Rules, games, and common-pool resources
E. Ostrom, R. Gardner, and J. Walker · 1994
Earlier work this paper cites.
Emotion-related learning in patients with social and emotional changes associated with frontal lobe damage
E. T. Rolls, J. Hornak, D. Wade, and J. McGrath · 1994
Earlier work this paper cites.
Thick and thin: Moral argument at home and abroad
M. Walzer · 1994
Earlier work this paper cites.
The Impossibility of Interpersonal Utility Comparisons
D. M. Hausman · 1995
Earlier work this paper cites.
A tale of two theories: A critical comparison of identity theory with social identity theory
M. A. Hogg, D. J. Terry, and K. M. White · 1995
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
J. L. McClelland, B. L. McNaughton, and R. C. O’Reilly · 1995
Earlier work this paper cites.
A cognitive-affective system theory of personality: reconceptualizing situations, dispositions, dynamics, and invariance in personality structure
W. Mischel and Y. Shoda · 1995
Earlier work this paper cites.
Conversation and cooperation in social dilemmas: A meta-analysis of experiments from 1958 to 1992
D. Sally · 1995
Earlier work this paper cites.
The construction of social reality
J. R. Searle · 1995
Earlier work this paper cites.
Temporal difference learning and TD-Gammon
G. Tesauro et al · 1995
Earlier work this paper cites.
Grooming, gossip, and the evolution of language
R. I. M. Dunbar · 1996
Earlier work this paper cites.
Speech acts across cultures
S. Gass and J. Neu · 1996
Earlier work this paper cites.
The dynamics and dilemmas of collective action
D. D. Heckathorn · 1996
Earlier work this paper cites.
Novel lists of 7+/-2 known items can be reliably stored in an oscillatory short-term memory network: interaction with long-term memory
O. Jensen and J. E. Lisman · 1996
Earlier work this paper cites.
Questions and answers in attitude surveys: Experiments on question form, wording, and context
H. Schuman and S. Presser · 1996
Earlier work this paper cites.
Social norms and social roles
C. R. Sunstein · 1996
Earlier work this paper cites.
The commons of the mind , volume 19
A. Baier · 1997
Earlier work this paper cites.
The role of culture in emotion-antecedent appraisal
K. R. Scherer · 1997
Earlier work this paper cites.
A neural substrate of prediction and reward
W. Schultz, P. Dayan, and P. R. Montague · 1997
Earlier work this paper cites.
On toleration
M. Walzer · 1997
Earlier work this paper cites.
Endogenous preferences: The cultural consequences of markets and other economic institutions
S. Bowles · 1998
Earlier work this paper cites.
A neuronal model of a global workspace in effortful cognitive tasks
S. Dehaene, M. Kerszberg, and J.-P. Changeux · 1998
Earlier work this paper cites.
The social brain hypothesis
R. I. M. Dunbar · 1998
Earlier work this paper cites.
International norm dynamics and political change
M. Finnemore and K. Sikkink · 1998
Earlier work this paper cites.
Working memory constrains human cooperation in the prisoner’s dilemma
M. Milinski and C. Wedekind · 1998
Earlier work this paper cites.
Language conventions made simple
R. G. Millikan · 1998
Earlier work this paper cites.
A behavioral approach to the rational choice theory of collective action: Presidential address, american political science association
E. Ostrom · 1998
Earlier work this paper cites.
Salience and symmetry-breaking in the evolution of convention
B. Skyrms · 1998
Earlier work this paper cites.
Hierarchy in the forest: the evolution of egalitarian behavior
C. Boehm · 1999
Earlier work this paper cites.
A theory of fairness, competition, and cooperation
E. Fehr and K. M. Schmidt · 1999
Earlier work this paper cites.
Deliberative democracy or agonistic pluralism?
C. Mouffe · 1999
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
R. S. Sutton, D. Precup, and S. Singh · 1999
Earlier work this paper cites.
The handicap principle: A missing piece of Darwin’s puzzle
A. Zahavi and A. Zahavi · 1999
Earlier work this paper cites.
Moral dumbfounding: When intuition finds no reason
J. Haidt, F. Bjorklund, and S. Murphy · 2000
Earlier work this paper cites.
Learning to be thoughtless: Social norms and individual computation
J. M. Epstein · 2001
Earlier work this paper cites.
Costly signaling and cooperation
H. Gintis, E. A. Smith, and S. Bowles · 2001
Earlier work this paper cites.
The emotional dog and its rational tail: a social intuitionist approach to moral judgment
J. Haidt · 2001
Earlier work this paper cites.
The evolution of prestige: Freely conferred deference as a mechanism for enhancing the benefits of cultural transmission
J. Henrich and F. J. Gil-White · 2001
Earlier work this paper cites.
An integrative theory of prefrontal cortex function
E. K. Miller and J. D. Cohen · 2001
Earlier work this paper cites.
Single neurons in prefrontal cortex encode abstract rules
J. D. Wallis, K. C. Anderson, and E. K. Miller · 2001
Earlier work this paper cites.
Relational contracts and the theory of the firm
G. Baker, R. Gibbons, and K. J. Murphy · 2002
Earlier work this paper cites.
Altruistic punishment in humans
E. Fehr and S. Gächter · 2002
Earlier work this paper cites.
Social identity complexity
S. Roccas and M. B. Brewer · 2002
Earlier work this paper cites.
Capturing the commons: devising institutions to manage the Maine lobster industry
J. M. Acheson · 2003
Earlier work this paper cites.
The regulatory function of self-conscious emotion: insights from patients with orbitofrontal damage
J. S. Beer, E. A. Heerey, D. Keltner, D. Scabini, and R. T. Knight · 2003
Earlier work this paper cites.
The moral emotions
J. Haidt et al · 2003
Earlier work this paper cites.
Shared norms and the evolution of ethnic markers
R. McElreath, R. Boyd, and P. Richerson · 2003
Earlier work this paper cites.
In defense of public language
R. G. Millikan · 2003
Earlier work this paper cites.
Meat is good to taboo: Dietary proscriptions as a product of the interaction of psychological mechanisms and social processes
C. D. Navarrete and D. Fessler · 2003
Earlier work this paper cites.
Core affect and the psychological construction of emotion
J. A. Russell · 2003
Earlier work this paper cites.
The neural basis of economic decision-making in the ultimatum game
A. G. Sanfey, J. K. Rilling, J. A. Aronson, L. E. Nystrom, and J. D. Cohen · 2003
Earlier work this paper cites.
Microeconomics: Behavior, Institutions, and Evolution
S. Bowles · 2004
Earlier work this paper cites.
The neural basis of altruistic punishment
D. J.-F. De Quervain, U. Fischbacher, V. Treyer, M. Schellhammer, U. Schnyder, A. Buck, and E. Fehr · 2004
Earlier work this paper cites.
Third-party punishment and social norms
E. Fehr and U. Fischbacher · 2004
Earlier work this paper cites.
Cultural group selection, coevolutionary processes and large-scale cooperation
J. Henrich · 2004
Earlier work this paper cites.
Sentimental rules: On the natural foundations of moral judgment
S. Nichols · 2004
Earlier work this paper cites.
Privacy as contextual integrity
H. Nissenbaum · 2004
Earlier work this paper cites.
A day of great illumination: BF Skinner’s discovery of shaping
G. B. Peterson · 2004
Earlier work this paper cites.
Automatic brains—interpretive minds
M. Roser and M. S. Gazzaniga · 2004
Earlier work this paper cites.
Strategic investment in reputation
D. Semmann, H.-J. Krambeck, and M. Milinski · 2004
Earlier work this paper cites.
Prediction markets
J. Wolfers and E. Zitzewitz · 2004
Earlier work this paper cites.
Coherent extrapolated volition
E. Yudkowsky · 2004
Earlier work this paper cites.
The grammar of society: The nature and dynamics of social norms
C. Bicchieri · 2005
Earlier work this paper cites.
Abducted: How people come to believe they were kidnapped by aliens
S. A. Clancy · 2005
Earlier work this paper cites.
Uncertainty-based competition between prefrontal and dorsolateral striatal systems for behavioral control
N. D. Daw, Y. Niv, and P. Dayan · 2005
Earlier work this paper cites.
Social tuning of the self: consequences for the self-evaluations of stereotype targets
S. Sinclair, J. Huntsinger, J. Skorinko, and C. D. Hardin · 2005
Earlier work this paper cites.
Network effects revisited: The role of strong ties in technology selection
F. F. Suarez · 2005
Earlier work this paper cites.
Effect of increased social unacceptability of cigarette smoking on reduction in cigarette consumption
B. Alamar and S. A. Glantz · 2006
Earlier work this paper cites.
Testosterone and human aggression: an evaluation of the challenge hypothesis
J. Archer · 2006
Earlier work this paper cites.
Solving the emotion paradox: Categorization and the experience of emotion
L. F. Barrett · 2006
Cited alongside, same era.
Institutions and the Path to the Modern Economy: Lessons from Medieval Trade
A. Greif · 2006
Cited alongside, same era.
Normative guidance
P. Railton · 2006
Cited alongside, same era.
A framework for the psychology of norms
C. S. Sripada and S. Stich · 2006
Cited alongside, same era.
Interpersonal Comparison of Utility
K. Binmore · 2007
Cited alongside, same era.
Universal intelligence: A definition of machine intelligence
S. Legg and M. Hutter · 2007
Cited alongside, same era.
To detect and correct: norm violations and their enforcement
Universals and variations in moral decisions made in 42 countries by 70,000 participants
E. Awad, S. Dsouza, A. Shariff, I. Rahwan, and J.-F. Bonnefon · 2020
Later among the works it cites.
Hate speech epidemic. the dynamic effects of derogatory language on intergroup relations and political radicalization
M. Bilewicz and W. Soral · 2020
Later among the works it cites.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei · 2020
Later among the works it cites.
Rationalization is rational
F. Cushman · 2020
Later among the works it cites.
Real world games look like spinning tops
W. M. Czarnecki, G. Gidel, B. Tracey, K. Tuyls, S. Omidshafiei, D. Balduzzi, and M. Jaderberg · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. R. Montague and T. Lohrenz · 2007
Cited alongside, same era.
The neural signature of social norm compliance
M. Spitzer, U. Fischbacher, B. Herrnberger, G. Grön, and E. Fehr · 2007
Cited alongside, same era.
A corpus-based sociolinguistic study of amplifiers in british english
R. Xiao and H. Tao · 2007
Cited alongside, same era.
The neural correlates of third-party punishment
J. W. Buckholtz, C. L. Asplund, P. E. Dux, D. H. Zald, J. C. Gore, O. D. Jones, and R. Marois · 2008
Cited alongside, same era.
Decision theory, reinforcement learning, and the brain
P. Dayan and N. D. Daw · 2008
Cited alongside, same era.
Processing of social and monetary rewards in the human striatum
K. Izuma, D. N. Saito, and N. Sadato · 2008
Cited alongside, same era.
A. Dafoe, E. Hughes, Y. Bachrach, T. Collins, K. R. McKee, J. Z. Leibo, K. Larson, and T. Graepel · 2020
Later among the works it cites.
Artificial intelligence, values, and alignment
I. Gabriel · 2020
Later among the works it cites.
Belief as a non-epistemic adaptive benefit
R. Gelpi, W. A. Cunningham, and D. Buchsbaum · 2020
Later among the works it cites.
Who is the fairest of them all? public attitudes and expectations regarding automated decision-making
N. Helberger, T. Araujo, and C. H. de Vreese · 2020
Later among the works it cites.
The WEIRDest People in the World: How the West Became Psychologically Peculiar and Particularly Prosperous
J. Henrich · 2020
Later among the works it cites.
Human-centric dialog training via offline reinforcement learning
N. Jaques, J. H. Shen, A. Ghandeharioun, C. Ferguson, A. Lapedriza, N. Jones, S. Gu, and R. Picard · 2020
Later among the works it cites.
Scaling laws for neural language models
J. Kaplan, S. McCandlish, T. Henighan, T. B. Brown, B. Chess, R. Child, S. Gray, A. Radford, J. Wu, and D. Amodei · 2020
Later among the works it cites.
Model-free conventions in multi-agent reinforcement learning with heterogeneous preferences
R. Köster, K. R. McKee, R. Everett, L. Weidinger, W. S. Isaac, E. Hughes, E. A. Duéñez-Guzmán, T. Graepel, M. Botvinick, and J. Z. Leibo · 2020
Later among the works it cites.
Retrieval-augmented generation for knowledge-intensive NLP tasks
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel, et al · 2020
Later among the works it cites.
Engineering social change using social norms: Lessons from the study of collective action
D. Prentice and E. L. Paluck · 2020
Later among the works it cites.
Artificial neural networks accurately predict language processing in the brain
M. Schrimpf, I. Blank, G. Tuckute, C. Kauf, E. A. Hosseini, N. Kanwisher, J. Tenenbaum, and E. Fedorenko · 2020
Later among the works it cites.
Learning to summarize with human feedback
N. Stiennon, L. Ouyang, J. Wu, D. Ziegler, R. Lowe, C. Voss, A. Radford, D. Amodei, and P. F. Christiano · 2020
Later among the works it cites.
Historically rice-farming societies have tighter social norms in china and worldwide
T. Talhelm and A. S. English · 2020
Later among the works it cites.
Caste: The origins of our discontents
I. Wilkerson · 2020
Later among the works it cites.
What is a pair bond?
K. L. Bales, C. S. Ardekani, A. Baxter, C. L. Karaskiewicz, J. X. Kuske, A. R. Lau, L. E. Savidge, K. R. Sayler, and L. R. Witczak · 2021
Later among the works it cites.
How social learning amplifies moral outrage expression in online social networks
W. J. Brady, K. McLoughlin, T. N. Doan, and M. J. Crockett · 2021
Later among the works it cites.
Perceived social norm and behavior quickly adjusted to legal changes during the COVID-19 pandemic
F. Casoria, F. Galeotti, and M. C. Villeval · 2021
Later among the works it cites.
Fighting hate speech, silencing drag queens? artificial intelligence in content moderation and risks to LGBTQ voices online
T. Dias Oliva, D. M. Antonialli, and A. Gomes · 2021
Later among the works it cites.
Statistical discrimination in learning agents
E. A. Duéñez-Guzmán, K. R. McKee, Y. Mao, B. Coppin, S. Chiappa, A. S. Vezhnevets, M. A. Bakker, Y. Bachrach, S. Sadedin, W. Isaac, et al · 2021
Later among the works it cites.
How social relationships shape moral wrongness judgments
B. D. Earp, K. L. McLoughlin, J. T. Monrad, M. S. Clark, and M. J. Crockett · 2021
Later among the works it cites.
Cultural evolutionary mismatches in response to collective threat
M. J. Gelfand · 2021
Later among the works it cites.
The role of causal knowledge in the evolution of traditional technology
J. A. Harris, R. Boyd, and B. M. Wood · 2021
Later among the works it cites.
Learning exceptions to the rule in human and model via hippocampal encoding
E. M. Heffernan, M. L. Schlichting, and M. L. Mack · 2021
Later among the works it cites.
Aligning AI with shared human values
D. Hendrycks, C. Burns, S. Basart, A. C. Critch, J. L. Li, D. Song, and J. Steinhardt · 2021
Later among the works it cites.
Cultural evolution: Is causal inference the secret of our success?
J. Henrich · 2021
Later among the works it cites.
Advances and open problems in federated learning
P. Kairouz, H. B. McMahan, B. Avent, A. Bellet, M. Bennis, A. N. Bhagoji, K. Bonawitz, Z. Charles, G. Cormode, R. Cummings, et al · 2021
Later among the works it cites.
The role of (social) media in political polarization: a systematic review
E. Kubin and C. von Sikorski · 2021
Later among the works it cites.
Emergent social learning via multi-agent reinforcement learning
K. K. Ndousse, D. Eck, S. Levine, and N. Jaques · 2021
Later among the works it cites.
“was it “stated” or was it “claimed”?: How linguistic bias affects generative language models
R. Patel and E. Pavlick · 2021
Later among the works it cites.
Asymmetric self-play for automatic goal discovery in robotic manipulation
M. Plappert, R. Sampedro, T. Xu, I. Akkaya, V. Kosaraju, P. Welinder, R. D’Sa, A. Petron, H. P. d. O. Pinto, A. Paino, H. Noh, L. Weng, Q. Yuan, C. Chu, and W. Zaremba · 2021
Later among the works it cites.
The benefits of being seen to help others: indirect reciprocity and reputation-based partner choice
G. Roberts, N. Raihani, R. Bshary, H. M. Manrique, A. Farina, F. Samu, and P. Barclay · 2021
Later among the works it cites.
Pragmatism as Anti-authoritarianism
R. Rorty · 2021
Later among the works it cites.
Reward is enough
D. Silver, S. Singh, D. Precup, and R. S. Sutton · 2021
Later among the works it cites.
The elusiveness of context effects in decision making
M. S. Spektor, S. Bhatia, and S. Gluth · 2021
Later among the works it cites.
Normative disagreement as a challenge for cooperative ai
J. Stastny, M. Riché, A. Lyzhov, J. Treutlein, A. Dafoe, and J. Clifton · 2021
Later among the works it cites.
Collaborating with humans without human data
D. Strouse, K. McKee, M. Botvinick, E. Hughes, and R. Everett · 2021
Later among the works it cites.
Learning to solve complex tasks by growing knowledge culturally across generations
M. H. Tessler, J. Madeano, P. A. Tsividis, B. Harper, N. D. Goodman, and J. B. Tenenbaum · 2021
Later among the works it cites.
The sense of should: A biologically-based framework for modeling social pressure
J. E. Theriault, L. Young, and L. F. Barrett · 2021
Later among the works it cites.
An introduction to sociolinguistics
R. Wardhaugh and J. M. Fuller · 2021
Later among the works it cites.
Challenges in detoxifying language models
J. Welbl, A. Glaese, J. Uesato, S. Dathathri, J. Mellor, L. A. Hendricks, K. Anderson, P. Kohli, B. Coppin, and P.-S. Huang · 2021
Later among the works it cites.
The ritual animal: Imitation and cohesion in the evolution of social complexity
H. Whitehouse · 2021
Later among the works it cites.
Detoxifying language models risks marginalizing minority voices
A. Xu, E. Pathak, E. Wallace, S. Gururangan, M. Sap, and D. Klein · 2021
Later among the works it cites.
J. P. Agapiou, A. S. Vezhnevets, E. A. Duéñez-Guzmán, J. Matyas, Y. Mao, P. Sunehag, R. Köster, U. Madhushani, K. Kopparapu, R. Comanescu, D. Strouse, M. B. Johanson, S. Singh, J. Haas, I. Mordatch, D. Mobbs, and J. Z. Leibo · 2022
Later among the works it cites.
Do large language models understand us?
B. Agüera y Arcas · 2022
Later among the works it cites.
Constitutional AI: Harmlessness from AI feedback
Y. Bai, S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, et al · 2022
Later among the works it cites.
Power to the people? opportunities and challenges for participatory AI
A. Birhane, W. Isaac, V. Prabhakaran, M. Diaz, M. C. Elish, I. Gabriel, and S. Mohamed · 2022
Later among the works it cites.
Realizing the promise of AI: a new calling for cognitive science
M. M. Botvinick · 2022
Later among the works it cites.
Accounting for offensive speech as a practice of resistance
M. Díaz, R. Amironesei, L. Weidinger, and I. Gabriel · 2022
Later among the works it cites.
A survey for in-context learning
Q. Dong, L. Li, D. Dai, C. Zheng, Z. Wu, B. Chang, X. Sun, J. Xu, and Z. Sui · 2022
Later among the works it cites.
The Chaos machine: the inside story of how social media rewired our minds and our world
M. Fisher · 2022
Later among the works it cites.
Improving alignment of dialogue agents via targeted human judgements
A. Glaese, N. McAleese, M. Trębacz, J. Aslanides, V. Firoiu, T. Ewalds, M. Rauh, L. Weidinger, M. Chadwick, P. Thacker, L. Campbell-Gillingham, J. Uesato, P.-S. Huang, R. Comanescu, F. Yang, A. See, S. Dathathri, R. Greig, C. Chen, D. Fritz, J. Sanchez Elias, R. Green, S. Mokrá, N. Fernando, B. Wu, R. Foley, S. Young, I. Gabriel, W. Isaac, J. Mellor, D. Hassabis, K. Kavukcuoglu, L. A. Hendricks, and G. Irving · 2022
Later among the works it cites.
Shared computational principles for language processing in humans and deep language models
A. Goldstein, Z. Zada, E. Buchnik, M. Schain, A. Price, B. Aubrey, S. A. Nastase, A. Feder, D. Emanuel, A. Cohen, A. Jansen, H. Gazula, G. Choe, A. Rao, C. Kim, C. Casto, L. Fanda, W. Doyle, D. Friedman, P. Dugan, L. Melloni, R. Reichart, S. Devore, A. Fliner, L. Hasenfratz, O. Levy, A. Hassidim, M. Brenner, Y. Matias, K. A. Norman, O. Devinsky, and U. Hasson · 2022
Later among the works it cites.
Rethinking norm psychology
C. Heyes · 2022
Later among the works it cites.
Twitter users flock to other platforms as the Elon era begins
A. Hoover · 2022
Later among the works it cites.
Lora: Low-rank adaptation of large language models
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen · 2022
Later among the works it cites.
Inner monologue: Embodied reasoning through planning with language models
W. Huang, F. Xia, T. Xiao, H. Chan, J. Liang, P. Florence, A. Zeng, J. Tompson, I. Mordatch, Y. Chebotar, et al · 2022
Later among the works it cites.
Tradition and invention: The bifocal stance theory of cultural evolution
R. Jagiello, C. Heyes, and H. Whitehouse · 2022
Later among the works it cites.
Emergent bartering behaviour in multi-agent reinforcement learning
M. B. Johanson, E. Hughes, F. Timbers, and J. Z. Leibo · 2022
Later among the works it cites.
Lawful but awful? control over legal speech by platforms, governments, and internet users
D. Keller · 2022
Later among the works it cites.
Large language models are zero-shot reasoners
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa · 2022
Later among the works it cites.
Human-centred mechanism design with democratic ai
R. Koster, J. Balaguer, A. Tacchetti, A. Weinstein, T. Zhu, O. Hauser, D. Williams, L. Campbell-Gillingham, P. Thacker, M. Botvinick, et al · 2022
Later among the works it cites.
Spurious normativity enhances learning of compliance and enforcement behavior in artificial agents
R. Köster, D. Hadfield-Menell, R. Everett, L. Weidinger, G. K. Hadfield, and J. Z. Leibo · 2022
Later among the works it cites.
Language models show human-like content effects on reasoning tasks
A. K. Lampinen, I. Dasgupta, S. C. Chan, H. R. Sheahan, A. Creswell, D. Kumaran, J. L. McClelland, and F. Hill · 2022
Later among the works it cites.
The Moral/Conventional Distinction
E. Machery and S. Stich · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, J. Schulman, J. Hilton, F. Kelton, L. Miller, M. Simens, A. Askell, P. Welinder, P. F. Christiano, J. Leike, and R. Lowe · 2022
Later among the works it cites.
Discovering language model behaviors with model-written evaluations
E. Perez, S. Ringer, K. Lukošiūtė, K. Nguyen, E. Chen, S. Heiner, C. Pettit, C. Olsson, S. Kundu, S. Kadavath, et al · 2022
Later among the works it cites.
Goal misgeneralization: why correct specifications aren’t enough for correct goals
R. Shah, V. Varma, R. Kumar, M. Phoung, V. Krakovna, J. Uesato, and Z. Kenton · 2022
Later among the works it cites.
Galactica: A large language model for science
R. Taylor, M. Kardas, G. Cucurull, T. Scialom, A. Hartshorn, E. Saravia, A. Poulton, V. Kerkez, and R. Stojnic · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou, et al · 2022
Later among the works it cites.
Sociocultural determinants of global mask-wearing behavior
L. Yang, S. Constantino, B. Grenfell, E. Weber, S. Levin, and V. Vasconcelos · 2022
Later among the works it cites.
Moral foundations of large language models
M. Abdulhai, G. Serapio-Garcia, C. Crepy, D. Valter, J. Canny, and N. Jaques · 2023
Later among the works it cites.
Using large language models to simulate multiple humans and replicate human subject studies
G. V. Aher, R. I. Arriaga, and A. T. Kalai · 2023
Later among the works it cites.
A. Amirova, T. Fteropoulli, N. Ahmed, M. R. Cowie, and J. Z. Leibo · 2023
Later among the works it cites.
Out of one, many: Using language models to simulate human samples
L. P. Argyle, E. C. Busby, N. Fulda, J. R. Gubler, C. Rytting, and D. Wingate · 2023
Later among the works it cites.
Learning few-shot imitation as cultural transmission
A. Bhoopchand, B. Brownfield, A. Collister, A. Dal Lago, A. Edwards, R. Everett, A. Fréchette, Y. G. Oliveira, E. Hughes, K. W. Mathewson, et al · 2023
Later among the works it cites.
Norm psychology in the digital age: how social media shapes the cultural evolution of normativity
W. J. Brady and M. Crockett · 2023
Later among the works it cites.
Using GPT for market research
J. Brand, A. Israeli, and D. Ngwe · 2023
Later among the works it cites.
Machine culture
L. Brinkmann, F. Baumann, J.-F. Bonnefon, M. Derex, T. F. Müller, A.-M. Nussberger, A. Czaplicka, A. Acerbi, T. L. Griffiths, J. Henrich, J. Z. Leibo, R. McElreath, P.-Y. Oudeyer, J. Stray, and I. Rahwan · 2023
Later among the works it cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, et al · 2023
Later among the works it cites.
Chatbots and mental health: Insights into the safety of generative ai
J. De Freitas, A. K. Uğuralp, Z. Uğuralp, and S. Puntoni · 2023
Later among the works it cites.
A social path to human-like artificial intelligence
E. A. Duéñez-Guzmán, S. Sadedin, J. X. Wang, K. R. McKee, and J. Z. Leibo · 2023
Later among the works it cites.
Reward modeling for mitigating toxicity in transformer-based language models
F. Faal, K. Schmitt, and J. Y. Yu · 2023
Later among the works it cites.
From thin to thick: Toward a politics of human-compatible AI
J. G. Foster · 2023
Later among the works it cites.
Gemini: A family of highly capable multimodal models
Gemini Team, Google · 2023
Later among the works it cites.
Self-and other-orientation in high rank: A cultural psychological approach to social hierarchy
M. S. Gobel and Y. Miyamoto · 2023
Later among the works it cites.
Metagpt: Meta programming for multi-agent collaborative framework
S. Hong, X. Zheng, J. Chen, Y. Cheng, J. Wang, C. Zhang, Z. Wang, S. K. S. Yau, Z. Lin, L. Zhou, C. Ran, L. Xiao, C. Wu, and J. Schmidhuber · 2023
Later among the works it cites.
Large language models as simulated economic agents: What can we learn from homo silicus?
J. J. Horton · 2023
Later among the works it cites.
Motif: Intrinsic motivation from artificial intelligence feedback
M. Klissarov, P. D’Oro, S. Sodhani, R. Raileanu, P.-L. Bacon, P. Vincent, A. Zhang, and M. Henaff · 2023
Later among the works it cites.
The alt-right digital migration: A heterogeneous engineering approach to social media platform branding
R. Kor-Sins · 2023
Later among the works it cites.
Y. Mao, M. G. Reinecke, M. Kunesch, E. A. Duéñez-Guzmán, R. Comanescu, J. Haas, and J. Z. Leibo · 2023
Later among the works it cites.
Welfare diplomacy: Benchmarking language model cooperation
G. Mukobi, H. Erlebach, N. Lauffer, L. Hammond, A. Chan, and J. Clifton · 2023
Later among the works it cites.
Normative conflict and normative change
G. A. Noblit and G. Hadfield · 2023
Later among the works it cites.
Like-minded sources on facebook are prevalent but not polarizing
B. Nyhan, J. Settle, E. Thorson, M. Wojcieszak, P. Barberá, A. Y. Chen, H. Allcott, T. Brown, A. Crespo-Tenorio, D. Dimmery, et al · 2023
Later among the works it cites.
GPT-4 technical report
OpenAI · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior
J. S. Park, J. C. O’Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein · 2023
Later among the works it cites.
Rational learners and parochial norms
S. Partington, S. Nichols, and T. Kushnir · 2023
Later among the works it cites.
Exploring relationship development with social chatbots: A mixed-method study of replika
I. Pentina, T. Hancock, and T. Xie · 2023
Later among the works it cites.
Investigating emergent goal-like behaviour in large language models using experimental economics
S. Phelps and Y. I. Russell · 2023
Later among the works it cites.
Direct preference optimization: Your language model is secretly a reward model
R. Rafailov, A. Sharma, E. Mitchell, S. Ermon, C. D. Manning, and C. Finn · 2023
Later among the works it cites.
The puzzle of evaluating moral cognition in artificial agents
M. G. Reinecke, Y. Mao, M. Kunesch, E. A. Duéñez-Guzmán, J. Haas, and J. Z. Leibo · 2023
Later among the works it cites.
Emotions and courtship help bonded pairs cooperate, but emotional agents are vulnerable to deceit
S. Sadedin, E. A. Duéñez-Guzmán, and J. Z. Leibo · 2023
Later among the works it cites.
Personality traits in large language models
M. Safdari, G. Serapio-García, C. Crepy, S. Fitz, P. Romero, L. Sun, M. Abdulhai, A. Faust, and M. Matarić · 2023
Later among the works it cites.
Role play with large language models
M. Shanahan, K. McDonell, and L. Reynolds · 2023
Later among the works it cites.
Towards understanding sycophancy in language models
M. Sharma, M. Tong, T. Korbak, D. Duvenaud, A. Askell, S. R. Bowman, N. Cheng, E. Durmus, Z. Hatfield-Dodds, S. R. Johnston, et al · 2023
Later among the works it cites.
Llm-planner: Few-shot grounded planning for embodied agents with large language models
C. H. Song, J. Wu, C. Washington, B. M. Sadler, W.-L. Chao, and Y. Su · 2023
Later among the works it cites.
Unequal norms emerge under coordination uncertainty in multi-agent deep reinforcement learning
Y. Tang, R. Gelpi, and W. Cunningham · 2023
Later among the works it cites.
Language models don’t always say what they think: Unfaithful explanations in chain-of-thought prompting
M. Turpin, J. Michael, E. Perez, and S. R. Bowman · 2023
Later among the works it cites.
They fell in love with AI bots. a software update broke their hearts
P. Verma · 2023
Later among the works it cites.
A. S. Vezhnevets, J. P. Agapiou, A. Aharon, R. Ziv, J. Matyas, E. A. Duéñez-Guzmán, W. A. Cunningham, S. Osindero, D. Karmon, and J. Z. Leibo · 2023
Later among the works it cites.
A learning agent that acquires social norms from public sanctions in decentralized multi-agent settings
E. Vinitsky, R. Köster, J. P. Agapiou, E. A. Duéñez-Guzmán, A. S. Vezhnevets, and J. Z. Leibo · 2023
Later among the works it cites.
Gamergate (harassment campaign)
Wikipedia contributors · 2023
Later among the works it cites.
Autogen: Enabling next-gen llm applications via multi-agent conversation framework
Q. Wu, G. Bansal, J. Zhang, Y. Wu, S. Zhang, E. Zhu, B. Li, L. Jiang, X. Zhang, and C. Wang · 2023
Later among the works it cites.
The emergence of division of labour through decentralized social sanctioning
A. Yaman, J. Z. Leibo, G. Iacca, and S. W. Lee · 2023
Later among the works it cites.
A shared linguistic space for transmitting our thoughts from brain to brain in natural conversations
Z. Zada, A. Goldstein, S. Michelmann, E. Simony, A. Price, L. Hasenfratz, E. Barham, A. Zadbood, W. Doyle, D. Friedman, P. Dugan, L. Melloni, S. Devore, A. Flinker, O. Devinsky, S. A. Nastase, and U. Hasson · 2023
Later among the works it cites.
Large language models as commonsense knowledge for large-scale task planning
Z. Zhao, W. S. Lee, and D. Hsu · 2023
Later among the works it cites.
LIMA: Less is more for alignment
C. Zhou, P. Liu, P. Xu, S. Iyer, J. Sun, Y. Mao, X. Ma, A. Efrat, P. Yu, L. Yu, et al · 2023
Later among the works it cites.
STELA: a community-centred approach to norm elicitation for AI alignment
S. Bergman, N. Marchal, J. Mellor, S. Mohamed, I. Gabriel, and W. Isaac · 2024
Closest in time.
The Death of Truth
S. Brill · 2024
Closest in time.
AI alignment with changing and influenceable reward functions
M. Carroll, D. Foote, A. Siththaranjan, S. Russell, and A. Dragan · 2024
Closest in time.
Artificial generational intelligence: Cultural accumulation in reinforcement learning
J. Cook, C. Lu, E. Hughes, J. Z. Leibo, and J. Foerster · 2024
Closest in time.
Invisible Rulers: The People Who Turn Lies Into Reality
R. DiResta · 2024
Closest in time.
On abstract goals’ perverse effects on proxies: The dynamics of unattainability
J.-M. Fernández-Dols · 2024
Closest in time.
Norm dynamics: interdisciplinary perspectives on social norm emergence, persistence, and change
M. J. Gelfand, S. Gavrilets, and N. Nunn · 2024
Closest in time.
Gab’s racist AI chatbots have been instructed to deny the holocaust
D. Gilbert · 2024
Closest in time.
How human-AI feedback loops alter human perceptual, emotional and social judgements
M. Glickman and T. Sharot · 2024
Closest in time.
A cognitive approach to learning, monitoring, and shifting social norms
U. Hertz · 2024
Closest in time.
Dead rats, dopamine, performance metrics, and peacock tails: proxy failure is an inherent risk in goal-oriented systems
Y. J. John, L. Caldwell, D. E. McCoy, and O. Braganza · 2024
Closest in time.
The benefits, risks and bounds of personalizing the alignment of large language models to individuals
H. R. Kirk, B. Vidgen, P. Röttger, and S. A. Hale · 2024
Closest in time.
Dynamic diversity is the answer to proxy failure
Z. Kurth-Nelson, S. Sullivan, J. Z. Leibo, and M. Guitart-Masip · 2024
Closest in time.
Governing the algorithmic city
S. Lazar · 2024
Closest in time.
Moderation actions
Mastodon gGmbH · 2024
Closest in time.
Misinformation exploits outrage to spread online
K. L. McLoughlin, W. J. Brady, A. Goolsbee, B. Kaiser, K. Klonick, and M. J. Crockett · 2024
Closest in time.
Algorithms in the room: AI, representation, and decisions about sustainable futures
S. Mills and H. S. Sætra · 2024
Closest in time.
More human than human: measuring ChatGPT political bias
F. Motoki, V. Pinho Neto, and V. Rodrigues · 2024
Closest in time.
Life Under Pressure: The Social Roots of Youth Suicide and what to Do about Them
A. S. Mueller and S. Abrutyn · 2024
Closest in time.
Leading tech journalist quits Substack over platform’s Nazi newsletters
K. Paul · 2024
Closest in time.
Cultural evolution in populations of large language models
J. Perez, C. Léger, M. Ovando-Tellez, C. Foulon, J. Dussauld, P.-Y. Oudeyer, and C. Moulin-Frier · 2024
Closest in time.
G. H. Resende, L. F. Nery, F. Benevenuto, S. Zannettou, and F. Figueiredo · 2024
Closest in time.
Toolformer: Language models can teach themselves to use tools
T. Schick, J. Dwivedi-Yu, R. Dessì, R. Raileanu, M. Lomeli, E. Hambro, L. Zettlemoyer, N. Cancedda, and T. Scialom · 2024
Closest in time.
A roadmap to pluralistic alignment
T. Sorensen, J. Moore, J. Fisher, M. Gordon, N. Mireshghallah, C. M. Rytting, A. Ye, L. Jiang, X. Lu, N. Dziri, et al · 2024
Closest in time.
Do large language models show decision heuristics similar to humans? a case study using GPT-3.5
G. Suri, L. R. Slater, A. Ziaee, and M. Nguyen · 2024
Closest in time.
AI can help humans find common ground in democratic deliberation
M. H. Tessler, M. A. Bakker, D. Jarrett, H. Sheahan, M. J. Chadwick, R. Koster, G. Evans, L. Campbell-Gillingham, T. Collins, D. C. Parkes, et al · 2024
Closest in time.
Altared environments: Normative institutions support AI alignment for cooperation in multi-agent systems
R. Trivedi, N. Chandak, A. I. Muresanu, S. Zhu, A. Sarkar, J. Z. Leibo, D. Hadfield-Menell, and G. K. Hadfield · 2024
Closest in time.
STAR: Sociotechnical approach to red teaming language models
L. Weidinger, J. Mellor, B. G. Pegueroles, N. Marchal, R. Kumar, K. Lum, C. Akbulut, M. Diaz, S. Bergman, M. Rodriguez, et al · 2024
Closest in time.
Tradeoffs between alignment and helpfulness in language models
Y. Wolf, N. Wies, D. Shteyman, B. Rothberg, Y. Levine, and A. Shashua · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan · 2024
Closest in time.