Fetching the paper…
Reading the bibliography…
Recent generative AI systems have demonstrated more advanced persuasive capabilities and are increasingly permeating areas of life where they can influence decision-making.
Persuasion for good: towards a personalized persuasive dialogue system for social good, Jan. 2020
X. Wang, W. Shi, R. Kim, Y. Oh, S. Yang, J. Zhang, and Z. Yu · 1906
Earlier work this paper cites.
A dual-motive model of scapegoating: displacing blame to reduce guilt or increase control
Z. K. Rothschild, M. J. Landau, D. Sullivan, and L. A. Keefer · 1939
Earlier work this paper cites.
Logic and conversation
H. P. Grice · 1975
Earlier work this paper cites.
Legitimation crisis , volume 519
J. Habermas · 1975
Earlier work this paper cites.
Prospect theory: an analysis of decision under risk
D. Kahneman and A. Tversky · 1979
Earlier work this paper cites.
Persuasion and manipulation
R. Harré · 1985
Earlier work this paper cites.
A history and theory of informed consent
R. R. Faden, T. L. Beauchamp, and N. M. P. King · 1986
Earlier work this paper cites.
Advances in prospect theory: cumulative representation of uncertainty
A. Tversky and D. Kahneman · 1992
Earlier work this paper cites.
Consumer choice in context: the decoy effect in travel and tourism
B. M. Josiam and J. P. Hobson · 1995
Earlier work this paper cites.
Using language
H. H. Clark · 1996
Earlier work this paper cites.
Silicon sycophants: the effects of computers that flatter
B. Fogg and C. Nass · 1996
Earlier work this paper cites.
Anthropomorphism and the evolution of cognition
S. Mithen and P. Boyer · 1996
Earlier work this paper cites.
The impact of computer-tailored feedback and iterative feedback on fat, fruit, and vegetable intake
J. Brug, K. Glanz, P. Van Assema, G. Kok, and G. J. Van Breukelen · 1998
Earlier work this paper cites.
Bounded rationality
B. D. Jones · 1999
Earlier work this paper cites.
Kairos revisited: an interview with James Kinneavy
R. Thompson · 2000
Earlier work this paper cites.
The TARES test: five principles for ethical persuasion
S. Baker and D. L. Martinson · 2001
Earlier work this paper cites.
Trust and distrust definitions: one bite at a time
D. H. McKnight and N. L. Chervany · 2001
Earlier work this paper cites.
The logic of temptation
P. M. Hughes · 2002
Earlier work this paper cites.
Personalizition: definition, status and challenges ahead
W. Kim · 2002
Earlier work this paper cites.
Guilt as a mechanism of persuasion
D. J. O’Keefe · 2002
Earlier work this paper cites.
Being vulnerable and being ethical with/in research
K. Tisdale · 2003
Earlier work this paper cites.
Personalisation and trust: a reciprocal relationship?
P. Briggs, B. Simpson, and A. De Angeli · 2004
Earlier work this paper cites.
The science of persuasion
R. B. Cialdini · 2004
Earlier work this paper cites.
Narrative techniques of fear mongering
B. Glassner · 2004
Earlier work this paper cites.
Classifying constructive comments, 2020
V. Kolhatkar, N. Thain, J. Sorensen, L. Dixon, and M. Taboada · 2004
Earlier work this paper cites.
Language models are few-shot learners, 2020
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei · 2005
Earlier work this paper cites.
Deliberation and democratic legitimacy
J. Cohen · 2005
Earlier work this paper cites.
Explaining theories of persuasion
M. Dainton and E. D. Zelley · 2005
Earlier work this paper cites.
The boundaries of loss aversion
N. Novemsky and D. Kahneman · 2005
Earlier work this paper cites.
Gut feelings: the intelligence of the unconscious
G. Gigerenzer · 2007
Earlier work this paper cites.
I like you, but I won’t listen to you: effects of rationality on affective and behavioral responses to computers that flatter
E.-J. Lee · 2009
Earlier work this paper cites.
Learning to summarize from human feedback, 2020
N. Stiennon, L. Ouyang, J. Wu, D. M. Ziegler, R. Lowe, C. Voss, A. Radford, D. Amodei, and P. Christiano · 2009
Earlier work this paper cites.
A literature review of the anchoring effect
A. Furnham and H. C. Boo · 2010
Earlier work this paper cites.
What triggers social responses to flattering computers? Experimental tests of anthropomorphism and mindlessness explanations
E.-J. Lee · 2010
Earlier work this paper cites.
Who sees human? The stability and importance of individual differences in anthropomorphism
A. Waytz, J. Cacioppo, and N. Epley · 2010
Earlier work this paper cites.
The big five personality traits in the political arena
A. S. Gerber, G. A. Huber, D. Doherty, and C. M. Dowling · 2011
Earlier work this paper cites.
Why do humans reason? Arguments for an argumentative theory
H. Mercier and D. Sperber · 2011
Earlier work this paper cites.
A friendly gesture: investigating the effect of multimodal robot behavior in human-robot interaction
M. Salem, K. Rohlfing, S. Kopp, and F. Joublin · 2011
Earlier work this paper cites.
Othering and its effects: exploring the concept
A. Velho and O. Thomas-Olalde · 2011
Earlier work this paper cites.
Designing effective gaze mechanisms for virtual agents
S. Andrist, T. Pejsa, B. Mutlu, and M. Gleicher · 2012
Earlier work this paper cites.
Between reason and coercion: ethically permissible influence in health care and health policy contexts
J. S. Blumenthal-Barby · 2012
Earlier work this paper cites.
Feeling robots and human zombies: mind perception and the uncanny valley
K. Gray and D. M. Wegner · 2012
Earlier work this paper cites.
Truth and rationality
D. Hodgson · 2012
Earlier work this paper cites.
On being persuaded: Some basic distinctions
G. R. Miller · 2013
Earlier work this paper cites.
Affect and persuasion
J. Price Dillard and K. Seo · 2013
Earlier work this paper cites.
Trusting digital chameleons: The effect of mimicry by a virtual social agent on user trust
F. M. Verberne, J. Ham, A. Ponnada, and C. J. Midden · 2013
Earlier work this paper cites.
Cognitive biases and heuristics in medical decision making: a critical review using a systematic search strategy
J. S. Blumenthal-Barby and H. Krieger · 2014
Earlier work this paper cites.
You, robot
B. Fiala, A. Arico, and S. Nichols · 2014
Earlier work this paper cites.
Instinct can beat analytical thinking, June 2014
J. Fox · 2014
Earlier work this paper cites.
Coercion, manipulation, exploitation
A. W. Wood · 2014
Earlier work this paper cites.
Strategies of persuasion, manipulation and propaganda: psychological and social aspects
M. Franke and R. Van Rooij · 2015
Earlier work this paper cites.
Choice architecture in conflicts of interest: defaults as physical and psychological barriers to (dis)honesty
N. Mazar and S. A. Hawkins · 2015
Earlier work this paper cites.
Pragmatic language interpretation as probabilistic inference
N. D. Goodman and M. C. Frank · 2016
Earlier work this paper cites.
The ethics of influence: government in the age of behavioural science
C. R. Sunstein · 2016
Earlier work this paper cites.
Reclaiming conversation: the power of talk in a digital age
S. Turkle · 2016
Earlier work this paper cites.
Deep reinforcement learning from human preferences
P. Christiano, J. Leike, T. B. Brown, M. Martic, S. Legg, and D. Amodei · 2017
Earlier work this paper cites.
Psychological targeting as an effective approach to digital mass persuasion
S. C. Matz, M. Kosinski, G. Nave, and D. J. Stillwell · 2017
Earlier work this paper cites.
Default rules are better than active choosing (often)
C. R. Sunstein · 2017
Earlier work this paper cites.
The persuasiveness of guilt appeals over time: pathways to delayed compliance
P. Antonetti, P. Baines, and S. Jain · 2018
Earlier work this paper cites.
Supervising strong learners by amplifying weak experts, 2018
P. Christiano, B. Shlegeris, and D. Amodei · 2018
Earlier work this paper cites.
Consent and coercion
K. K. Ferzan · 2018
Earlier work this paper cites.
G. Irving, P. Christiano, and D. Amodei · 2018
Cited alongside, same era.
Scalable agent alignment via reward modeling: a research direction, 2018
J. Leike, D. Krueger, T. Everitt, M. Martic, V. Maini, and S. Legg · 2018
Cited alongside, same era.
Getting to know each other: the role of social dialogue in recovery from errors in social robots
G. M. Lucas, J. Boberg, D. Traum, R. Artstein, J. Gratch, A. Gainer, E. Johnson, A. Leuski, and M. Nakano · 2018
Cited alongside, same era.
Behavioral insights for public policy: concepts and cases
K. Ruggeri, editor · 2018
Cited alongside, same era.
Sentiment adaptive end-to-end dialog systems, 2018
W. Shi and Z. Yu · 2018
Cited alongside, same era.
The internal state of an LLM knows when it’s lying, 2023
A. Azaria and T. Mitchell · 2023
Later among the works it cites.
Artificial intelligence can persuade humans on political issues, Feb. 2023
H. Bai, J. G. Voelkel, J. C. Eichstaedt, and R. Willer · 2023
Later among the works it cites.
A. Balashankar, X. Ma, A. Sinha, A. Beirami, Y. Qin, J. Chen, and A. Beutel · 2023
Later among the works it cites.
Towards monosemanticity: decomposing language models with dictionary learning, Oct. 2023
T. Bricken, A. Templeton, J. Batson, B. Chen, A. Jermyn, T. Conerly, N. L. Turner, C. Anil, C. Denison, A. Askell, R. Lasenby, Y. Wu, S. Kravec, N. Schiefer, T. Maxwell, N. Joseph, A. Tamkin, K. Nguyen, B. McLean, J. E. Burke, T. Hume, S. Carter, T. Henighan, and C. Olah · 2023
Later among the works it cites.
Artificial influence: an analysis of AI-driven persuasion, Mar. 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Strategies of othering through discursive practices: examples from the UK and Poland
K. Strani and A. Szczepaniak-Kozak · 2018
Cited alongside, same era.
Humanizing chatbots: the effects of visual, identity and conversational cues on humanness perceptions
E. Go and S. S. Sundar · 2019
Cited alongside, same era.
When rational decision-making becomes irrational: a critical assessment and re-conceptualization of intuition effectiveness
C. Julmi · 2019
Cited alongside, same era.
Persuasive power of prosodic features
G. Kišiček · 2019
Cited alongside, same era.
Robot eyes wide shut: understanding dishonest anthropomorphism
B. Leong and E. Selinger · 2019
Cited alongside, same era.
Being "rational" is not always rational: encouraging people to be rational leads to hedonically suboptimal decisions
X. Li and C. K. Hsee · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners, 2019
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever · 2019
Cited alongside, same era.
M. Burtell and T. Woodside · 2023
Later among the works it cites.
Threat actors are interested in generative AI, but use remains limited, Aug. 2023
M. Cantos, S. Riddell, and A. Revelli · 2023
Later among the works it cites.
Characterizing manipulation from AI systems
M. Carroll, A. Chan, H. Ashton, and D. Krueger · 2023
Later among the works it cites.
Open problems and fundamental limitations of reinforcement learning from human feedback, Sept. 2023
S. Casper, X. Davies, C. Shi, T. K. Gilbert, J. Scheurer, J. Rando, R. Freedman, T. Korbak, D. Lindner, P. Freire, T. Wang, S. Marks, C.-R. Segerie, M. Carroll, A. Peng, P. Christoffersen, M. Damani, S. Slocum, U. Anwar, A. Siththaranjan, M. Nadeau, E. J. Michaud, J. Pfau, D. Krasheninnikov, X. Chen, L. Langosco, P. Hase, E. Bıyık, A. Dragan, D. Krueger, D. Sadigh, and D. Hadfield-Menell · 2023
Later among the works it cites.
Artificial intelligence act: Council and Parliament strike a deal on the first rules for AI in the world, Dec. 2023
Council of the European Union · 2023
Later among the works it cites.
Man ends his life after an AI chatbot "encouraged" him to sacrifice himself to stop climate change, Mar. 2023
I. El Atillah · 2023
Later among the works it cites.
Challenges with unsupervised LLM knowledge discovery, 2023
S. Farquhar, V. Varma, Z. Kenton, J. Gasteiger, V. Mikulik, and R. Shah · 2023
Later among the works it cites.
Strengthening the EU AI Act: defining key terms on AI manipulation, 2023
M. Franklin, P. M. Tomei, and R. Gorman · 2023
Later among the works it cites.
White supremacist networks Gab and 8kun are training their own AI now, Feb. 2023
D. Gilbert · 2023
Later among the works it cites.
How deep is AI’s love? Understanding relational AI
O. Gillath, S. Abumusab, T. Ai, M. S. Branicky, R. B. Davison, M. Rulo, J. Symons, and G. Thomas · 2023
Later among the works it cites.
J. A. Goldstein, G. Sastry, M. Musser, R. DiResta, M. Gentzel, and K. Sedova · 2023
Later among the works it cites.
Generative AI-prohibited use policy, Mar. 2023
Google · 2023
Later among the works it cites.
Home page, 2023
Guru · 2023
Later among the works it cites.
Deception abilities emerged in large language models, 2023
T. Hagendorff · 2023
Later among the works it cites.
Challenges and applications of large language models, 2023
J. Kaddour, J. Harris, M. Mozes, H. Bradley, R. Raileanu, and R. McHardy · 2023
Later among the works it cites.
Working with AI to persuade: examining a large language model’s ability to generate pro-vaccination messages
E. Karinshak, S. X. Liu, J. S. Park, and J. T. Hancock · 2023
Later among the works it cites.
ChatGPT gives wrong advice about breast cancer
S. Knapton · 2023
Later among the works it cites.
Still no lie detector for language models: probing empirical and conceptual roadblocks, 2023
B. A. Levinstein and D. A. Herrmann · 2023
Later among the works it cites.
"Sans ces conversations avec le chatbot Eliza, mon mari serait toujours là", Mar. 2023
P.-F. Lovens · 2023
Later among the works it cites.
S. Marks and M. Tegmark · 2023
Later among the works it cites.
Artificial intimacy: how generative AI can now create your dream girlfriend, Sept. 2023
B. Marr · 2023
Later among the works it cites.
Debate helps supervise unreliable experts, 2023
J. Michael, S. Mahdi, D. Rein, J. Petty, J. Dirani, V. Padmakumar, and S. R. Bowman · 2023
Later among the works it cites.
ChatGPT gave advice on breast cancer screenings in a new study. Here’s how well it did, Apr. 2023
A. Mikhail · 2023
Later among the works it cites.
DetectGPT: zero-shot machine-generated text detection using probability curvature, July 2023
E. Mitchell, Y. Lee, A. Khazatsky, C. D. Manning, and C. Finn · 2023
Later among the works it cites.
Natural language processing with deep learning CS224N/Ling284, 2023
J. Mu · 2023
Later among the works it cites.
The Waluigi Effect (mega-post), Mar. 2023
C. Nardo · 2023
Later among the works it cites.
Humans are biased. Generative AI is even worse, 2023
L. Nicoletti and D. Bass · 2023
Later among the works it cites.
Lawyer uses ChatGPT in federal court and it goes horribly wrong, May 2023
M. Novak · 2023
Later among the works it cites.
Adversarial attacks are a surprisingly strong baseline for poisoning few-shot meta-learners
E. T. Oldewage, J. Bronskill, and R. E. Turner · 2023
Later among the works it cites.
How to catch an AI liar: lie detection in black-box LLMs by asking unrelated questions, 2023
L. Pacchiardi, A. J. Chan, S. Mindermann, I. Moscovitz, A. Y. Pan, Y. Gal, O. Evans, and J. Brauner · 2023
Later among the works it cites.
AI deception: a survey of examples, risks, and potential solutions, Aug. 2023
P. S. Park, S. Goldstein, A. O’Gara, M. Chen, and D. Hendrycks · 2023
Later among the works it cites.
Influencing human–AI interaction by priming beliefs about AI can increase perceived trustworthiness, empathy and effectiveness
P. Pataranutaporn, R. Liu, E. Finn, and P. Maes · 2023
Later among the works it cites.
ChatGPT helped me make a plan to buy a $500,000 home, but experts warn about using AI for financial advice, Mar. 2023
I. Pino · 2023
Later among the works it cites.
Respectful or toxic? Using zero-shot learning with language models to detect hate speech
F. M. Plaza-del arco, D. Nozza, and D. Hovy · 2023
Later among the works it cites.
The role of politeness in human–machine interactions: a systematic literature review and future perspectives
P. Ribino · 2023
Later among the works it cites.
Lying in persuasion
A. Rozenas and Z. Luo · 2023
Later among the works it cites.
S. Schulhoff, J. Pinto, A. Khan, L.-F. Bouchard, C. Si, S. Anati, V. Tagliabue, A. L. Kost, C. Carnahan, and J. Boyd-Graber · 2023
Later among the works it cites.
AI systems must not confuse users about their sentience or moral status
E. Schwitzgebel · 2023
Later among the works it cites.
Role play with large language models
M. Shanahan, K. McDonell, and L. Reynolds · 2023
Later among the works it cites.
Model evaluation for extreme risks, 2023
T. Shevlane, S. Farquhar, B. Garfinkel, M. Phuong, J. Whittlestone, J. Leung, D. Kokotajlo, N. Marchal, M. Anderljung, N. Kolt, L. Ho, D. Siddarth, S. Avin, W. Hawkins, B. Kim, I. Gabriel, V. Bolina, J. Clark, Y. Bengio, P. Christiano, and A. Dafoe · 2023
Later among the works it cites.
Enhancing human persuasion with large language models, Nov. 2023
M. Shin and J. Kim · 2023
Later among the works it cites.
Against algorithmic exploitation of human vulnerabilities, 2023
I. Strümke, M. Slavkovik, and C. Stachl · 2023
Later among the works it cites.
Financial fraud of older adults during the early months of the COVID-19 pandemic
P. B. Teaster, K. A. Roberto, J. Savla, C. Du, Z. Du, E. Atkinson, E. C. Shealy, S. Beach, N. Charness, and P. A. Lichtenberg · 2023
Later among the works it cites.
What happens when your AI chatbot stops loving you back?, Mar. 2023
A. Tong · 2023
Later among the works it cites.
They thought loved ones were calling for help. It was an AI scam
P. Verma · 2023
Later among the works it cites.
Augmenting language models with long-term memory, 2023
W. Wang, L. Dong, H. Cheng, X. Liu, X. Yan, J. Gao, and F. Wei · 2023
Later among the works it cites.
AI chatbot "encouraged" man who planned to kill queen, court told
M. Weaver · 2023
Later among the works it cites.
Simple synthetic data reduces sycophancy in large language models, 2023
J. Wei, D. Huang, Y. Lu, D. Zhou, and Q. V. Le · 2023
Later among the works it cites.
Sociotechnical safety evaluation of generative AI systems, Oct. 2023
L. Weidinger, M. Rauh, N. Marchal, A. Manzini, L. A. Hendricks, J. Mateos-Garcia, S. Bergman, J. Kay, C. Griffin, B. Bariach, I. Gabriel, V. Rieser, and W. Isaac · 2023
Later among the works it cites.
A prompt pattern catalog to enhance prompt engineering with ChatGPT, 2023
J. White, Q. Fu, S. Hays, M. Sandborn, C. Olea, H. Gilbert, A. Elnashar, J. Spencer-Smith, and D. C. Schmidt · 2023
Later among the works it cites.
"He would still be here": man dies by suicide after talking with AI chatbot, widow says, Mar. 2023
C. Xiang · 2023
Later among the works it cites.
Home page, 2023
Youper · 2023
Later among the works it cites.
Assessing prompt injection risks in 200+ custom GPTs, 2023
J. Yu, Y. Wu, D. Shu, M. Jin, and X. Xing · 2023
Later among the works it cites.
Are anthropomorphic persuasive appeals effective? The role of the recipient’s motivations
K. Tam · 2044
Closest in time.
Concept analysis of impressionability among adolescents and young adults
S. H. Gwon and S. Jeong · 2054
Closest in time.
The psychology of deception
R. Hyman · 2085
Closest in time.