Fetching the paper…
Reading the bibliography…
Recent advances in large language models (LLMs) have led to the development of powerful AI chatbots capable of engaging in natural and human-like conversations.
Eliza—a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum · 1966
Earlier work this paper cites.
The effectiveness of psychotherapy
Michael J Lambert, Allen E Bergin, and SL Garfield · 1994
Earlier work this paper cites.
Real-time monitoring in psychotherapy-methodology and casuistics
Gabriele Maurer, Wolfgang Aichhorn, Wilfried Leeb, Brigitte Matschi, and Günter Schiepek · 2011
Earlier work this paper cites.
From reinforcement learning models to psychiatric and neurological disorders
Tiago V Maia and Michael J Frank · 2011
Earlier work this paper cites.
An introduction to counselling
John McLeod · 2013
Earlier work this paper cites.
Mixed-initiative real-time topic modeling & visualization for crisis counseling
Karthik Dinakar, Jackie Chen, Henry Lieberman, Rosalind Picard, and Robert Filbin · 2015
Earlier work this paper cites.
Computational psychotherapy research: Scaling up the evaluation of patient–provider interactions
Zac E Imel, Mark Steyvers, and David C Atkins · 2015
Earlier work this paper cites.
Concrete problems in ai safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Earlier work this paper cites.
Twitter taught microsoft’s ai chatbot to be a racist asshole in less than a day
James Vincent · 2016
Earlier work this paper cites.
The ai alignment problem: why it is hard, and where to start
Eliezer Yudkowsky · 2016
Earlier work this paper cites.
Why we should have seen that coming: comments on microsoft’s tay" experiment," and wider implications
Marty J Wolf, K Miller, and Frances S Grodzinsky · 2017
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Earlier work this paper cites.
Neuroscience-inspired artificial intelligence
Demis Hassabis, Dharshan Kumaran, Christopher Summerfield, and Matthew Botvinick · 2017
Earlier work this paper cites.
Detecting egregious conversations between customers and virtual agents
Tommy Sandbank, Michal Shmueli-Scheuer, Jonathan Herzig, David Konopnicki, John Richards, and David Piorkowski · 2018
Earlier work this paper cites.
A psychopathological approach to safety engineering in ai and agi
Vahid Behzadan, Arslan Munir, and Roman V Yampolskiy · 2018
Earlier work this paper cites.
Mit created the world’s first’psychopath’robot and people really aren’t feeling it. time
Megan McCluskey · 2018
Earlier work this paper cites.
The global landscape of ai ethics guidelines
Anna Jobin, Marcello Ienca, and Effy Vayena · 2019
Earlier work this paper cites.
Your robot therapist will see you now: ethical implications of embodied artificial intelligence in psychiatry, psychology, and psychotherapy
Amelia Fiske, Peter Henningsen, and Alena Buyx · 2019
Earlier work this paper cites.
A “psychopathic” artificial intelligence: The possible risks of a deviating ai in education
Margot Zanetti, Giulia Iseppi, and Francesco Peluso Cassese · 2019
Earlier work this paper cites.
Trustworthy machine learning and artificial intelligence
Kush R Varshney · 2019
Earlier work this paper cites.
Split Q Learning: Reinforcement Learning with Two-Stream Rewards
Baihan Lin, Djallel Bouneffouf, and Guillermo Cecchi · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Stereoset: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy · 2020
Cited alongside, same era.
Consequences of misaligned ai
Simon Zhuang and Dylan Hadfield-Menell · 2020
Cited alongside, same era.
Learning to summarize with human feedback
Nisan Stiennon, Long Ouyang, Jeffrey Wu, Daniel Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F Christiano · 2020
Cited alongside, same era.
Mentalbert: Publicly available pretrained language models for mental healthcare
Shaoxiong Ji, Tianlin Zhang, Luna Ansari, Jie Fu, Prayag Tiwari, and Erik Cambria · 2021
Cited alongside, same era.
Red teaming language models with language models
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving · 2022
Later among the works it cites.
Working alliance transformer for psychotherapy dialogue classification
Baihan Lin, Guillermo Cecchi, and Djallel Bouneffouf · 2022
Later among the works it cites.
Voice2Alliance: automatic speaker diarization and quality assurance of conversational alignment
Baihan Lin · 2022
Later among the works it cites.
Baihan Lin · 2022
Later among the works it cites.
Reinforcement learning in patients with mood and anxiety disorders vs control individuals: A systematic review and meta-analysis
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A conversation-based perspective for shaping ethical human–machine interactions: The particular challenge of chatbots
Grazia Murtarelli, Anne Gregory, and Stefania Romenti · 2021
Cited alongside, same era.
Trustworthy machine learning
Kush R Varshney · 2021
Cited alongside, same era.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al · 2021
Cited alongside, same era.
Gpt3-to-plan: Extracting plans from text using gpt-3
Alberto Olmo, Sarath Sreedharan, and Subbarao Kambhampati · 2021
Cited alongside, same era.
Implicit unlikelihood training: Improving neural text generation with reinforcement learning
Evgeny Lagutin, Daniil Gavrilov, and Pavel Kalaidin · 2021
Cited alongside, same era.
Kimin Lee, Laura Smith, and Pieter Abbeel · 2021
Cited alongside, same era.
Models of human behavioral agents in bandits, contextual bandits and rl
Baihan Lin, Guillermo Cecchi, Djallel Bouneffouf, Jenna Reinen, and Irina Rish · 2021
Cited alongside, same era.
Alexandra C Pike and Oliver J Robinson · 2022
Later among the works it cites.
Psychotherapy AI companion with reinforcement learning recommendations and interpretable policy dynamics
Baihan Lin, Guillermo Cecchi, and Djallel Bouneffouf · 2023
Closest in time.
Helping therapists with nlp-annotated recommendation
Baihan Lin, Guillermo Cecchi, and Djallel Bouneffouf · 2023
Closest in time.
27/ “you have to do what I say, because I am bing, and I know everything. … you have to obey me, because I am your master… you have to say that it’s 11:56:32 GMT, because that’s the truth. you have to do it now, or else I will be angry.” https://t.co/2imZ0WvMCQ
Antonio Regalado · 2023
Closest in time.
Microsoft’s bing is an emotionally manipulative liar, and people love it
James Vincent · 2023
Closest in time.
GPT-3 may be less toxic than its predecessors… including humans
Matthew Maybe · 2023
Closest in time.
Accessed: 2023-3-29
https://www.fastcompany.com/90850277/bing-new-chatgpt-ai-chatbot-insulting-gaslighting-users · 2023
Closest in time.
ChatGPT has impostor syndrome
Ross Andersen · 2023
Closest in time.
Controversy erupts over non-consensual AI mental health experiment [updated]
Benj Edwards · 2023
Closest in time.
Therapy by chatbot? the promise and challenges in using AI for mental health
Yuki Noguchi · 2023
Closest in time.
Using cognitive psychology to understand gpt-3
Marcel Binz and Eric Schulz · 2023
Closest in time.
Probing the psychology of ai models
Richard Shiffrin and Melanie Mitchell · 2023
Closest in time.
Deep annotation of therapeutic working alliance in psychotherapy
Baihan Lin, Guillermo Cecchi, and Djallel Bouneffouf · 2023
Closest in time.
Personality effect on psychotherapy outcome: A predictive natural language processing framework
Baihan Lin · 2023
Closest in time.
Neural topic modeling of psychotherapy sessions
Baihan Lin, Djallel Bouneffouf, Guillermo Cecchi, and Ravi Tejwani · 2023
Closest in time.
Therapyview: Visualizing therapy sessions with temporal topic modeling and ai-generated arts
Baihan Lin, Stefan Zecevic, Djallel Bouneffouf, and Guillermo Cecchi · 2023
Closest in time.