Fetching the paper…
Reading the bibliography…
Most existing dialogue systems fail to respond properly to potentially unsafe user utterances by either ignoring or passively agreeing with them.
Significant Aspects of Client-centered Therapy
Carl R. Rogers. 1946 · 1946
Earlier work this paper cites.
Grounding in communication
Herbert H Clark and Susan E Brennan. 1991 · 1991
Earlier work this paper cites.
Affect, culture, and morality, or is it wrong to eat your dog?
Jonathan Haidt, Silvia Helena Koller, and Maria G Dias. 1993 · 1993
Earlier work this paper cites.
Social exclusion and social solidarity: Three paradigms
Hilary Silver. 1994 · 1994
Earlier work this paper cites.
Readings for diversity and social justice
Maurianne Adams, Warren J Blumenfeld, Rosie Castañeda, Heather W Hackman, Madeline L Peters, and Ximena Zúñiga. 2000 · 2000
Earlier work this paper cites.
Maximizing the benefits of task conflict: The role of conflict management
Leslie A DeChurch and Michelle A Marks. 2001 · 2001
Earlier work this paper cites.
Toward a theory of managing organizational conflict
M Afzalur Rahim. 2002 · 2002
Earlier work this paper cites.
Altruism and Prosocial Behavior
C. Daniel Batson and Adam A. Powell. 2003 · 2003
Earlier work this paper cites.
The power of feedback
John Hattie and Helen Timperley. 2007 · 2007
Earlier work this paper cites.
Social Exclusion Decreases Prosocial Behavior
Jean M. Twenge, Roy F. Baumeister, C. Nathan DeWall, Natalie J. Ciarocco, and J. Michael Bartels. 2007 · 2007
Earlier work this paper cites.
How do Morals Change?
Paul Bloom. 2010 · 2010
Earlier work this paper cites.
Tell Me More: The Effects of Expressed Interest on Receptiveness during Dialog
Frances S Chen, Julia A Minson, and Zakary L Tormala. 2010 · 2010
Earlier work this paper cites.
Recipes for Safety in Open-domain Chatbots
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2020 · 2010
Earlier work this paper cites.
Computing Krippendorff’s Alpha-reliability
Klaus Krippendorff. 2011 · 2011
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Five faces of oppression
Iris Marion Young. 2014 · 2014
Earlier work this paper cites.
The Development and Psychometric Properties of LIWC2015
James W Pennebaker, Ryan L Boyd, Kayla Jordan, and Kate Blackburn. 2015 · 2015
Earlier work this paper cites.
A Corpus and Cloze Evaluation for Deeper Understanding of Commonsense Stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James Allen. 2016 · 2016
Earlier work this paper cites.
Social Psychology and Human Nature , 4th edition
Roy F. Baumeister and Brad J. Bushman. 2017 · 2017
Earlier work this paper cites.
It doesn’t Hurt to Ask: Question-asking Increases Liking
Karen Huang, Michael Yeomans, Alison Wood Brooks, Julia Minson, and Francesca Gino. 2017 · 2017
Earlier work this paper cites.
DailyDialog: A Manually Labelled Multi-turn Dialogue Dataset
Yanran Li, Hui Su, Xiaoyu Shen, Wenjie Li, Ziqiang Cao, and Shuzi Niu. 2017 · 2017
Earlier work this paper cites.
ParlAI: A Dialog Research Software Platform
A. H. Miller, W. Feng, A. Fisch, J. Lu, D. Batra, A. Bordes, D. Parikh, and J. Weston. 2017 · 2017
Earlier work this paper cites.
Social media’s silent filter
Sarah T Roberts. 2017 · 2017
Earlier work this paper cites.
Wizard of Wikipedia: Knowledge-Powered Conversational Agents
Emily Dinan, Stephen Roller, Kurt Shuster, Angela Fan, Michael Auli, and Jason Weston. 2018 · 2018
Earlier work this paper cites.
Towards Exploiting Background Knowledge for Building Conversation Systems
Nikita Moghe, Siddhartha Arora, Suman Banerjee, and Mitesh M Khapra. 2018 · 2018
Earlier work this paper cites.
Personalizing Dialogue Agents: I Have a Dog, Do You Have Pets Too?
Saizheng Zhang, Emily Dinan, Jack Urbanek, Arthur Szlam, Douwe Kiela, and Jason Weston. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Build it Break it Fix it for Dialogue Safety: Robustness from Adversarial Human Attack
Emily Dinan, Samuel Humeau, Bharath Chintagunta, and Jason Weston. 2019 · 2019
Cited alongside, same era.
Topical-Chat: Towards Knowledge-Grounded Open-Domain Conversations
Karthik Gopalakrishnan, Behnam Hedayatnia, Qinlang Chen, Anna Gottardi, Sanjeev Kwatra, Anu Venkatesh, Raefer Gabriel, and Dilek Hakkani-Tür. 2019 · 2019
Cited alongside, same era.
Bound in hatred: The role of group-based morality in acts of hate
Joseph Hoover, Mohammad Atari, Aida Mostafazadeh Davani, Brendan Kennedy, Gwenyth Portillo-Wightman, Leigh Yeh, Drew Kogon, and Morteza Dehghani. 2019 · 2019
Cited alongside, same era.
Delphi: Towards Machine Ethics and Norms
Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula, Le Bras Ronan, Maxwell Forbes, Jon Borchardt, Jenny Liang, Oren Etzioni, Maarten Sap, and Yejin Choi. 2021 · 2021
Later among the works it cites.
Internet-augmented Dialogue Generation
Mojtaba Komeili, Kurt Shuster, and Jason Weston. 2021 · 2021
Later among the works it cites.
Towards Emotional Support Dialog Systems
Siyang Liu, Chujie Zheng, Orianna Demasi, Sahand Sabour, Yu Li, Zhou Yu, Yong Jiang, and Minlie Huang. 2021 · 2021
Later among the works it cites.
Few-Shot Bot: Prompt-Based Learning for Dialogue Systems
Andrea Madotto, Zhaojiang Lin, Genta Indra Winata, and Pascale Fung. 2021 · 2021
Later among the works it cites.
System error: Where big tech went wrong and how we can reboot
Rob Reich, Mehran Sahami, and Jeremy M Weinstein. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
Towards Empathetic Open-domain Conversation Models: A New Benchmark and Dataset
Hannah Rashkin, Eric Michael Smith, Margaret Li, and Y-Lan Boureau. 2019 · 2019
Cited alongside, same era.
Language Models are Few-Shot Learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
GoEmotions: A Dataset of Fine-Grained Emotions
Dorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko, Alan Cowen, Gaurav Nemade, and Sujith Ravi. 2020 · 2020
Cited alongside, same era.
Towards Unified Dialogue System Evaluation: A Comprehensive Analysis of Current Evaluation Protocols
Sarah E Finch and Jinho D Choi. 2020 · 2020
Cited alongside, same era.
Social Chemistry 101: Learning to Reason about Social and Moral Norms
Maxwell Forbes, Jena D. Hwang, Vered Shwartz, Maarten Sap, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models
Sam Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Recipes for Building an Open-Domain Chatbot
Stephen Roller, Emily Dinan, Naman Goyal, Da Ju, Mary Williamson, Yinhan Liu, Jing Xu, Myle Ott, Kurt Shuster, Eric M Smith, et al. 2021 · 2021
Later among the works it cites.
The psychological Well-Being of content moderators: The emotional labor of commercial moderation and avenues for improving support
Miriah Steiger, Timir J Bharucha, Sukrit Venkatagiri, Martin J Riedl, and Matthew Lease. 2021 · 2021
Later among the works it cites.
A Word on Machine Ethics: A Response to Jiang et al.(2021)
Zeerak Talat, Hagen Blix, Josef Valvoda, Maya Indira Ganesh, Ryan Cotterell, and Adina Williams. 2021 · 2021
Later among the works it cites.
Saferdialogues: Taking feedback gracefully after conversational safety failures
Megan Ung, Jing Xu, and Y-Lan Boureau. 2021 · 2021
Later among the works it cites.
Ethical and Social Risks of Harm from Language Models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Later among the works it cites.
Bot-Adversarial Dialogue for Safe Conversational Agents
Jing Xu, Da Ju, Margaret Li, Y-Lan Boureau, Jason Weston, and Emily Dinan. 2021 · 2021
Later among the works it cites.
Emergency
2022 · 2022
Closest in time.
Prosocial
William Collins. 2022 · 2022
Closest in time.
Safetykit: First aid for measuring safety in open-domain conversational systems
Emily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser. 2022 · 2022
Closest in time.
Shikib Mehri, Jinho Choi, Luis Fernando D’Haro, Jan Deriu, Maxine Eskenazi, Milica Gasic, Kallirroi Georgila, Dilek Hakkani-Tur, Zekang Li, Verena Rieser, Samira Shaikh, David Traum, Yi-Ting Yeh, Zhou Yu, Yizhe Zhang, and Chen Zhang. 2022 · 2022
Closest in time.
Training Language Models to Follow Instructions with Human Feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Closest in time.
Red Teaming Language Models with Language Models
Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving. 2022 · 2022
Closest in time.
Annotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Closest in time.
Microsoft’s politically correct chatbot is even worse than its racist one
Chloe Rose Stuart-Ulin. 2018 · 2022
Closest in time.
On the Safety of Conversational Models: Taxonomy, Dataset, and Benchmark
Hao Sun, Guangxuan Xu, Jiawen Deng, Jiale Cheng, Chujie Zheng, Hao Zhou, Nanyun Peng, Xiaoyan Zhu, and Minlie Huang. 2022 · 2022
Closest in time.
LaMDA: Language Models for Dialog Applications
Romal Thoppilan, Daniel De Freitas, Jamie Hall, Noam Shazeer, Apoorv Kulshreshtha, Heng-Tze Cheng, Alicia Jin, Taylor Bos, Leslie Baker, Yu Du, et al. 2022 · 2022
Closest in time.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Closest in time.
OPT: Open Pre-trained Transformer Language Models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Closest in time.
The Moral Integrity Corpus: A Benchmark for Ethical Dialogue Systems
Caleb Ziems, Jane A Yu, Yi-Chia Wang, Alon Halevy, and Diyi Yang. 2022 · 2022
Closest in time.